Viewing profile — gck1
gck1
HN member- Joined
- Wed, May 09, 2018, 3:20 PM UTC
- HN karma
- 895
- Public activity
- 372 items
- HN profile
- View on Hacker News ↗
About gck1
Recent public activity
-
comment
Comment #49125966
Heck, agents don't start editing before they're already at 70k for me. I've played with explorer agents giving exploration summaries to help the implementer agents use more of thei…
- story
- story
-
comment
Comment #49118102
This will absolutely not end well.
-
comment
Comment #49118063
> According to whom? It's very easy to answer this without my help by trying to get access to Mythos. Do you see requirements clearly listed anywhere? Can you even apply? What you'…
-
comment
Comment #49117681
You seem to be putting a lot of weight on Anthropic employees being the smartest people in the world. And I don't doubt that, not in the slightest. But I've seen exceptionally smar…
-
comment
Comment #49117569
They had a model escape in April, roughly the same time when they were fearmongering about Mythos and how Anthropic should be the sole keyholder of cybersecurity capabilities, and …
-
comment
Comment #49117451
They also gave access to Mythos ( the Mythos) to some companies, based on... vibes. Who knows how these companies are using it. If Anthropic can't effectively contain their own mod…
-
comment
Comment #49117088
> On July 21, OpenAI disclosed that several of their models had broken out of an isolated test environment > In response to this incident, we began a large-scale retrospective revi…
-
comment
Comment #49116793
Recent-ish models learned to use the same trick engineers played on non-engineers, where they try to sound very smart by overcomplicating very simple concepts. It's very taxing, es…
-
comment
Comment #49116281
They do have subagents, released v2 of that feature with the launch of 5.6 model series in fact. It's just... very poorly executed, is a significant regression from subagents v1 an…
-
comment
Comment #49115936
It's funny how codex itself can't do Sol orchestrator / Luna implementor out of the box.
-
comment
Comment #49115579
Luna is comparable to GPT 5.4 from 4 months ago on many benchmarks. I know many who have said during that time, myself included, that if that's the model they had to use for the re…
-
comment
Comment #49113842
They're supposed to bring 5h today.
-
comment
Comment #49090585
I did a full circle and essentially dropped all of my personal static workflows encoded in skills because I observed recent models picking better ad-hoc workflows for particular pr…
-
comment
Comment #49077469
It took me a few hours to find some very questionable communities, which in turn gave me access to: - Ways to obtain cheap guarded-AI tokens that are not linked back to me and with…
-
comment
Comment #49077009
> In cybersecurity, a level playing field favors the attacker Yes, but didn't it always? Hence why my position is that this will get us back to relatively where we were pre-LLMs. A…
-
comment
Comment #49076808
I've got zero knowledge of bio, so can't answer that. But with cyber the answer is very simple - the attackers already have more cyber-offense capabilities and there's no putting i…
-
comment
Comment #49076682
It's refreshing to see how there's almost no person in this thread who can't see the BS. All the goodwill that Anthropic could have had is basically gone. Anthropic is likely on th…
-
comment
Comment #49076569
> We should not sell powerful chips or chipmaking equipment to China Yes, please. We don't know whether we'd have open weight models today, had the chip-prohibition not been in pla…
-
comment
Comment #49073777
Its such a shame Android's backup/restore is such a mess, even more so in GrapheneOS. I remember the era before Google and manufacturs started cracking down on bootloaders and cust…
-
comment
Comment #49073331
They can still make your life very difficult. They could throw you in jail for "obstructing investigation" or something similar. Not in US, but had my phone seized by authorities a…
-
comment
Comment #49048180
Claude's "I'm going to draw the line here", "This is where I'm going to hold the ground" always rubs me the wrong way. Classifiers rejecting a request are one thing, but there's so…
-
comment
Comment #49044987
That's brilliant, I should try that. I usually just start by preloadig context with plausible legitimate use, have it work and obviously fail, and then ask to figure it out without…
-
comment
Comment #49043473
> so you haven’t actually used Claude Code yet. Where do you think the principle came from? I've used claude code for a year, and stopped February this year.