Viewing profile — _345
_345
HN member- Joined
- Mon, Nov 07, 2022, 3:13 AM UTC
- HN karma
- 129
- Public activity
- 86 items
- HN profile
- View on Hacker News ↗
About _345
No profile information was provided.
Recent public activity
-
comment
Comment #49128424
I'm pretty sure the lemon lime game referred to in the article is supposed to be some trick where it always says you're wrong and you have to cheat (maybe in browser developer tool…
-
comment
Comment #49079203
"87.3% Share of the maximum achievable score our GRPO-trained 9B open-source model reached on catalog review, vs 76.9% for the best frontier configuration: a 13.5% relative improve…
-
comment
Comment #48974442
i got half way through the readme for this project and had the same thought and got sad and just left the page
-
comment
Comment #48934358
The oneplus open (2023) is such a great phone, what a shame
-
comment
Comment #48927977
What inspired you to make this?
-
comment
Comment #48876465
i dont understand why you would use this. i think it needs more examples
-
comment
Comment #48810816
I feel like its only useful if the work you are doing doesn't have correctness as a high priority. If your work is okay with something only mostly being correct and it can just be …
-
comment
Comment #48756823
This makes so much sense as to why I've always felt that Opus 4.8 was leagues ahead of GPT 5.5. It's so good at taking underspecified requirements and filling in the gaps with sens…
-
comment
Comment #48693131
I did the same thing but it's 25 feet hahaha, love to see this
-
comment
Comment #48513935
Just you as an adult. It's always been that infantile
-
comment
Comment #48513547
any guesses how?
-
comment
Comment #48503794
I'm not letting Jenna from HR log into my personal machine with access to all of my lifelong data though. I do let my claude bypass permissions though
-
comment
Comment #48499897
Best comment in this thread
-
comment
Comment #48499825
In my case no, I actually saw worse performance with fable medium and switched back to opus high and xhigh
-
comment
Comment #48499792
way worse things can happen than your machine being bricked, if a malicious actor can weaponize an agent to do their bidding
-
comment
Comment #48496126
I moved to Firefox as soon as they began threatening uBlock Origin support and people started switching to Lite, I find it silly that people were tweaking their registries just to …
-
comment
Comment #48486066
... how. how is that even possible. pirated usage plans?
-
comment
Comment #48337603
Agree wholeheartedly. I think that Anthropic has just invested more effort in creating a better DevEx than OpenAI, and so people just "feel" that claude code is better but they're …
-
comment
Comment #48204671
It doesn't sound like he was booed, more like the topic of AI was booed when mentioned
-
comment
Comment #48183480
yes just ask claude to add quick-silence-dissension to your project
-
comment
Comment #48013146
It's a seriously degraded experience from a developer's perspective. Okay you've got one local LLM installed finally after configuring everything perfectly, what happens when you w…
-
comment
Comment #48008902
I've been experimenting with Hermes, I'm convinced hermes is also just bad. Like as a harness it has got to be doing something to lobotomize these models- Even GPT-5.4 performs bad…
-
comment
Comment #48002519
If you're okay with sonnet level performance, this sounds like a straight upgrade. But I find that sonnet messes up too much, that it ends up not being worth cost optimizing down t…
-
comment
Comment #48000407
> Also why light text on black background? im really curious how you think this is worse than the other way around
-
comment
Comment #48000351
I think I like this article and I haven't finished it yet, but I don't think the bottleneck has shifted to non-human with the advent of agentic AI. It's still the human (deciding w…