Viewing profile — oshrimpton
oshrimpton
HN member- Joined
- Sun, Mar 29, 2026, 11:49 PM UTC
- HN karma
- 294
- Public activity
- 7 items
- HN profile
- View on Hacker News ↗
About oshrimpton
No profile information was provided.
Recent public activity
-
comment
Comment #48607959
Yeah they are 100% in the wrong for removing the fine tuned codex models. It makes sense why they wouldn't want to allocate so many resources towards fine tuning but still the ensh…
-
comment
Comment #48607645
I would be so curious to find a comprehensive benchmark on this, humans do have an unfortunate ahem Dunning-Kruger effect ahem tendency to do this
-
comment
Comment #48607571
Yeah the benchmark for sure isn't perfect and without super rigid prompting it is far too easy for it to get off course. 28% hallucination rate isn't nothing either
-
comment
Comment #48607405
Surprisingly not! It is the biggest hallucinator on the AA Omniscience Index just 2pp away from V4 Pro. I think this is partially due to the fact that Flash was trained on >32T tok…
-
comment
Comment #48607293
I'd definitely agree that it isn't directly model size, but there is the fact that a larger model in terms of parameter count needs a large amount of training data to not overfit o…
-
comment
Comment #48606532
Agreed on the title, my bad! But yeah, I've had some truly terrible experiences using these "frontier" models in coding agents especially, where they just fabricate facts about cod…
- story