Live data from Hacker News

Viewing profile — oshrimpton

oshrimpton

HN member
Joined
Sun, Mar 29, 2026, 11:49 PM UTC
HN karma
294
Public activity
7 items

About oshrimpton

No profile information was provided.

Recent public activity

  1. comment
    Comment #48607959

    Yeah they are 100% in the wrong for removing the fine tuned codex models. It makes sense why they wouldn't want to allocate so many resources towards fine tuning but still the ensh…

  2. comment
    Comment #48607645

    I would be so curious to find a comprehensive benchmark on this, humans do have an unfortunate ahem Dunning-Kruger effect ahem tendency to do this

  3. comment
    Comment #48607571

    Yeah the benchmark for sure isn't perfect and without super rigid prompting it is far too easy for it to get off course. 28% hallucination rate isn't nothing either

  4. comment
    Comment #48607405

    Surprisingly not! It is the biggest hallucinator on the AA Omniscience Index just 2pp away from V4 Pro. I think this is partially due to the fact that Flash was trained on >32T tok…

  5. comment
    Comment #48607293

    I'd definitely agree that it isn't directly model size, but there is the fact that a larger model in terms of parameter count needs a large amount of training data to not overfit o…

  6. comment
    Comment #48606532

    Agreed on the title, my bad! But yeah, I've had some truly terrible experiences using these "frontier" models in coding agents especially, where they just fabricate facts about cod…

  7. story