Live data from Hacker News

Viewing profile — olliepro

olliepro

HN member
Joined
Fri, Apr 04, 2025, 4:20 PM UTC
HN karma
62
Public activity
36 items

About olliepro

Data-Scientist > CS/AI PhD

Recent public activity

  1. comment
    Comment #48762787

    The authors have some inconsistencies with training token length… Most errors are probably responses that didn’t finish before their 3K token limit. They’ve measured how well RL is…

  2. comment
    Comment #48609447

    This is the classic pattern of LLM generated MCQs.

  3. comment
    Comment #48214911

    With super high res onboard camera footage too.

  4. comment
    Comment #47726991

    They do quite a lot of distillation. As we've seen from the American open weight models from AI2 (OLMo series of models). They have a lot of incentive to distill beyond just copyin…

  5. comment
    Comment #47726955

    A lot of distillation happens. E.g. OLMo models have a completely open dataset and they are heavily distilled. It only makes sense to try to absorb behaviors from the best models o…

  6. comment
    Comment #47692556

    decentralized training makes a lot more sense when the required hardware isn't a $40K GPU...

  7. comment
    Comment #47689587

    This would likely only get used for small finetuning jobs. It’s too slow for the scale of pretraining.

  8. comment
    Comment #47268648

    I bet they lack good long context training data and need to start a flywheel of collecting it via their api (from willing customers)

  9. comment
    Comment #46979483

    Tensors are in no shortage nowadays. I did read this a tensors though and got a good laugh.

  10. comment
    Comment #46948854

    There’s a section of I-15 in Utah’s Salt Lake County which reliably has a crash on weekdays at 6pm. It was unfortunately at a pinch point in the mountains with no good alternate ro…

  11. comment
    Comment #46878534

    Much of the scientific medical literature is behind paywalls. They have tapped into that datasource (whereas ChatGPT doesn't have access to that data). I suspect that were the medi…

  12. comment
    Comment #46787051

    It depends on your thing. If the marathon was just the motivation, your thing is running... if the marathon was the bucketlist item, it is the thing.

  13. comment
    Comment #46786897

    Getting everyone to fall in love with the thing is not doing the thing... learned this as a data scientist brought in to work on a project which ended soon thereafter. A team of 20…

  14. comment
    Comment #46786822

    Everyone's threshold is different. I aspire to "move fast and break things", but more often than not, I obsess over the rough edges.

  15. comment
    Comment #46786752

    The more I use AI to do the thing, the more it feels like I didn't do the thing.

  16. comment
    Comment #46774424

    What abstraction levels do you expect will remain only in the Human domain? The progression from basic arithmetic, to complex ratios and basic algebra, graphing, geometry, trig, ca…

  17. comment
    Comment #46739134

    I made a skill that reflects on past conversations via parallel headless codex sessions. Its great for context building. Repo: https://github.com/olliepro/Codex-Reflect-Skill

  18. comment
    Comment #46738606

    I was thinking about something like this, but I don't have codex running on a server. Keep me posted on how it goes!

  19. story
    Show HN: Codex Self-Reflect Skill and CLI to run subagents on past Codex convos

    This skill is useful for identifying agent friction points and brainstorming new skills, developing context of past work for a new conversation, identifying code bloat from failed …

  20. comment
    Comment #46597384

    I believe the idea is that it “files away” the files into folders.

  21. comment
  22. comment
    Comment #46597338

    Can Claude code jump through the hoops for you?

  23. comment
    Comment #46465990

    Three things that shook me awake to the idea that the information barrage of the internet is a tranquilizer/red herring: - Bad Mental Health: At the start of the war in Ukraine I r…

  24. comment
    Comment #46447461

    Although there are many examples of troubling sycophantic responses confirming or encouraging delusions, this document is the original complaint (the initial filing) in a lawsuit a…

  25. comment
    Comment #46247226

    It feels like this should work, but the breadth of knowledge in these models is so vast. Everyone knows how to taste, but not everyone knows physics, biology, math, every language……