Live data from Hacker News

Viewing profile — hagen8

hagen8

HN member
Joined
Wed, Feb 25, 2026, 7:05 PM UTC
HN karma
26
Public activity
19 items

About hagen8

No profile information was provided.

Recent public activity

  1. comment
    Comment #49215108

    Cached input tokens are what drives most costs.

  2. comment
    Comment #49184114

    Wrong. They are commonly used by millions.

  3. comment
    Comment #49184096

    Check out academic papers about: 1. Hierarchical skills, workflow, skill learning 2. Meta Harness, self-learning harnesses 3. Trace/trajectory representation 4. Common agentic benc…

  4. comment
    Comment #49171940

    Check out https://agents-last-exam.org/ there is still room for improvements!

  5. comment
    Comment #49166624

    There are certain physical limits. Calculations need to be done. Either less calculations are necessary for the intelligence, or u accept less intelligence. But there is a limit in…

  6. comment
    Comment #49127637

    Most importantly, after entering the elevator. First press the close button and then the floor. That way u, safe the time of pressing a button as the door is already closing.

  7. comment
    Comment #49069999

    Where are the sources for that?

  8. comment
    Comment #49036007

    This is the claude code frontend-skill.

  9. comment
    Comment #48994232

    Just switch the model, its not that much effort tbh. And u can also get a cheaper model than 2.5 lite for the same intelligence

  10. comment
    Comment #48991225

    This will soon happen with theoretical physics, computer science, and everything which can be verified cheaply. Then, we will have long running projects augmented by agents for 2-4…

  11. comment
    Comment #48926198

    Inference costs will go down massively once they use the upcoming GPUs. I estimated that a model like GLM5.2 will be around 0.03USD/M output tokens in 2 years when the Feynman GPUs…

  12. comment
    Comment #48924791

    Some ppl don't like to hear it. But I would assume that token costs when using an inference provider are cheaper than electricity of using locally. If we just take into account out…

  13. comment
    Comment #48856870

    In my opinion Opus is waaayy better in agentic orchestration. It feels like it can natively deal with multiple subagents whereas gpt needs to be taught extensively.

  14. comment
    Comment #48716754

    This is way to complex... Why don't just use some harness which manages all that and give u a good UI?

  15. comment
    Comment #47372204

    Well, the question is what is contributing to the usage. Because as the context grows, the amount of input tokens are increasing. A model call with 800K token as input is 8 times m…

  16. comment
    Comment #47372134

    Did u use the API or subscription?

  17. comment
    Comment #47321197

    They will sooner or later change that policy or get very slow in keeping up.

  18. comment
    Comment #47271707

    But does it use the same agent harness? Because the harness determines the behavior a lot.

  19. comment