Viewing profile — hagen8
hagen8
HN member- Joined
- Wed, Feb 25, 2026, 7:05 PM UTC
- HN karma
- 26
- Public activity
- 19 items
- HN profile
- View on Hacker News ↗
About hagen8
No profile information was provided.
Recent public activity
-
comment
Comment #49215108
Cached input tokens are what drives most costs.
-
comment
Comment #49184114
Wrong. They are commonly used by millions.
-
comment
Comment #49184096
Check out academic papers about: 1. Hierarchical skills, workflow, skill learning 2. Meta Harness, self-learning harnesses 3. Trace/trajectory representation 4. Common agentic benc…
-
comment
Comment #49171940
Check out https://agents-last-exam.org/ there is still room for improvements!
-
comment
Comment #49166624
There are certain physical limits. Calculations need to be done. Either less calculations are necessary for the intelligence, or u accept less intelligence. But there is a limit in…
-
comment
Comment #49127637
Most importantly, after entering the elevator. First press the close button and then the floor. That way u, safe the time of pressing a button as the door is already closing.
-
comment
Comment #49069999
Where are the sources for that?
-
comment
Comment #49036007
This is the claude code frontend-skill.
-
comment
Comment #48994232
Just switch the model, its not that much effort tbh. And u can also get a cheaper model than 2.5 lite for the same intelligence
-
comment
Comment #48991225
This will soon happen with theoretical physics, computer science, and everything which can be verified cheaply. Then, we will have long running projects augmented by agents for 2-4…
-
comment
Comment #48926198
Inference costs will go down massively once they use the upcoming GPUs. I estimated that a model like GLM5.2 will be around 0.03USD/M output tokens in 2 years when the Feynman GPUs…
-
comment
Comment #48924791
Some ppl don't like to hear it. But I would assume that token costs when using an inference provider are cheaper than electricity of using locally. If we just take into account out…
-
comment
Comment #48856870
In my opinion Opus is waaayy better in agentic orchestration. It feels like it can natively deal with multiple subagents whereas gpt needs to be taught extensively.
-
comment
Comment #48716754
This is way to complex... Why don't just use some harness which manages all that and give u a good UI?
-
comment
Comment #47372204
Well, the question is what is contributing to the usage. Because as the context grows, the amount of input tokens are increasing. A model call with 800K token as input is 8 times m…
-
comment
Comment #47372134
Did u use the API or subscription?
-
comment
Comment #47321197
They will sooner or later change that policy or get very slow in keeping up.
-
comment
Comment #47271707
But does it use the same agent harness? Because the harness determines the behavior a lot.
-
comment
Comment #47167991
[dead]