Viewing profile — helloplanets
helloplanets
HN member- Joined
- Tue, Sep 15, 2020, 7:45 AM UTC
- HN karma
- 2,634
- Public activity
- 540 items
- HN profile
- View on Hacker News ↗
About helloplanets
No profile information was provided.
Recent public activity
- comment
- comment
-
comment
Comment #49044849
A model being a distilled version of another specific model is a different thing from using synthetic data off of another model. Anthropic goes to insane lengths to block other lab…
- comment
-
comment
Comment #49041420
In the same vein, Jonathan Blow's modestly named "Preventing the Collapse of Civilization" talk: https://www.youtube.com/watch?v=ZSRHeXYDLko Crazy how time flies, given that talk i…
-
comment
Comment #49040607
Pretty sure Mythos and Fable have way more params, but they've just been able to use the synthetic data off of them to get the leap in quality from Opus. So, not a distilled versio…
-
comment
Comment #49024738
With the new Ultracode modes in Claude Code and Codex it's been taken to the next level. I mean, multi agent systems have been available for a long time, but the newest models seem…
-
comment
Comment #49007363
People don't realize the exponential diffictlty curve with juggling. The highest amount of balls ever juggled is 11.
- story
- comment
-
comment
Comment #48929964
The actual part on fine-tuning seems very short in the article. Did I miss a page where they have examples of fine-tuning it for different niche use cases? Optimizing models to be …
- story
- comment
-
comment
Comment #48896927
Shouldn't the valuation be in Bs instead of Ms?
-
comment
Comment #48888389
Yes, it is a 10x markup on the API prices. Depending on whether you factor in cooling costs, data center staff, etc. Or GPU costs and the electricity the GPUs are using only. Eithe…
-
comment
Comment #48888357
True, it'd be a whole other situation if the tokens limits were cumulative. I guess it would all come down to whether their Claude Code subscription plans are turning in a profit o…
-
comment
Comment #48884360
The subscription based plans are heavily subsidized, but the direct API inference pricing (which larger companies need to pay) is profitable. Using a full Claude Max 20x plan to 10…
- comment
- comment
- comment
- comment
-
comment
Comment #48841274
I just did this on one .claude directory and >20% of the answers there included some variation of "real", "actual", "exact", "honest", "genuine", "valid", "true". ~15% in that dire…
-
comment
Comment #48840587
It's kind of offputting how much Anthropic models these days keep repeating "real", "genuine" and "honest". They've RL'd that way over the top.
- story
-
comment
Comment #48829786
Definitely not just a classifier layered on top, although there is one of those as well. Pretty sure it's different post-training / finetune run and the model weights are different…