Viewing profile — comp_raccoon
comp_raccoon
HN member- Joined
- Wed, Nov 16, 2022, 5:06 PM UTC
- HN karma
- 45
- Public activity
- 12 items
- HN profile
- View on Hacker News ↗
About comp_raccoon
No profile information was provided.
Recent public activity
-
comment
Comment #46006081
Olmo author here! Your are absolutely spot on on > It was impossible for me to actually fact-check any of the claims in the response based on the matched training data. this is tru…
-
comment
Comment #46006028
Olmo author here… would be nice to have some more competition!! I don’t like that we are so lonely either. We are competitive with open weights models in general, just a couple poi…
-
comment
Comment #46005995
Olmo author here, but I can help! First release of Qwen 3 left a lot of performance on the table bc they had some challenges balancing thinking and non-thinking modes. VL series ha…
-
comment
Comment #46005940
Olmo author here! Qwenmodels are in general amazing, but 30B is v fast cuz it’s an MoE. MoEs very much on the roadmap for next Olmo.
-
comment
Comment #46005929
Olmo author here! we release all training data and all our training scripts, plus intermediate checkpoints, so you could take a checkpoint, reproduce a few steps on the training da…
-
comment
Comment #46005885
Olmo researcher here. The point of OlmoTrace is not no attribute the entire response to one document in the training data—that’s not how language models “acquire” knowledge, and fi…
-
comment
Comment #41653047
This is correct! we wanted to show that you can use PixMo dataset and our training code to improve any open model, not just ours!
-
comment
Comment #41650734
google image APIs are not great, yeah it’s only for demo, though—checkpoints on huggingface are uncensored.
-
comment
Comment #41650718
it’s coming! just takes a bit more time to properly release it.
- comment
- story
-
comment
Comment #39840788
they have a technical report coming! knowing the team, they will do a great job disclosing as much as possible.