Viewing profile — cs-fan-101
cs-fan-101
HN member- Joined
- Tue, Mar 28, 2023, 6:06 PM UTC
- HN karma
- 67
- Public activity
- 16 items
- HN profile
- View on Hacker News ↗
About cs-fan-101
No profile information was provided.
Recent public activity
- story
- story
- story
- story
- story
- story
-
story
Jais – the world’s most advanced Arabic large language model
Cerebras, G42's Inception, and MBZUAI are pleased to announce Jais, the world’s best-performing Arabic LLM. Jais is a 13B parameter model that was trained on a new 395 billion toke…
-
comment
Comment #36853064
Cerebras and Opentensor are pleased to announce BTLM-3B-8K (Bittensor Language Model), a new state-of-the-art 3 billion parameter open-source language model that achieves breakthro…
- story
-
comment
Comment #36800111
[Cerebras employee here] Condor Galaxy 1 can support beyond 600 billion parameters. In standard config its 600B but it can scale to train upwards of 100T parameter models
-
comment
Comment #36799873
Cerebras announced today that it has built and sold a 4 exaFLOPS AI Supercomputer, named Condor Galaxy 1 (CG-1), to its strategic partner G42, the Abu Dhabi-based AI pioneer. Locat…
- story
-
comment
Comment #35487847
Recently, we announced in this post ( https://news.ycombinator.com/item?id=35343763#35345980 ) the release of Cerebras-GPT — a family of open-source GPT models trained on the Pile …
- story
-
comment
Comment #35443633
Simply focusing on the "better in every regard" part of the comment. One example where Cerebras systems perform well is when a user is interested in training models that require lo…
-
comment
Comment #35346149
Someone posted this repost from the Cerebras Discord earlier, but sharing for visibility - "We chose to train these models to 20 tokens per param to fit a scaling law to the Pile d…