Viewing profile — forrestp
forrestp
HN member- Joined
- Thu, Aug 25, 2022, 4:44 AM UTC
- HN karma
- 44
- Public activity
- 9 items
- HN profile
- View on Hacker News ↗
About forrestp
No profile information was provided.
Recent public activity
-
comment
Comment #44049302
My understanding is that your levers are roughly better / more diverse embeddings or computing more embeddings (embed chunks / groups / etc) + aggregating more cosine similarities …
-
comment
Comment #41811710
It's expensive in this field to verify other people's work. There are a few other papers in the last 3 years that have the same high-level idea but call the anchor tokens something…
-
comment
Comment #41131353
Gradient.ai | SF Bay Area | Onsite/hybrid | Staff SWE | Senior SWE | Enterprise Account Executive Our vision is to power the future of enterprise automation. Gradient is a full sta…
-
comment
Comment #40204188
Right. We are sleep deprived -- couldn't stop over the weekend. Please forgive the typos
-
comment
Comment #40204137
All (training / evals / inference) performed on their L40s clusters. These machines are underrated but capable of serious work
-
comment
Comment #40204080
We are training on top of llama 3. The 256k reasoning benchmarks are on the open LLM leaderboard. And re: token count: our copy was wrong -- it's pre-prepped copy for a model run t…
- story
-
comment
Comment #40164753
Our team @ https://gradient.ai/ has more checkpoints coming soon with longer context lengths.
- story