Live data from Hacker News

Theoretical Bottlenecks for Scaling LLM Inference to Get Higher Token per Second

twitter.com

1–2 of 2 posts