Live data from Hacker News

Viewing profile — somnial

somnial

HN member
Joined
Thu, Oct 30, 2025, 10:18 AM UTC
HN karma
10
Public activity
16 items

About somnial

No profile information was provided.

Recent public activity

  1. story
  2. comment
    Comment #49171387

    throughput scales superlinearly with number of GPUs when networked well and deployed with wideEP, so 1x won't compare. also it would be interesting to figure from the DSpark paper …

  3. story
  4. story
  5. story
  6. comment
    Comment #48699802

    this is a blog post from a company that hosts open weights LLMs ( https://www.doubleword.ai/ ). I think its possible it might have been tongue in cheek

  7. comment
    Comment #48491672

    https://fergusfinn.com/blog/economics-of-speculative-decodin... good point tho - plus for Deepseek the shared expert increases the overlap slightly

  8. comment
    Comment #48443249

    true, but no reason the predictor model couldn't use linear attention (i.e. mamba, GDN etc) to predict KV caches

  9. story
  10. story
  11. story
  12. story
  13. story
  14. story
  15. story
  16. story