Live data from Hacker News

SubQ 1.1 Card: Linear-scaling sparse attention with 98% retrieval at 12M tokens [pdf]

subq.ai

1–2 of 2 posts