SubQ: a sub-quadratic LLM with 12M-token context
1–10 of 47 posts
Re: SubQ: a sub-quadratic LLM with 12M-token context
#2Re: SubQ: a sub-quadratic LLM with 12M-token context
#3Re: SubQ: a sub-quadratic LLM with 12M-token context
#4Re: SubQ: a sub-quadratic LLM with 12M-token context
#5Looks like long context isn’t a problem anymore
Re: SubQ: a sub-quadratic LLM with 12M-token context
#6It seems at or above SOTA on the given benchmarks, doesn’t have context rot, is orders of magnitude faster, and uses less compute that current transformer models. I suppose it’s just an announcement and we can’t test it ourselves yet.
Re: SubQ: a sub-quadratic LLM with 12M-token context
#7Re: SubQ: a sub-quadratic LLM with 12M-token context
#8I’m very surprised this isn’t getting more attention. Am I missing something? It seems at or above SOTA on the given benchmarks, doesn’t have context rot, is orders of magnitude faster, and uses less compute that current transformer models. I suppose it’s just an announcement and we can’t test it ourselves yet.
Re: SubQ: a sub-quadratic LLM with 12M-token context
#9I’m very surprised this isn’t getting more attention. Am I missing something? It seems at or above SOTA on the given benchmarks, doesn’t have context rot, is orders of magnitude faster, and uses less compute that current transformer models. I suppose it’s just an announcement and we can’t test it ourselves yet.
Re: SubQ: a sub-quadratic LLM with 12M-token context
#10[flagged]