No post body was provided.
Untitled topic
1–10 of 55 posts
Re: undefined
#2[stub]
[flagged]
Re: undefined
#3[stub]
[flagged]
Re: undefined
#4Re: undefined
#5Calling beam search 'AI' is doing a lot of heavy lifting here. This is just superoptimization with a very expensive heuristic function.
Re: undefined
#6[stub]
[deleted]
Re: undefined
#7Re: undefined
#8Very interesting research on this, keen to colab with you folks, I've been building a few experiments for old GTX GPUs to extend lifetime of them with matching performance of tokens for Smol, igor [] autohand.ai let's chat.
Re: undefined
#9Earlier quoted context omitted.
[dead]
Whenever I see '17x faster than experts,' I read 'the experts didn't actually try very hard on the baseline.'
Well, most of our results are not 17x. But still (IMO) solid across the board!
Also, the 17x came from a pretty obscure fusion optimization that isn't called out anywhere in the documentation (we had to run the profiler to see what was actually going on). Wouldn't be surprised if whoever within AWS wrote the kernel didn't know about that optimization.
Re: undefined
#10[stub]
[flagged]