Kimi K3, and what we can still learn from the pelican benchmark
21–30 of 245 posts
Re: Kimi K3, and what we can still learn from the pelican benchmark
#22Re: Kimi K3, and what we can still learn from the pelican benchmark
#23Sorry, how again is this the end of the frontier labs?
Re: Kimi K3, and what we can still learn from the pelican benchmark
#24Another day, another model and another pelican :-) I can't help but wonder where is the trend going? What will we have in five years? Maybe it will all have puttered out, and we will have moved to the next thing? Or maybe the prompt then will be "make a pelican ride a bicycle", and out will come the genetic code for a giant pelican with extremities suitable for a handle bar and pedals, and an inborn affinity to ride…
You are thinking too hard on this. This entire "benchmark" is a performative joke for attention that only works on HN. > What will we have in five years? Maybe it will all have puttered out, and we will have moved to the next thing? We will just have more of the same.
Re: Kimi K3, and what we can still learn from the pelican benchmark
#25K3 is as expensive as Sonnet, not great at writing English, is handing IP back to the Chinese, and once open source will be difficult to run at scale without the compute that OpenAI and Anthropic have largely grabbed. Sorry, how again is this the end of the frontier labs?
Competition is always good.
Re: Kimi K3, and what we can still learn from the pelican benchmark
#26Is there a gallery of all pelicans generated by simon over time?
Re: Kimi K3, and what we can still learn from the pelican benchmark
#27It's incredible Simon still believes pelicans on bikes aren't part of the training set, despite hundreds of them on blogs, forums, and Github. Stuff we put in our company blog shows up known by LLMs 6 months later, and we have 1000x less traffic than Simon's own website
Re: Kimi K3, and what we can still learn from the pelican benchmark
#28K3 is as expensive as Sonnet, not great at writing English, is handing IP back to the Chinese, and once open source will be difficult to run at scale without the compute that OpenAI and Anthropic have largely grabbed. Sorry, how again is this the end of the frontier labs?
Re: Kimi K3, and what we can still learn from the pelican benchmark
#29It's incredible Simon still believes pelicans on bikes aren't part of the training set, despite hundreds of them on blogs, forums, and Github. Stuff we put in our company blog shows up known by LLMs 6 months later, and we have 1000x less traffic than Simon's own website
Re: Kimi K3, and what we can still learn from the pelican benchmark
#30Perhaps more importantly can they do that during reinforcement training. Learning how to critically analyse the appearance of what it generates would be quite useful.
Manually feeding images back to models has been hilariously bad in the past which suggests that relating something it sees to something it wrote is not an ability it is very good at.