It seems pretty clear that Anthropic models have been specifically trained to be good at generating three.js (JavaScript 3-D Graphics) code, so given current state of AI code generation in general, I don't find three.js models/animations as indicative of anything other than the model's ability to write three.js code. When Fable was first released the day-1 demos of it on Twitter (presumably from people who were given…
Karpathy’s Pelican
151–160 of 461 posts
Re: Karpathy’s Pelican
#152Earlier quoted context omitted.
Well, in that case - if it is just a "hear, hear" - I do not see why the utterance from that actor would be of note. Maybe nozzlegear wanted to suggest some importance on Musk remaining a bet-ter on the general tech, regardless of the competition?
This gets me thinking (although there is not much to divine from three letters). Maybe it's a Jennifer Lawrence GIF-inspired "Yeah, right" then? Downplaying the significance?
> "Yah": slang spelling of the word "yeah" (which of course can also be used ironically)
> Merriam-Webster: "Yah": used to express disgust, contempt, defiance, or derision; probably imitative of the sound of retching
Re: Karpathy’s Pelican
#153Regarding the argument about LLMs having difficulties auditing their work: I wonder whether we are entering the era of throwaway software. Just like cheap plastics and improved processes has enabled us to rapidly manufacture anything we want for a very low price, maybe LLMs give us the same for software. Produce it cheaply and if it breaks throws it away and reproduce it.
I'm sure this will become the standard. And plastic is an excellent analogy. Maybe we can take it a step further and compare it to on-demand 3D printing. Why would anyone still use off-the-shelf software when they can have a system that has access to all data, can transform it into any form, and can export it in any format? After years of thinking that I needed to develop a decent movie management system for my own f…
Re: Karpathy’s Pelican
#154It seems pretty clear that Anthropic models have been specifically trained to be good at generating three.js (JavaScript 3-D Graphics) code, so given current state of AI code generation in general, I don't find three.js models/animations as indicative of anything other than the model's ability to write three.js code. When Fable was first released the day-1 demos of it on Twitter (presumably from people who were given…
Re: Karpathy’s Pelican
#155Re: Karpathy’s Pelican
#156I'd rather have them battle on the topic "Who builds a better Google Wave for LLM chats" to explore the space of how AI studios could be. "Getting Started with Google Wave": https://www.youtube.com/watch?v=eKUAqNGVwX0
i remember this failing, but this product looks pretty useful in 2026
Re: Karpathy’s Pelican
#157I'd rather have them battle on the topic "Who builds a better Google Wave for LLM chats" to explore the space of how AI studios could be. "Getting Started with Google Wave": https://www.youtube.com/watch?v=eKUAqNGVwX0
Re: Karpathy’s Pelican
#158A lot of people are posting here about how bad the end product is, but that is kind of the point. Models have moved beyond generating images to a new kind of benchmark that better exposes understanding of the physical world, and we can use benchmarks like this to measure future progress. (Of course, it will have to be a qualitative/subjective measurement.)
Re: Karpathy’s Pelican
#159I'd rather have them battle on the topic "Who builds a better Google Wave for LLM chats" to explore the space of how AI studios could be. "Getting Started with Google Wave": https://www.youtube.com/watch?v=eKUAqNGVwX0
Re: Karpathy’s Pelican
#160I always thought of the Pelican more of like a gimmicky quick test. There are people who took it as a serious benchmark for overall model performance?
"Draw a pelican on a bicycle" is not a serious benchmark. "Draw an animation of this long ass scene from a movie, and only call me when everything works e2e" can be.
Consuming radium and using uranium glass, that’s what we’re doing.