Live data from Hacker News

Karpathy’s Pelican

twitter.com

141–150 of 461 posts

Re: Karpathy’s Pelican

#141
post #46

We must not be using the same opus 5, because if I tried to generate this it would refuse based on copyright grounds.

Karpathy works for Anthropic so he obviously has full access to everything, and doesn’t have to pay for token burn either.

I saw that "~free" and laughed.

Re: Karpathy’s Pelican

#142

IMO, the area where AI is going to be most useful over the next couple years is in developing manufacturing processes top to bottom. Maybe a million token budget is too small, but something like "design me a sneaker and all the equipment to manufacture it autonomously".

They can't even run a vending machine. I don't want to do a shallow dismissal, but I think there's a gulf between my understanding and yours. I hope it's me so I learn something.

I think LLMs will be excellent glue of "find the right function/button and run/push it" but design without constraints and they just explode immediately

Re: Karpathy’s Pelican

#143
post #136

If you like this check: https://news.ycombinator.com/item?id=47400868 it can be used to generate animations (not games) as well.

I know this is a plug but I thought of it, too! I hope this is the next benchmark they saturate.

I like where Karpathy is going; I had the same thoughts about LLM generated slop scenery. I just want some variety of scenery for the goblins to get massacred in in whatever fantasy slop game I play.

Re: Karpathy’s Pelican

#144
Is anyone else getting "mongodb is webscale" vibes? (Except 16 years ago that was a lot smoother, because it used some sort of "render this conversation" engine...)

Re: Karpathy’s Pelican

#145
post #119

I'm pretty tired of the "Y made this game in Z tokens" all over the internet last week. They look impressive, and it's cool that it's even possible, but they're useless as games. None of them are any fun. They're like the most boring variant of basic controllers you can imagine. None have any cool mechanics. None have any tweaks made from hours and hours of testing. All have the same cel-shader.

Right, but the new phenomena is that we've lost a signal. Used to be if a game looked that good it probably had time spent on the game part too.

I've been hearing similar from a while back, but the blame wasn't being directed at AI, it was directed at Unreal's defaults now making random indies look like AAA in the screenshots.

(It's been a while since I was in the game industry, so IDK quite how accurate this is).

Re: Karpathy’s Pelican

#146

Earlier quoted context omitted.

you're tired because you lack imagination.

It’s the people spamming these shitty games that lack imagination.

people are excited, let them be excited! but more than the excitement is realizing the implications of what this means for building software, first it looks like a toy and then it doesn't. the OP is talking about how the games are not playable and fun. who cares? that's not the point.

Re: Karpathy’s Pelican

#147
This is equally bad as a pelican test.

LLMs should be tested in the same way people should be tested for a job interview (but often aren’t) - with tasks RELEVANT to usage.

So you don’t just randomly pick some random thing to make the LLM randomly do (like many job interviewers do).

You start with clear statements about real world usage scenarios. THEN you come up with tests that give insight to how well the LLM/hob seeker gets the job done.

Please, stop coming up with random tests like it’s Microsoft in 1990 and you’re asking job seekers how the would move Mount Fuji, as a way of assessing their programming skills.

No stupid irrelevant pelicans on bicycles and no stupid renderings of Lord Of The Rings. Unless those are relevant use cases.

Any test that anyone comes up with must clearly state the context and how the outcome is measured.

Re: Karpathy’s Pelican

#148

It seems pretty clear that Anthropic models have been specifically trained to be good at generating three.js (JavaScript 3-D Graphics) code, so given current state of AI code generation in general, I don't find three.js models/animations as indicative of anything other than the model's ability to write three.js code. When Fable was first released the day-1 demos of it on Twitter (presumably from people who were given…

Taking a single paragraph of literary text, which is abstract and ambiguous, and converting it into a 3D animation requires an enormous amount of implicit knowledge about spatial relationships, intuitive physics, everyday objects, and so forth. Not to mention the mathematics of 3D transformations and computer graphics more generally. Saying that it's indicative of no more than three.js coding ability is absurd.

Re: Karpathy’s Pelican

#150

IMO, the area where AI is going to be most useful over the next couple years is in developing manufacturing processes top to bottom. Maybe a million token budget is too small, but something like "design me a sneaker and all the equipment to manufacture it autonomously".

> IMO, the area where AI is going to be most useful over the next couple years is in developing manufacturing processes top to bottom.

Why do you say this? I have some experience of manufacturing processes and am not seeing where AI would be useful other than to drive the robots which we can already do quite well without AI (see all the dark/lights-out factories that already exist).

Where do you see it being useful? An example would be nice.

Post reply on HN