Live data from Hacker News

Genie 3: A new frontier for world models

deepmind.google

491–500 of 512 posts

Re: Genie 3: A new frontier for world models

#491
post #20

Earlier quoted context omitted.

Bitter lesson strikes again!

_Especially_ given the goal of a world model using a rasters-only frame-by-frame approach. Holy shit.

It is truly remarkable. Even if you expected this to happen eventually, like I did, it seems like your assumptions are 2-10x of the time it’s needed to progress.

It makes me think that Stargate might actually lead to AGI/ASI

Re: Genie 3: A new frontier for world models

#492

Earlier quoted context omitted.

Is it actually unbelievable? It's basically what every major AI lab head is saying from the start. It's the peanut gallery that keeps saying they are lying to get funding.

Even as a layman and AI skeptic, to me this entirely matches my expectations, and something like this seemed like it was basically inevitable as of the first demos of video rendering responding to user input (a year ago? maybe?). Not to detract from what has been done here in any way, but it all seems entirely consistent with the types of progress we have seen. It's also no surprise to me that it's from Google, who I…

Google seems to have had the keys to changing the world years ago and decided not to.

Hard to fault them as the process towards ASI now appears to be runaway and uncontrollable.

Re: Genie 3: A new frontier for world models

#494

Earlier quoted context omitted.

Probably depends on how you engage with GTA. “Drive on the street simulator” along with arrays of weapons and explosions is the majority of my hours in GTA. I despise the creative and artistic vision of GTA online, but I’m clearly in a minority there gauging by how much money they’ve made off it.

I took the "creative and artistic vision" line to refer to the story mode.

I assumed the opposite because I haven't heard about GTA's story in ages, but could be a sampling bias. It's hand-wavy, but last I recall most of the microtransactions didn't show up in single player (like if you bought a car, you couldn't use it in single player) so the people spending money on it are doing it for online, not the story.

I didn't think the story was earth-shattering; it was fine, but no Baldur's Gate.

Edit: In retrospect, the characters were fairly iconic. I still distinctly remember Trevor.

Re: Genie 3: A new frontier for world models

#495
The next version of this, paired with a generative sound model + VR. Matched with the tech that already exists that is able to convert thoughts to words, you could pre-empt what someone wants and display it to them in real-time. Allowing them to explore their conscious and unconscious, meeting and interacting and forming deeper relationships with the aspects of themselves that they had forgotten about.

Re: Genie 3: A new frontier for world models

#496

I wonder how hard it would be to get VR output? That's an insane product right there just waiting to happen. Too bad Google sleeps so hard on the tech they create.

Consistent output and spatial coherence across each eye, maybe a couple years? But meeting head tracking accuracy and latency requirements, I’d bet decades. There’s no way any of this tech reduces end to end latency to acceptable levels, without a massive change in hardware. We’ll probably see someone use reprojection techniques in a year or so and claim they’ve done it. But true generated pixels straight to the head…

I think your timeline is off, at least for a tech demo.

This model already runs at 24fps, and I bet could be made to run at >75fps by scaling hardware and distilling/quantizing the model to only work on certain environments.

The two eye problem seems pretty trivial to me: add another image decoding head with the sole task of decoding the other eye. Training data for this can be plentifully gathered through simulated 3D data, or running existing 2D data (e.g. youtube videos) through slow mono to stereo models. This should add minimal latency as it's another head vs. subsequent layers.

If you can train the model to allow WASD movement + mouse, head tracking is not very different. I think with enough effort we could probably build a VR experience using this today. Getting it onto affordable hardware could be a totally different story, but certainly not decades.

Maybe I'm missing something though!

Re: Genie 3: A new frontier for world models

#500
post #411
post #368

Earlier quoted context omitted.

Your brain does not need to render any environments, just the experience of being in them.

What do you think the difference is?

I guess I mean that we are awake experience the input from our senses, and that in a dream only the replication of the experience of seeing or hearing etc. is needed, not a replication of the input of the senses which then leads to the experience.
Post reply on HN