Live data from Hacker News

Nvidia releases its own brand of world models

techcrunch.com

41–50 of 74 posts

Re: Nvidia releases its own brand of world models

#41
post #15

What stops Nvidia from cutting out the middlemen? They have the chips.

Given they are releasing models, whose to say that they don't have teams working -- not secretly -- just not in the open?

They literally have paid platforms to use their models.

Re: Nvidia releases its own brand of world models

#43

Earlier quoted context omitted.

Good enough for what? I've been playing around with local models that often get mentioned here (llama3.3, mistral, etc.) and they routinely provide incorrect code that does not even compile or implements algorithms that have nothing to do with the task at hand, generate invalid JSON like '{ "foo": bar - 42 }', write nonsensical statements like "CR1616 has double the capacity of CR2032", etc. I'm yet to find a useful…

Yeah, I would like to know this. From my perspective, even frontier models by the big players (4o, 3.5 Sonnet) can be unreliable at times, and are at best just walking the line of usefulness for a lot of "exact" tasks (for me: programming, approximation and back-of-the-envelope calculations, expertise on subjects I'm unfamiliar with, a better Google, etc.). The only deal that would make sense for me is to get somethi…

o1 is 10x better than 4o and 3.5 Sonnet for non-trivial coding tasks. I don't even bother with the other models for coding-related tasks, it really is a big difference.

Re: Nvidia releases its own brand of world models

#44

The AI race is one of the most impressive examples of Capitalism making the market efficient. Or at least I've witnessed in my life. We went from Google having complete control. To Open AI releasing GPT2 which really inspired a lot of people to try it. Then GPT3+ convinced the world to try it. After that, Gemini, LLaMa, every type of fine-tune... The noteworthy thing is that LLaMA was good enough that ChatGPT had com…

> The competition has been the best type of brutal. I’d counter with “The competition has been the worst type of brutal.” https://finance.yahoo.com/news/ai-destroyed-google-promise-c... Tech’s quest for the best chatbot so they get all the grift dollars has torched climate change progress.

"Torched climate change progress" is a wildly overblown take: https://www.sustainabilitybynumbers.com/p/ai-energy-demand

Re: Nvidia releases its own brand of world models

#45

What stops Nvidia from cutting out the middlemen? They have the chips.

Nothing, in the same way that TSMC could cut Nvidia and OpenAI out.

Vertical integration is incredibly powerful, but it requires mastery of the whole stack.

Does Nvidia understand AI consumers the way OpenAI / Anthropocene does? Do they have the distribution channels? Can they operate services at that scale?

If they can truly do it all and the middlemen don’t add any unique value, Nvidia (or TSMC) can and should make an integration play. TBB I’m skeptical though.

Re: Nvidia releases its own brand of world models

#46

Earlier quoted context omitted.

Yeah, I would like to know this. From my perspective, even frontier models by the big players (4o, 3.5 Sonnet) can be unreliable at times, and are at best just walking the line of usefulness for a lot of "exact" tasks (for me: programming, approximation and back-of-the-envelope calculations, expertise on subjects I'm unfamiliar with, a better Google, etc.). The only deal that would make sense for me is to get somethi…

o1 is 10x better than 4o and 3.5 Sonnet for non-trivial coding tasks. I don't even bother with the other models for coding-related tasks, it really is a big difference.

Yep. Calling 4o a “frontier model” when o1 is available seems questionable.

I only use 4o when I need web search incorporated in results.

Re: Nvidia releases its own brand of world models

#47
post #45

What stops Nvidia from cutting out the middlemen? They have the chips.

Nothing, in the same way that TSMC could cut Nvidia and OpenAI out. Vertical integration is incredibly powerful, but it requires mastery of the whole stack. Does Nvidia understand AI consumers the way OpenAI / Anthropocene does? Do they have the distribution channels? Can they operate services at that scale? If they can truly do it all and the middlemen don’t add any unique value, Nvidia (or TSMC) can and should make…

> Nothing, in the same way that TSMC could cut Nvidia and OpenAI out.

one thing is not like the other.

Re: Nvidia releases its own brand of world models

#48
post #22

The AI race is one of the most impressive examples of Capitalism making the market efficient. Or at least I've witnessed in my life. We went from Google having complete control. To Open AI releasing GPT2 which really inspired a lot of people to try it. Then GPT3+ convinced the world to try it. After that, Gemini, LLaMa, every type of fine-tune... The noteworthy thing is that LLaMA was good enough that ChatGPT had com…

I'm not sure how a few multinational mega corporations "competing" with each other is an impressive example of capitalist market efficiency. After all, this is isn't GPTx vs Gemini vs Llama vs Claude - it's Microsoft vs Google vs Meta vs Amazon. None of which are fair actors in the global marketplace.

Wait, which megacorp is Claude?

Re: Nvidia releases its own brand of world models

#49

Earlier quoted context omitted.

Yeah, I would like to know this. From my perspective, even frontier models by the big players (4o, 3.5 Sonnet) can be unreliable at times, and are at best just walking the line of usefulness for a lot of "exact" tasks (for me: programming, approximation and back-of-the-envelope calculations, expertise on subjects I'm unfamiliar with, a better Google, etc.). The only deal that would make sense for me is to get somethi…

o1 is 10x better than 4o and 3.5 Sonnet for non-trivial coding tasks. I don't even bother with the other models for coding-related tasks, it really is a big difference.

For isolated coding tasks I've been using o1-preview instead of Sonnet for a while now, I just didn't mention it. Haven't had a chance to test o1 proper, but I assume it's also a jump in performance. However, for more "holistic" tasks which need to take into account a larger view of some other modules/systems/interfaces, I've found o1-preview can get really confident about weirdly incorrect things that end up being harder to debug than the more straightforward hallucinations of Sonnet, and so I mostly revert to Sonnet in those cases.

I tried not to make too big a fuss about the exact models I'm mentioning, since it's pretty clear that the strongest open model, Llama (discussed here), is not comparable to inference-time compute models.

And for agentic settings which I mentioned above, o1 just tips way too much into expensive & slow territory to make it useful, so my prediction is that it will be limited for direct (chat-based) consumer use for the time being.

Re: Nvidia releases its own brand of world models

#50

Earlier quoted context omitted.

> The competition has been the best type of brutal. I’d counter with “The competition has been the worst type of brutal.” https://finance.yahoo.com/news/ai-destroyed-google-promise-c... Tech’s quest for the best chatbot so they get all the grift dollars has torched climate change progress.

"Torched climate change progress" is a wildly overblown take: https://www.sustainabilitybynumbers.com/p/ai-energy-demand

It’s well known that humans have a tendency to hallucinate!
Post reply on HN