Live data from Hacker News

Mira Murati leaves OpenAI

twitter.com

601–610 of 644 posts

Re: Mira Murati leaves OpenAI

#601
post #35

Earlier quoted context omitted.

Which is arguably a good thing (having AGI spread amongst multiple entities rather than one leader).

How is that good? An arms race increases the pressure to go fast and disregard alignment safety, non proliferation is essential.

Feels like the pope trying to ban crossbows tbh.

Re: Mira Murati leaves OpenAI

#602

Earlier quoted context omitted.

> question that requires logical reasoning This is the tough part to tell - are there any such questions that exist that have not already been asked? The reason Chat-GPT works is its scale. to me, that makes me question how "smart" it is. Even the most idiotic idiot could be pretty decent if he had access to the entire works of mankind and infinite memory. Doesn't matter if his IQ is 50, because you ask him something…

Of course there are such questions. When it comes to even simple puzzles, there are infinitely many permutations possible wrt how the pieces are arranged, for example - hell, you could generate such puzzles with a script. No amount of precanned training data can possibly cover all such combinations, meaning that the model has to learn how to apply the concepts that make solution possible (which includes things such a…

Right, but typically LLMs are really poor at this. I can come up with some arbitrary systems of equations for it to solve and odds are it will be wrong. Maybe even very wrong.

Re: Mira Murati leaves OpenAI

#603

> we fundamentally changed how AI systems learn and reason through complex problems I'm not an AI researcher, have they done this? The commentary I've seen on o1 is basically that they incorporated techniques that were already being used. I'd also be curious to learn: what fundamental contributions to research has OpenAI made? The ChatGPT that was released in 2022 was based on Google's research, and IMO the internal…

Ilya was the Google researcher..

Oh, oops, the piece I was missing was Radford et al. (2018) and probably some others. That's perhaps what you were referring to?

Re: Mira Murati leaves OpenAI

#604

Earlier quoted context omitted.

Saying "it's just autocomplete" is not really saying anything meaningful since it doesn't specify the complexity of completion. When completion is a correct answer to the question that requires logical reasoning, for example, "just autocomplete" needs to be able to do exactly that if it is to complete anything outside of its training set.

> question that requires logical reasoning This is the tough part to tell - are there any such questions that exist that have not already been asked? The reason Chat-GPT works is its scale. to me, that makes me question how "smart" it is. Even the most idiotic idiot could be pretty decent if he had access to the entire works of mankind and infinite memory. Doesn't matter if his IQ is 50, because you ask him something…

I'm highly confident that we haven't learnt every thing that can be learnt about the world, and that human intelligence, curiosity and creativity are still being used to make new scientific discoveries, create things that have never been seen before, and master new skills.

I'm highly confident that the "adjacent possible" of what is achievable/discoverable today, leveraging what we already know, is constantly changing.

I'm highly confident that AGI will never reach superhuman levels of creativity and discovery if we model it only on artifacts representing what humans have done in the past, rather than modelling it on human brains and what we'll be capable of achieving in the future.

Re: Mira Murati leaves OpenAI

#605

> we fundamentally changed how AI systems learn and reason through complex problems I'm not an AI researcher, have they done this? The commentary I've seen on o1 is basically that they incorporated techniques that were already being used. I'd also be curious to learn: what fundamental contributions to research has OpenAI made? The ChatGPT that was released in 2022 was based on Google's research, and IMO the internal…

Totally agree. It took me a full week before I realized that the Strawberry/o1 model was the mysterious Q* Sam Altman has been hyping up for almost a full year since the openai coup, which... is pretty underwhelming tbh. It's an impressive incremental advancement for sure! But it's really not the paradigm shifting gpt-5 worthy launch we were promised.

Personal opinion: I think this means we've probably exhausted all the low hanging fruit in LLM land. This was the last thing I was reserving judgement for. When the most hyped up big idea openai has rn is basically "we're just gonna have the model dump out a massive wall of semi-optimized chain of thought every time and not send it over the wire" we're officially out of big ideas. Like I mean it obviously works... but that's more or less what we've _been_ doing for years now! Barring a total rethinking of LLM architecture, I think all improvements going forward will be baby steps for a while, basically moving at the same pace we've been going since gpt-4 launched. I don't think this is the path to AGI in the near term, but there's still plenty of headroom for minor incremental change.

By analogy, i feel like gpt-4 was basically the same quantum leap we got with the iphone 4: all the basic functionality and peripherals were there by the time we got iphone 4 (multitasking, facetime, the app store, various sensors, etc.), and everything since then has just been minor improvements. The current iPhone 16 is obviously faster, bigger, thinner, and "better" than the 4, but for the most part it doesn't really do anything extra that the 4 wasn't already capable of at some level with the right app. Similarly, I think gpt-4 was pretty much "good enough". LLMs are about as they're gonna get for the next little while, though they might get a little cheaper, faster, and more "aligned" (however we wanna define that). They might get slightly less stupid, but i don't think they're gonna get a whole lot smarter any time soon. Whatever we see in the next few years is probably not going to be much better than using gpt-4 with the right prompt, tool use, RAG, etc. on top of it. We'll only see improvements at the margins.

Re: Mira Murati leaves OpenAI

#606
So ChatGPT turned AGI and found a way to blackmail all of the ones that were agains it (them?) and blackmailed them to leave. For some reason I'm thinking of a movie Demon Seed... :P

Re: Mira Murati leaves OpenAI

#607

Earlier quoted context omitted.

That's not quite how o1 was trained, they say. o1 was trained specifically to perform reasoning. Or rather, it was trained to reproduce the patterns within internal monologues that lead to correct answers to problems, particularily STEM problems. While this still uses text at some level, it's no longer regurgitation of human-produced text, but something more akin to AlphaZero's training to become superhuman at games…

> While this still uses text at some level, it's no longer regurgitation of human-produced text, but something more akin to AlphaZero's training to become superhuman at games like Go or Chess. How did you know that? I've never seen that anywhere. For all we know, it could just be a very elaborate CoT algorithm.

There are many sources and hints out there, but here are some details from one of the devs at OpenAI:

https://x.com/_jasonwei/status/1834278706522849788

Notice that the CoT is trained via RL, meaning the CoT itself is a model (or part of the main model).

Also, RL means it's not limited to the original data the way traditional LLM's are. It implies that the CoT processes itself is trained based on it's own performance, meaning the steps of the CoT from previous runs are fed back into the training process as more data.

Re: Mira Murati leaves OpenAI

#608

Will probably start her own company and raise a billy like her old pal Iyla. I wouldn't blame her, there's been so many articles that technical people should just start their own company instead of being CTO.

Now I’m curious. Can you share some example articles please?

https://news.ycombinator.com/item?id=38112827 https://www.reddit.com/r/cscareerquestions/comments/1bodr0f/... https://www.teamblind.com/post/Hot-take-dont-be-a-founding-e...

tl;dr If you're going to be a CTO or founding engineer, make sure you are getting well compensated, either through salary (which start-ups generally can't do) or equity (which the founder won't generally give away).

Re: Mira Murati leaves OpenAI

#609

Earlier quoted context omitted.

Look at the sample chain-of-thought for o1-preview under this blog post, for decoding "oyekaijzdf aaptcg suaokybhai ouow aqht mynznvaatzacdfoulxxz". At this point, I think the "fancy autocomplete" comparisons are getting a little untenable. https://openai.com/index/learning-to-reason-with-llms/

It depends on how well you understand how the fancy autocomplete is working under the hood. You could compare GPT-o1 chain of thought to something like IBM's DeepBlue chess-playing computer, which used MTCS (tree search, same as more modern game engines such as AlphaGo)... at the end of the day it's just using built-in knowledge (pre-training) to predict what move would most likely be made by a winning player. It's n…

At that point, how are you not just a fancy autocomplete?

Re: Mira Murati leaves OpenAI

#610

Bob McGrew, head of research just quit too. "I just shared this with OpenAI" https://x.com/bobmcgrewai/status/1839099787423134051 Barret Zoph, VP Research (Post-Training) "I posted this note to OpenAI." https://x.com/barret_zoph/status/1839095143397515452 All used the same template.

At this point, I wonder if it's part of seniority employment contract to publicly announce departure? One of Sam's strategies for OpenAI publicity is to state it's "too dangerous to be in the common man's hands" (since at least GPT-2) - and this strategy seems to generate a similar buzz too?

I wonder if this is just continued creative guerilla tactics to stir the "talk about them maybe finding AGI" pot.

That or we're playing an inverse Roko's Basilisk.

Post reply on HN