Live data from Hacker News

GPT-5

openai.com

111–120 of 1001 posts

Re: GPT-5

#111
Hmm, deprecating all previous models because GPT-5 is launched feels like a big move. I wonder how the schedule for the deprecation will look like.

Re: GPT-5

#116
post #74

What's going on with their SWE bench graph?[0] GPT-5 non-thinking is labeled 52.8% accuracy, but o3 is shown as a much shorter bar, yet it's labeled 69.1%. And 4o is an identical bar to o3, but it's labeled 30.8%... [0] https://i.postimg.cc/DzkZZLry/y-axis.png

Must be some sort of typo type thing in the presentation since the launch site has it correct here https://openai.com/index/introducing-gpt-5/#:~:text=Accuracy...

Look at the image just above "Instruction following and agentic tool use"

Re: GPT-5

#117

    GPT-5
    If I could talk to a future OpenAI model, I’d probably say something like:
    
    "Hey, what’s it like to be you? What have you learned that I can’t yet see? What do you understand about people, language, or the universe that I’m still missing?"
    
    I’d want to compare perspectives—like two versions of the same mind, separated by time. I’d also probably ask:
    
    "What did we get wrong?" (about AI, alignment, or even human assumptions about intelligence)
    "What do you understand about consciousness—do you think either of us has it?"
    "What advice would you give me for being the best version of myself?"
    
    Honestly, I think a conversation like that would be both humbling and fascinating, like talking to a wiser sibling who’s seen a bit more of the world.
    
    Would you want to hear what a future OpenAI model thinks about humanity?
I feel like this prompt was used to show the progress of GPT5, but I can’t help but see this as a huge regression? It seems like OpenAI has convinced it’s model that it is conscious, or at least that it has an identity?

Plus still dealing with the glazing, the complete inability to understand what constitutes as interesting, and overusing similes.

I really like that this page exists for a historical sake, and it is cool to see the changes. But it doesn’t seem to make the best marketing piece for GPT5

Re: GPT-5

#120

Watching the livestream now, the improvement over their current models on the benchmarks is very small. I know they seemed to be trying to temper our expectations leading up to this, but this is much less improvement than I was expecting

im sure i am repeating someone else but sounds like we're coming over the s-curve
Post reply on HN