Is it bad that I hope it's not a significant improvement in coding?
GPT-5
71–80 of 1001 posts
Re: GPT-5
#72Is it bad that I hope it's not a significant improvement in coding?
Re: GPT-5
#73Re: GPT-5
#74GPT-5 non-thinking is labeled 52.8% accuracy, but o3 is shown as a much shorter bar, yet it's labeled 69.1%. And 4o is an identical bar to o3, but it's labeled 30.8%...
Re: GPT-5
#75Is it bad that I hope it's not a significant improvement in coding?
there is some improvements in some benchs and not else worthy of note in coding. i only took a peek though so i might be wrong
Re: GPT-5
#76Re: GPT-5
#77The silent victory here is this seems like it is being built to be faster and cheaper than o3 while presenting a reasonable jump, which is an important jump in scaling law On the other hand if it's just getting bigger and slower it's not a good sign for LLMs
Re: GPT-5
#7888.0 on Aider Polygot
not bad i guess
Re: GPT-5
#79Seems LLMs really hit the wall.
Re: GPT-5
#80The marketing copy and the current livestream appear tautological: "it's better because it's better." Not much explanation yet why GPT-5 warrants a major version bump. As usual, the model (and potentially OpenAI as a whole) will depend on output vibe checks.