GPT-6 Astra
171–180 of 1001 posts
Re: GPT-6 Astra
#172Finally, OpenAI has a Fable/Mythos class model. 5.6 Sol felt like 5.5 on steroids, probably just a different checkpoint with a lot more RL post training. I wouldn't be surprised if there are some conceptual similarities to the kind of latent reasoning Anthropic sees in claude's J-space, although those aren't the same thing. Recurrent/looped transformers themselves aren't a new concept, but it's interesting to finally…
Also Opus 5 has been really tough to work with. I can't understand half of what it says, it's just so damn obscure.
Re: GPT-6 Astra
#173GPT 6 Astra benchmarks https://cdn.thenewstack.io/media/2026/09/358eb84a-screenshot... Performance is significantly higher than Fable 5.1 Source: https://thenewstack.io/openai-gpt6-astra-benchmarks/
any benchmark where opus 5 achieves higher scores than fable 5 in any way is not a benchmark worth trusting.
Re: GPT-6 Astra
#174> GPT‑6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API and AWS. Not on Azure? If so, that's a big deal.
They broke up a while ago, why is this surprising?
Re: GPT-6 Astra
#175You should know: AA index is only 61. Pretty surprised it’s that low.
Re: GPT-6 Astra
#176Re: GPT-6 Astra
#177> GPT‑6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API and AWS. Not on Azure? If so, that's a big deal.
They broke up a while ago, why is this surprising?
Re: GPT-6 Astra
#178https://youtu.be/1QNsdr-Qx_I?si=coXwStCl7clpGVC1 Launch video
Re: GPT-6 Astra
#179GPT 6 Astra benchmarks https://cdn.thenewstack.io/media/2026/09/358eb84a-screenshot... Performance is significantly higher than Fable 5.1 Source: https://thenewstack.io/openai-gpt6-astra-benchmarks/
> Performance is significantly higher than Fable 5.1 That's not clear. Need to see independent benchmarks first.
Still below Fable 5, let alone Fable 5.1.
EDIT: This is suspiciously low. Calls the relevance of existing benchmarks into question.