GPT-6 Astra
21–30 of 1001 posts
Re: GPT-6 Astra
#22Re: GPT-6 Astra
#23Re: GPT-6 Astra
#24GPT 6 Astra benchmarks https://cdn.thenewstack.io/media/2026/09/358eb84a-screenshot... Performance is significantly higher than Fable 5.1 Source: https://thenewstack.io/openai-gpt6-astra-benchmarks/
Is the ARC-AGI-3 score with their custom harness? I'm guessing that is what the footnote is for? (per https://openai.com/index/how-two-settings-tripled-our-arc-ag... )
Re: GPT-6 Astra
#25Re: GPT-6 Astra
#26Re: GPT-6 Astra
#27someone screenshot?
Re: GPT-6 Astra
#28I'm seeing reporting it gets 98.6% on ARC-AGI3[1] (previously like 30% with Fable) https://venturebeat.com/technology/welcome-to-the-agi-era-op...
The blog post says 99.9%. Oddly, it does better on ARC-AGI-3 than it does on version 1 or 2 of the same benchmark (though gets 95+ on all three)
Re: GPT-6 Astra
#29Re: GPT-6 Astra
#30$10 per million input tokens and $50 per million output tokens sol is $4 / $20
2.5x more expensive than Sol. Can expect 2.5x more usage in Codex subscription. Sol is already brutal (even after their recent fixes, it's just a token-hungry model: I go through a full 20x account per day, on Sol Med/High standard speed, with ~2 threads). I hope the efficiency gains are true, since their token efficiency claims for Sol were bullshit.