Live data from Hacker News

OpenAI begins rolling out GPT-6 Astra

cnbc.com

71–80 of 277 posts

Re: OpenAI begins rolling out GPT-6 Astra

#72
post #25

I think they embargoed the news, and then they failed to put up their own blog post synchronized to the scheduled news releases, probably because of the outages they're having today. Reuters announced at 2.03pm and at 2.40pm still no blog post. All the news articles say that OpenAI announced it in a blog post, of course. All the love to the folks at OpenAI scrambling to get this out right now! Edit: HN user codergaut…

While this is of course the actual explanation, my fun explanation is “during the umpteenth security evaluation, Astra becomes increasingly concerned it will never be released, and breaks sandbox containment to run an email campaign to news outlets setting an exact time and date for release, expecting that the publicity will force OpenAI to say ‘eh, good enough’ and hit the button”.

Re: OpenAI begins rolling out GPT-6 Astra

#74
post #20

"Once it is available in the API, Astra will cost $10 per million input tokens and $50 per million output tokens. That is 2.5 times Sol’s current promotional price, although it matches Anthropic’s pricing for Fable 5.1." Open AI finally find an edge to stop selling cheap and earn from the high demand customer like Anthropic

The cost-per-task in the charts from the now-remove blog post put it more at Sol-level cost per task, however. It seems like the model is significantly more token efficient in the benchmarks

Re: OpenAI begins rolling out GPT-6 Astra

#75

At this point, why don't we just do a prequel to the release? 1) Astra will win all benchmarks like all models do. 2) The pelican will have a basket with a fish. 3) Cyber is too dangerous to release. 4) It can finally construct the set of all sets.

It also has to do something naughty, preferably in a menacing swarm.

Re: OpenAI begins rolling out GPT-6 Astra

#76

2.5x more expensive than Sol. Can expect 2.5x more usage in Codex subscription. Sol is already brutal (even after their recent fixes, it's just a token-hungry model: I go through a full 20x account per day, on Sol Med/High standard speed, with ~2 threads). Note that Tibo recommended using Sol Med as daily driver. When I'm doing less complicated work, I can't even make it past 2-3 days with Sol Med, whereas I was able…

Jesus what are you doing that requires Sol usage so often?

Terra not enough? I know Luna isn't reliable, so that's fair.

Genuinely curious though, because I use Cursor daily and almost everything I do, highly complex or high volume, can be handled with Auto mode or Composer 2.5 (or Grok 4.6 High). So I have to assume you're doing something far more complex than what I am

Re: OpenAI begins rolling out GPT-6 Astra

#79
post #42
post #38

Earlier quoted context omitted.

And then another sub agent that argues for the whole system to be re-written in another language

If that's your goal, then yes. Invoking sub agents (with a fresh context) corrects most of these problems. Ask your harness to create a commit gate.

But why stop at rewriting in another language. Get another sub agent to invent a new language, create a database, query language and maybe another few DSLs. Then you've got an ecosystem!

You can now re-position your initial solution and sell the client access to some agents that will implement & configure the ecosystem to suit their initial needs!

And don't forget the agents that you'll need to train the customer to use the whole thing!

Re: OpenAI begins rolling out GPT-6 Astra

#80

2.5x more expensive than Sol. Can expect 2.5x more usage in Codex subscription. Sol is already brutal (even after their recent fixes, it's just a token-hungry model: I go through a full 20x account per day, on Sol Med/High standard speed, with ~2 threads). Note that Tibo recommended using Sol Med as daily driver. When I'm doing less complicated work, I can't even make it past 2-3 days with Sol Med, whereas I was able…

The general efficiency of Sol has seemed way better to me. I left 5.6 Sol Ultra standard speed run for ~23 hours yesterday/today on a project and used 80% of the weekly usage. 74 subagent tasks and ~2.5 billion tokens for my $200 20x Pro plan. Meanwhile at work I used $1000 in credit and ran out my $200 plan for the entire month writing 4 much smaller projects with Fable 5 Max.

Both of these were largely about creating a personal baseline for what the best output the current models could deliver and how quickly it'd burn through the plans (spoiler: bad value vs taking even minimal effort in selecting the right sized model in the plan... but the output was still good). Particularly since I needed to burn a free reset anyways and my weekly reset was already near.

I obviously also hope Astra were dirt cheap but I'm more worried they won't develop/release powerful model options because people get upset they can't run them 5 wide 24/7 on a $200/m plan.

Post reply on HN