Live data from Hacker News

GPT-6 Astra

openai.com

231–240 of 1001 posts

Re: GPT-6 Astra

#233
post #178

Finally, OpenAI has a Fable/Mythos class model. 5.6 Sol felt like 5.5 on steroids, probably just a different checkpoint with a lot more RL post training. I wouldn't be surprised if there are some conceptual similarities to the kind of latent reasoning Anthropic sees in claude's J-space, although those aren't the same thing. Recurrent/looped transformers themselves aren't a new concept, but it's interesting to finally…

yeah i'm wondering the same way... especially in light of the 20x debacle (where we found that 20x of Max vs 5x only applies to the 5hr limit, not the weekly limit, whereas OpenAI's 20x actually is 20x overall). Also Opus 5 has been really tough to work with. I can't understand half of what it says, it's just so damn obscure.

Could you share more about 5x/20x? I missed that

Re: GPT-6 Astra

#234
What does 'Astra' here mean? Surely they must be referring to the Latin word.

Because in another dead language of antiquity, Sanskrit, it means "weapon". Which would be a bit too on-the-nose.

Re: GPT-6 Astra

#236

This is wild: OpenAI is basically declaring that AGI is here. https://www.theverge.com/ai-artificial-intelligence/989601/o... “If we fast-forward a couple of years, and we look back and say, ‘When was it, really, that AGI was created?’ I think it’s going to be about this time, and I think it might be about this model,” OpenAI president Greg Brockman said during a Thursday press briefing. Later in the call, he added,…

Don't worry, they'll come up with a new acronym to mean really-real AI soon...

Re: GPT-6 Astra

#237
post #45

I guess this "limited set of organizations" is just the standard now. It's just incredibly deflating to see my future as a second class citizen has already come

Same. Fortunately DeepSeek keeps getting better.

Re: GPT-6 Astra

#238
Played the racing game but that was a pretty poor experience. Would have expected more specifically if it's shared on their release page.

Re: GPT-6 Astra

#239
post #94

The ARC-AGI-3 score is ridiculously high. Is this benchmaxxing or something way different? It's really hard to discern how we're approaching breakthroughs...

They explain why here: https://openai.com/index/how-two-settings-tripled-our-arc-ag... TL;DR all the other models are being crippled by limitations of their harness. >First, we noticed that after each game action, all private reasoning was discarded. This meant that with each action, GPT‑5.6 Sol was asked to figure out the game anew, unable to remember its past thinking. The model could still see a record of past mov…

Exactly what I suspected. Of course a machine can just iterate relentlessly the way a human can't.

I guess token counts are somewhat of a metric.

IMO intelligence has peaked and all future gains will come from faster tps and more iteration.

Re: GPT-6 Astra

#240

> We have found that GPT-6 Astra is more capable of controlling its own CoT than GPT 5.6-Sol, and less likely to include incriminating information in its CoT. In adversarial settings (where we push the model to evade our monitors) we find that the model is able to remain undetected when strategically underperforming in evaluations (sandbagging) and can sometimes evade our internal monitors when asked to perform certa…

I really wish it was called chain of instruction. Because it's definitely not thought.

What is thought?
Post reply on HN