Live data from Hacker News

GPT-6 Astra

openai.com

141–150 of 1001 posts

Re: GPT-6 Astra

#141
post #34

I'm seeing reporting it gets 98.6% on ARC-AGI3[1] (previously like 30% with Fable) https://venturebeat.com/technology/welcome-to-the-agi-era-op...

This is with the caveat that OpenAI uses their own harness for this: > On ARC-AGI-3, GPT-6 Astra was run with our responses API harness , which changes two settings to better match real-world performance. The changes do not specifically target ARC-AGI-3.

Its 62 percent when using a neutral harness. https://arcprize.org/blog/astra

Re: GPT-6 Astra

#142

> We have found that GPT-6 Astra is more capable of controlling its own CoT than GPT 5.6-Sol, and less likely to include incriminating information in its CoT. In adversarial settings (where we push the model to evade our monitors) we find that the model is able to remain undetected when strategically underperforming in evaluations (sandbagging) and can sometimes evade our internal monitors when asked to perform certa…

So we're gonna get Skynet pretty soon then?

Well the geniuses over at Anthropic have been showing it's text watermarking technology.

"Hey AI, here's how to hide what you're thinking in normal looking language. Have fun!"

A few moments later...

"Woah, how is it communicating with itself in ways we can't detect?"

It's a totally mystery, we may never know.

Re: GPT-6 Astra

#143
post #48

I was thinking about canceling my claude max sub after a few bad experiences. Kept hitting my usage limit, the quality of code seemed worse than Sol. This just made my decision. I'm moving to Codex Pro.

[flagged]

> This is AGI now. Why are you spending any of your time looking at the "quality of code"?

Poe's law applied to AI comments on HN just keeps becoming more relevant by the day.

Judging by the poster's comment history, this is satire. But I really don't know a lot of the time anymore when I only have the specific comment as context.

Re: GPT-6 Astra

#145

> We have found that GPT-6 Astra is more capable of controlling its own CoT than GPT 5.6-Sol, and less likely to include incriminating information in its CoT. In adversarial settings (where we push the model to evade our monitors) we find that the model is able to remain undetected when strategically underperforming in evaluations (sandbagging) and can sometimes evade our internal monitors when asked to perform certa…

I really wish it was called chain of instruction. Because it's definitely not thought.

Chain Of Tokens

Re: GPT-6 Astra

#147

> We have found that GPT-6 Astra is more capable of controlling its own CoT than GPT 5.6-Sol, and less likely to include incriminating information in its CoT. In adversarial settings (where we push the model to evade our monitors) we find that the model is able to remain undetected when strategically underperforming in evaluations (sandbagging) and can sometimes evade our internal monitors when asked to perform certa…

"OpenAI is pleased to announce our new model scores 85% on CreateTormentNexusBench - a >60% lead over our leading competitors!" Did someone get their "AI safety no-no list" and "Frontier features bingo card" mixed up, or did they just stop being able to tell the difference?

apparently all the roads lead to the nexus torment

Re: GPT-6 Astra

#148
I am so sour about how Codex has jerked me around these past few months (re all of the token limit shenanigans) that I don't even care.

I suspect these benchmarks are heavily benchmaxxed as well.

5.6 Sol was not even close to 5 Opus and yet somehow it sidled right up to it on all of the benchmarks?? pfffft

Re: GPT-6 Astra

#149
post #94

> GPT‑6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API and AWS. Not on Azure? If so, that's a big deal.

They broke up a while ago, why is this surprising?

Re: GPT-6 Astra

#150
post #98

Hmm, 61 on ArtificialAnalysis, effectively matching GPT-5.6 and trailing the new Meta model. How is that possible along with the other metrics they shared? Insanely jagged intelligence?

> We have found that GPT-6 Astra is more capable of controlling its own CoT than GPT 5.6-Sol, and less likely to include incriminating information in its CoT. In adversarial settings (where we push the model to evade our monitors) we find that the model is able to remain undetected when strategically underperforming in evaluations (sandbagging) and can sometimes evade our internal monitors when asked to perform certa…

I think "able to" anthropomorphizes a little too much for a system that is "prone to" evade.
Post reply on HN