Live data from Hacker News

GPT-6 Astra

openai.com

981–990 of 1001 posts

Re: GPT-6 Astra

#981

I can’t help but notice how much this echoes Francois Chollet’s On the Measure of Intelligence: https://arxiv.org/abs/1911.01547 Most of frontier-model progress still looks like skill acquisition optimization: broader benchmark coverage and performance, more domains absorbed into the training distribution, and increasingly strong performance within that surface area. It seems more about coverage-driven competence. So…

You are conflating multiple things. 1) First, you are talking about positive forward transfer in continual learning. I've been giving talks for the past 6-7 years about how that community (I was one of the founders) went astray and wasn't focusing enough on that topic, but continual learning of the kind you are thinking isn't in any of these systems right now. I think some people left the Grok team to make a start-up…

"Given that this is HN, and a non-trivial number of us have ADHD, creativity is positively correlated with ADHD." Is this a statement of fact? As someone with ADHD and pretty confident in my creativity, I still believe this is much more cope than fact(which could be more a self-doubt thing than anything else).

Further, if there is a correlation, I'd bet it's not so much an intrinsic "creativity" trait, but more effectively higher creativity because more trials. That is, along the lines of Chollet's paper, a measure of creativity should be based on a fixed budget with fixed knowledge.

Among many other possibilities I haven't considered, perhaps another mechanism could be that because ADHD people spend more time thinking in less goal-oriented ways and mixing thoughts on accident, perhaps we do in fact gain some learned creativity via experience with vagueness[1]? But that might also imply that part of creativity is actually being able to diffuse more freely through thought space and lowering the barrier to attempted connections between ideas. That lower barrier leads to less likelihood of any "collision" being meaningful but maybe it's overcome by higher collision rates? Or maybe effectively higher order (not just pairwise) collisions?

Disclaimer in case it's not obvious: I don't know any of the literature on what creativity even means or how it's quantified.

[1] Which is me injecting an assumption that creativity ~= connecting things with no obvious or well-troden reasoning path between them.

edit -- oops just looked at your profile after seeing someone elses comment. I assume you are stating a fact then, leaving original anyway

Re: GPT-6 Astra

#982

Finally, OpenAI has a Fable/Mythos class model. 5.6 Sol felt like 5.5 on steroids, probably just a different checkpoint with a lot more RL post training. I wouldn't be surprised if there are some conceptual similarities to the kind of latent reasoning Anthropic sees in claude's J-space, although those aren't the same thing. Recurrent/looped transformers themselves aren't a new concept, but it's interesting to finally…

> Canceling my Anthropic Max sub when this ships. At this point, it reads like people are cancelling old ones and getting new subscriptions every two to three days, whenever a new ,model drops, and quite possibly by the end of the week they are back to the old provider while still having active subscriptions with at least two to three others. Interesting times.

Especially with these big models chewing up limits fast, I feel good about simultaneously having an Anthropic sub, an OpenAI sub, and an Opencode balance. The models also seem to catch things when code reviewing each other that they don't always catch when a new instance of the same model does a review.

Re: GPT-6 Astra

#983
post #335

It's fun, but every new model release makes me even less interested to create cool stuff. Like, what's the point, if the next AI can do it in 5 seconds?

listen to the DHH round two at Lex Friedman. Dropped a few days ago. Its inspiring

Re: GPT-6 Astra

#985

I can’t help but notice how much this echoes Francois Chollet’s On the Measure of Intelligence: https://arxiv.org/abs/1911.01547 Most of frontier-model progress still looks like skill acquisition optimization: broader benchmark coverage and performance, more domains absorbed into the training distribution, and increasingly strong performance within that surface area. It seems more about coverage-driven competence. So…

Why do you expect models to do "genuinely new" things? 99.9% of real world tasks are extremely repetitive.

Re: GPT-6 Astra

#986
Perfect thread to read through some good use cases instead of the one-shot game plays across x.com

First impression: the model seems kind. Always intriguing

Re: GPT-6 Astra

#987
On the introduction video: The lady asks "make sure that [the presentation] feels really high-end", but something that gets generated without any effort is not high-end anymore. it will be mediocre, or more probably GARBAGE.

Re: GPT-6 Astra

#988
This is great! It can finally substitute the entire upper management!

Their job is easier to automate than an engineer's or any creative's, and considering their salary it's a lot more gain for any company!

Re: GPT-6 Astra

#989
I finally got access. Here are the pelicans! https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

The "max" one at the bottom took 4 minutes 2 seconds and cost 63.206 cents.

For comparison, here those new Astra pelicans are in a grid with the GPT-5.6 pelicans: https://static.simonwillison.net/static/2026/gpt-6-and-5.6-p...

Re: GPT-6 Astra

#990
post #989

I finally got access. Here are the pelicans! https://tools.simonwillison.net/markdown-svg-renderer?url=ht... The "max" one at the bottom took 4 minutes 2 seconds and cost 63.206 cents. For comparison, here those new Astra pelicans are in a grid with the GPT-5.6 pelicans: https://static.simonwillison.net/static/2026/gpt-6-and-5.6-p...

Even the "low" pelican is better than most other models.
Post reply on HN