Live data from Hacker News

GPT-6 Astra

openai.com

991–1000 of 1001 posts

Re: GPT-6 Astra

#991

I think the thing I'm most excited about is the increase in _user prompting_. If I give a poorly constrained/ambiguous prompt, I don't want the model one-shotting assumptions left and right. The demos of Fable/GPT-6 are impressive, but "real AGI" should act more like a collaborator than either a peon or overachiever. It's a tough balance to get right, and although this has been possible to achieve with additional pro…

>>> The demos of Fable/GPT-6 are impressive, but "real AGI" should act more like a collaborator than either a peon or overachiever. I don't really agree. The thing that makes Fable feel like an actual collaborator is its ability to sus out your real intent when you give ambiguous instructions. It's really good at it. I watched some reviews today and came way with the impression that Astra is not better than Sol in th…

It shouldn't make this assumption, it should say "because xyz. You want it committed?".

Re: GPT-6 Astra

#992

Finally, OpenAI has a Fable/Mythos class model. 5.6 Sol felt like 5.5 on steroids, probably just a different checkpoint with a lot more RL post training. I wouldn't be surprised if there are some conceptual similarities to the kind of latent reasoning Anthropic sees in claude's J-space, although those aren't the same thing. Recurrent/looped transformers themselves aren't a new concept, but it's interesting to finally…

> Canceling my Anthropic Max sub when this ships. At this point, it reads like people are cancelling old ones and getting new subscriptions every two to three days, whenever a new ,model drops, and quite possibly by the end of the week they are back to the old provider while still having active subscriptions with at least two to three others. Interesting times.

Yeah I have a main $100 sub, a bunch of $20 subs and sometimes another simultaneous $100 when a really major model happens to drop. For the most part it seems better to have multiple subscriptions than a single $200-$300 one to stay more in touch with state of the art and get a feel for what's good at what.

Re: GPT-6 Astra

#993

I have nothing to say about the actual model, but unrelated--why do so many of these demos include people buying things autonomously? Even if I did trust an AI to get everything right, it's not like the AI can read my mind. If I was ordering food normally and without AI, I would want more control over the process--looking over the options, prices, thinking about what I really want. People don't know what they really…

I think that one thing that "thinkers", academics and other smart people who LIKE to think/select/choose forget is that the majority of the population is not wired this way.

Most people go about their day absolutely minimizing the amount of mental energy they have to spend. They have priorities like kids, work, family, groceries, etc. If anything here can be automated its a fat win in their life. I have friends that are loving the features where meals are planned for them, food is home delivered, Uber is auto ordered, trip plans are made, etc. They don't mind paying more just to reduce mental load. Its a huhe market.

Re: GPT-6 Astra

#994

I can’t help but notice how much this echoes Francois Chollet’s On the Measure of Intelligence: https://arxiv.org/abs/1911.01547 Most of frontier-model progress still looks like skill acquisition optimization: broader benchmark coverage and performance, more domains absorbed into the training distribution, and increasingly strong performance within that surface area. It seems more about coverage-driven competence. So…

You are conflating multiple things. 1) First, you are talking about positive forward transfer in continual learning. I've been giving talks for the past 6-7 years about how that community (I was one of the founders) went astray and wasn't focusing enough on that topic, but continual learning of the kind you are thinking isn't in any of these systems right now. I think some people left the Grok team to make a start-up…

"Given that this is HN, and a non-trivial number of us have ADHD, creativity is positively correlated with ADHD." Is this a statement of fact? As someone with ADHD and pretty confident in my creativity, I still believe this is much more cope than fact(which could be more a self-doubt thing than anything else).

Further, if there is a correlation, I'd bet it's not so much an intrinsic "creativity" trait, but more effectively higher creativity because more trials. That is, along the lines of Chollet's paper, a measure of creativity should be based on a fixed budget with fixed knowledge.

Among many other possibilities I haven't considered, perhaps another mechanism could be that because ADHD people spend more time thinking in less goal-oriented ways and mixing thoughts on accident, perhaps we do in fact gain some learned creativity via experience with vagueness[1]? But that might also imply that part of creativity is actually being able to diffuse more freely through thought space and lowering the barrier to attempted connections between ideas. That lower barrier leads to less likelihood of any "collision" being meaningful but maybe it's overcome by higher collision rates? Or maybe effectively higher order (not just pairwise) collisions?

Disclaimer in case it's not obvious: I don't know any of the literature on what creativity even means or how it's quantified.

[1] Which is me injecting an assumption that creativity ~= connecting things with no obvious or well-troden reasoning path between them.

edit -- oops just looked at your profile after seeing someone elses comment. I assume you are stating a fact then, leaving original anyway

Re: GPT-6 Astra

#995

Finally, OpenAI has a Fable/Mythos class model. 5.6 Sol felt like 5.5 on steroids, probably just a different checkpoint with a lot more RL post training. I wouldn't be surprised if there are some conceptual similarities to the kind of latent reasoning Anthropic sees in claude's J-space, although those aren't the same thing. Recurrent/looped transformers themselves aren't a new concept, but it's interesting to finally…

> Canceling my Anthropic Max sub when this ships. At this point, it reads like people are cancelling old ones and getting new subscriptions every two to three days, whenever a new ,model drops, and quite possibly by the end of the week they are back to the old provider while still having active subscriptions with at least two to three others. Interesting times.

Especially with these big models chewing up limits fast, I feel good about simultaneously having an Anthropic sub, an OpenAI sub, and an Opencode balance. The models also seem to catch things when code reviewing each other that they don't always catch when a new instance of the same model does a review.

Re: GPT-6 Astra

#996
post #335

It's fun, but every new model release makes me even less interested to create cool stuff. Like, what's the point, if the next AI can do it in 5 seconds?

listen to the DHH round two at Lex Friedman. Dropped a few days ago. Its inspiring

Re: GPT-6 Astra

#998

I can’t help but notice how much this echoes Francois Chollet’s On the Measure of Intelligence: https://arxiv.org/abs/1911.01547 Most of frontier-model progress still looks like skill acquisition optimization: broader benchmark coverage and performance, more domains absorbed into the training distribution, and increasingly strong performance within that surface area. It seems more about coverage-driven competence. So…

Why do you expect models to do "genuinely new" things? 99.9% of real world tasks are extremely repetitive.

Re: GPT-6 Astra

#999
Perfect thread to read through some good use cases instead of the one-shot game plays across x.com

First impression: the model seems kind. Always intriguing

Re: GPT-6 Astra

#1000
On the introduction video: The lady asks "make sure that [the presentation] feels really high-end", but something that gets generated without any effort is not high-end anymore. it will be mediocre, or more probably GARBAGE.
Post reply on HN