Live data from Hacker News

GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

twitter.com

121–130 of 219 posts

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#121
post #30

Thoughts[^0] from Theo, who had early access: > It's a damn good model. Not quite as "smart" as Fable, but it is incredibly capable. Fixed all the problems I had with GPT-5.5. > It is incredibly determined. Will run for a day without even using a /goal. It understands subagents incredibly well and is great at orchestrating. It's super pleasant in use cases like OpenClaw and Hermes Agent. It knows iOS dev incredibly w…

> Not quite as "smart" as Fable, but it is incredibly capable. THIS IS BECAUSE GPT-5.6 SOL IS... just a more posttrained version of GPT-5.5, not a brand new bigger model than GPT-5.5. It's not like how Mythos is bigger than Opus. OpenAI switching to Sol/Terra/Luna renaming is just a way to rip off people and charge more usage for the same sized model. GPT-5.6 --------> GPT-5.6 Sol GPT-5.6-mini ---> GPT-5.6 Terra GPT-…

What is the language of those words (sol, terra, luna). It does not seem to be a single language.

Spanish: Sol, tierra, luna

Italian: Sole, terra, luna

Catalan: Sol, terra, lluna

Portuguese: Sol, terra, lua

Might as well call it gelatto, siesta, fiesta if they think it sounds cool.

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#123
post #80
post #45

Earlier quoted context omitted.

> it's pretty clear he has very little understanding of development or engineering I cannot prove it but I have a feeling that you may be conflating "he clearly has different opinions on things I consider non-negotiable" to "he doesn't know what he's talking about". I also watched a lot of his videos. I wildly disagree with him a lot of times, but he has his reasoning, and I can see (and verify!) that those ideas are…

I sort of disagree, the issue is that he like so many professionals (prime agent, being the other) becoming youtubers uses their experiences to make their opinion the only opinion when said opinion is nuanced or plain wrong objectively.

No he usually acknowledges other opinions (including the ones that I share) and tells from his perspective why they are wrong. It feels condescending when you see a face on screen, roasting what you think is right, but I personally could get over it and learn to take the bits that challenge my ideas.

Of course, youtube isn't interactive and when you see something that you think is objectively wrong, your options are writing a comment nobody will read or ignoring it, which is frustrating, but that, in my opinion, doesn't discredit the content producer itself.

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#124
post #99

I know a few of my comments are related to this, but these new names are horrible. Why introduce ANOTHER layer of confusion and drop the mini, nano suffixes that people got used to? How does this go through so many layers of management at a trillion dollar company without who has a say raising this? I simply can't believe how stupid the naming scheme from OpenAI was and continues to be even after they acknowledged it…

Honestly, "mini" and "nano" to me just seem like really awful names from a marketing perspective - they might as well call it "lobotomized crap version of GPT" and "even more lobotomized crap version of GPT". Whereas Sol/Luna/Terra reads more like "GPT for hard/medium/basic problems".

I disagree. I needed a small text to json model that would parse basic info into a json, nothing else. I instantly knew to start with nano and if not good enough, use mini. This was obvious to me, just looking at the name. Now you have to actually know what those names mean. And I can guarantee you they'll add 5 more to create more hype, or make new names in whatever the next-gen-world-breaking-dangerous-model will be.

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#125
post #27

Damn this is exciting. I love that gpt models are much faster, efficient and cheaper than Claude models. They are so fast even on high/xhigh that I don’t find myself using the parallel agent setup anymore much since its cognitively less demanding to just follow along what the model is doing and most tasks it will complete in <5-<10mins anyway.

This is because GPT-5.6 is just a more posttrained version of GPT-5.5, not a bigger model than GPT-5.5. It's not like how Mythos is bigger than Opus. GPT-5.6 --------> GPT-5.6 Sol GPT-5.6-mini ---> GPT-5.6 Terra GPT-5.6-nano ---> GPT-5.6 Luna Two important things to note, if you want to verify what I say/correct me: GPT-5.6 Terra actually scores worse than GPT-5.5 on many benchmarks. It's not GPT-5.5 trained with mor…

> This is because GPT-5.6 is just a more posttrained version of GPT-5.5, not a bigger model than GPT-5.5.

What is this very confident assumption based on?

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#126
post #61

I find codex way more usable. It’s not pretentiously verbose like Claude. It’s also responsive - I can see the progress easily and steer the conversation. With Claude, it might take 15 minutes and I would lose patience.

Both are verbose in their own way, and both - terrible. Claude models love to throw huge blobs of text in architecture planning / interview conversations, but in not a mentally draining language. OpenAI models are more compact, but very dense & formal - they will speak in RFC language for a button that clicks and submits a form. So claude: 10 paragraphs of prose codex: 1 paragraph of jargon over jargon.

I've seen this with GPT, and I usually ask it to put together a more easy to understand document for a specific target audience or reading level and it seems to do okay.

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#127
post #30

Thoughts[^0] from Theo, who had early access: > It's a damn good model. Not quite as "smart" as Fable, but it is incredibly capable. Fixed all the problems I had with GPT-5.5. > It is incredibly determined. Will run for a day without even using a /goal. It understands subagents incredibly well and is great at orchestrating. It's super pleasant in use cases like OpenClaw and Hermes Agent. It knows iOS dev incredibly w…

> Not quite as "smart" as Fable, but it is incredibly capable. THIS IS BECAUSE GPT-5.6 SOL IS... just a more posttrained version of GPT-5.5, not a brand new bigger model than GPT-5.5. It's not like how Mythos is bigger than Opus. OpenAI switching to Sol/Terra/Luna renaming is just a way to rip off people and charge more usage for the same sized model. GPT-5.6 --------> GPT-5.6 Sol GPT-5.6-mini ---> GPT-5.6 Terra GPT-…

Post-training can have big gains no? I don't think the current sizes at ~1T are saturated in intelligence (it's like saying AlphaGo Master is just a post-trained version of AlphaGo Lee)

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#128
post #2

Any previewers have hot takes? I've really preferred gpt-5.5 over Opus 4.8 for data analysis and scientific software work. It seems much more reliable. Fable is unusable for the type of work that I do (due to guardrails). Really looking forward to trying these new OpenAI models out.

For compiler work I found that Sol is noticably better than 5.5 (and I generally use OAI models because I like the Codex app), but Fable was still obviously better.

Better in what way? Does it follow the goals better, does the code produce have higher quality in a testable/maintainable sense or is it just closer to how you would usually program something?

Re: GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday

#130

I’m bouncing back between Codex and Claude like a ping-pong ball. I much prefer the experience using Codex, less verbose and to-the-point I’ve found. But Fable, being as strong as it is, is a big draw for Claude right now. I’ll likely switch back to Codex if 5.6 Sol is comparable.

I personally like Codex much more as vscode integration. It's actually nice to use, looks great, handles images much, much better, can even send me images in the chat to see what it's doing(my side project is based heavily on image processing), and what's most important to me is that /steer actually works. When I type something to Claude mid-task, it'll maybe acknowledge that within next few minutes, although it seems to be actually quicker last few days but still takes a minute or two, whereas Codex will almost instantly read it, switch what it's doing or answer me. It feels much more polished in some ways(though usage page in settings almost never loads for me).
Post reply on HN