Live data from Hacker News

GPT-5.6

openai.com

411–420 of 1001 posts

Re: GPT-5.6

#411
GPT 5.6 Sol is a token hog. After implementing the task, it started some "reviews" I didn't ask for - they consumed 19.5M and 11.9M tokens, while the task itself was below 5M tokens.

Re: GPT-5.6

#412
post #379

Here are 18 pelicans - six each for Luna, Terra and Sol at the six different reasoning effort levels (plus the price to generate each one): https://static.simonwillison.net/static/2026/gpt-5.6-pelican... Or if you want to see some in 3D, OpenAI featured a pelican riding a tricycle, bicycle, pony and another pelican in their livestream this morning: https://www.youtube.com/live/Wq45rvPGNHs?t=1070s

Is the direction of the pelicans encoded in your prompt? Curious why they are all left to right with the exception of terra xhigh.

Re: GPT-5.6

#413
post #379

Here are 18 pelicans - six each for Luna, Terra and Sol at the six different reasoning effort levels (plus the price to generate each one): https://static.simonwillison.net/static/2026/gpt-5.6-pelican... Or if you want to see some in 3D, OpenAI featured a pelican riding a tricycle, bicycle, pony and another pelican in their livestream this morning: https://www.youtube.com/live/Wq45rvPGNHs?t=1070s

I like Terra High the best. That pelican is utterly yoked.

Re: GPT-5.6

#414
post #149
post #6

Ok long time Claude Code user here; lately I've started to realize there's other great models out there I should be trying, but I'm hesitant to leave Claude Code behind for something new. What's the consensus today on codex vs claude code, does it really matter anymore?

Claude Code fan here... Codex is very good. Sometimes better. The killer feature is price. After 6+ months of exclusive Claude Code usage, I was begrudgingly forced to try Codex once Anthropic rejiggered their limits such that I kept maxing out my $200/mo plan in just a few days. These days I pay both $200/mo plans, and it's just about enough to get me through a week's work (small game studio - infinite code to write…

> (small game studio - infinite code to write!)

Curious: what multiplier do you think your productivity has increased by, from before AI?

Re: GPT-5.6

#415

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna, and it just moves further into confusing territory when it comes to naming. The naming convention is especially difficult to decipher depending on what your native language is. Of course a latin language speaker might be able to easily determine oh yeah each one is slightly bigger than the other but I still think it borderlines too confus…

My guess is that it's the same for Haiku/Sonnet/Opus: Biggest model for architecture and high level planning and technically challenging problems, medium model for simple implementation tasks, small model is for nothing

Re: GPT-5.6

#416
post #379

Here are 18 pelicans - six each for Luna, Terra and Sol at the six different reasoning effort levels (plus the price to generate each one): https://static.simonwillison.net/static/2026/gpt-5.6-pelican... Or if you want to see some in 3D, OpenAI featured a pelican riding a tricycle, bicycle, pony and another pelican in their livestream this morning: https://www.youtube.com/live/Wq45rvPGNHs?t=1070s

[deleted]

Re: GPT-5.6

#417

Earlier quoted context omitted.

Hahaha seeing this play out in real time is absolutely incredible.

Im here for it, good on Anthropomorphic to feel some heat again after all that drug dealer Fable business.

100. I'm so tired of being treated like a drug addict by them. I'm currently sitting on _four_ "reset vouchers" from OpenAI so I get basically all week to play with 5.6 Sol to my hearts content. The amount of positive sentiment that brings to me should really be a concern for Anthropic who is increasingly alienating me away with their shit strategies.

Re: GPT-5.6

#419

Huh, a good alternative just as anthropic's 50% weekly subscription subsidy is ending this weekend. Time to see if it's benchmaxxed or actually a strong leap over GPT5.5. They also seem to really not care about alignment, or care about it in the wrong way. It's entirely missing in the blogpost and there are some concerning bits in the model card, seemingly treating CoT controllability as something to be "investigated…

There's also this:

> GPT-5.6 Sol’s detected cheating rate was higher than any public model we have evaluated -- https://www.lesswrong.com/posts/JFjNmPTbH8kL6xtp6/gpt-5-6-th...

Re: GPT-5.6

#420
post #379

Here are 18 pelicans - six each for Luna, Terra and Sol at the six different reasoning effort levels (plus the price to generate each one): https://static.simonwillison.net/static/2026/gpt-5.6-pelican... Or if you want to see some in 3D, OpenAI featured a pelican riding a tricycle, bicycle, pony and another pelican in their livestream this morning: https://www.youtube.com/live/Wq45rvPGNHs?t=1070s

At what stages will models start to internally reflect the drawn SVG and automatically fix their own mistakes?

I assume multimodal models can do it already do it today if constantly asked "make it better"

Post reply on HN