What's even more noticable is that Anthropic still hasn't responded to the Kimi K3 release or the DeepSeek release.
Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
161–170 of 281 posts
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#162Earlier quoted context omitted.
I don't understand this at all. Whenever I ask Gemini 3.1 Pro Extended, or Claude 5 Max something in chat, the most I ever wait is maybe 30 seconds. Is that really so bad?
If you just want to know "When it's the next full moon", yes. Very bad as google can answer in 1s.
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#163Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#164Earlier quoted context omitted.
I have the opposite problem. I'm not well-calibrated on when I'd want lower reasoning than what's available to me (and how to compare that to lower-tier models). OpenAI now has Luna, Terra and Sol, each at Low, Medium, High and Xhigh, with Pro/Ultra depending on harness and plan. That's ~15 possible combinations of model and reasoning level, and there isn't a satisfactory explanation of which one you want for any par…
I don’t see why they just don’t allow a smaller model to answer the question while letting the bigger one vet it. The vetting can be asynchronous and can be delivered after a few seconds (if it’s an easy query). If it’s a hard query, the UI can show the answer is currently being vetted or something.
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#165Earlier quoted context omitted.
I'm not sure Google has lower talent costs. What makes you think so?
You are right to question that. I am basing that statement off of the many famous engineers that have recently left and were very highly paid. My guess is Google is ceding the frontier and the high salaries that go with it and letting Anthropic/OpenAI fight over the high priced talent that exit. The remaining non-famous engineers will not be able to command celebrity salaries. But this is just speculation and I don't…
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#166Earlier quoted context omitted.
How are the models too dumb? How are they dumber than the average person? I wonder if anyone gave Claude an IQ test (the one for humans).
Yes there are a few sites with IQ test benchmarks. The frontier models come up around 130 or 140 depending on which model/test.
The test wasn't made to accurately measure IQs that high.
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#167Earlier quoted context omitted.
And after half an hour using it he'd just admit that his test was way too simple as these models are still way too dumb
How are the models too dumb? How are they dumber than the average person? I wonder if anyone gave Claude an IQ test (the one for humans).
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#168Earlier quoted context omitted.
And after half an hour using it he'd just admit that his test was way too simple as these models are still way too dumb
How are the models too dumb? How are they dumber than the average person? I wonder if anyone gave Claude an IQ test (the one for humans).
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#169Earlier quoted context omitted.
And after half an hour using it he'd just admit that his test was way too simple as these models are still way too dumb
How are the models too dumb? How are they dumber than the average person? I wonder if anyone gave Claude an IQ test (the one for humans).
Re: Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
#170Earlier quoted context omitted.
Yes there are a few sites with IQ test benchmarks. The frontier models come up around 130 or 140 depending on which model/test.
That tracks, I wonder what the people who say that the models are too dumb expect to see. Miracles?
They are definitely smart enough to be useful, but dumb enough in their weak spots not to deserve the "general intelligence" qualifier.