It’s been hard to keep up with the evolution in LLMs. SOTA models basically change every other week, and each of them has its own quirks. Differences in features, personality, output formatting, UI, safety filters… make it nearly impossible to migrate workflows between distinct LLMs. Even models of the same family exhibit strikingly different behaviors in response to the same prompt. Still, having to find each model’…
Claude 4
81–90 of 1001 posts
Re: Claude 4
#82This is starting to get ridiculous. I am busy with life and have hundreds of tabs unread including one [1] about Claude 3.7 Sonnet and Claude Code and Gemini 2.5 Pro. And before any of that Claude 4 is out. And all the stuff Google announced during IO yday. So will Claude 4.5 come out in a few months and 5.0 before the end of the year? At this point is it even worth following anything about AI / LLM? [1] https://news…
Re: Claude 4
#83Re: Claude 4
#84Looks like both opus and sonnet are already in Cursor.
A bit busy at the moment then.
Re: Claude 4
#85Good, I was starting to get uncomfortable with how hard Gemini has been dominating lately ETA: I guess Anthropic still thinks they can command a premium, I hope they're right (because I would love to pay more for smarter models). > Pricing remains consistent with previous Opus and Sonnet models: Opus 4 at $15/$75 per million tokens (input/output) and Sonnet 4 at $3/$15.
Re: Claude 4
#86Is this really worthy of a claude 4 label? Was there a new pre-training run? Cause this feels like 3.8... only swe went up significantly, and that as we all understand by now is done by cramming on specific post training data and doesn't generalize to intelligence. The agentic tooluse didn't improve and this says to me that it's not really smarter.
Re: Claude 4
#87I've found myself having brand loyalty to Claude. I don't really trust any of the other models with coding, the only one I even let close to my work is Claude. And this is after trying most of them. Looking forward to trying 4.
[0] https://mattsayar.com/personalized-software-really-is-coming...
Re: Claude 4
#88[deleted]
Re: Claude 4
#89Wonder why they renamed it from Claude (e.g. Claude 3.7 Sonnet) to Claude (Claude Opus 4).
Re: Claude 4
#90> Finally, we've introduced thinking summaries for Claude 4 models that use a smaller model to condense lengthy thought processes. This summarization is only needed about 5% of the time—most thought processes are short enough to display in full. This is not better for the user. No users want this. If you're doing this to prevent competitors training on your thought traces then fine. But if you really believe this is…