Earlier quoted context omitted.
How does Grok 4.5 compare to Opus >= 4.8 though? I'm willing to pay 2x for a 10% smarter model. Intelligence matters that much (because 10% smarter probably saves, on average, several hours of human time).
It's a bit worse. I haven't tried so it's pure speculation based on benchmarks, but I'd assume Grok 4.6 is around Opus 4.8 in real world use, but clearly below Opus 5.
For building a full stack custom CRM and media pipeline tool with video conversion, transcription, and indexing. Supabase, AWS, Meili, NextJS, GCS - lots of surfaces and planes.
4.8 basically couldn't do it, I abandoned the project as the fallback was, "current business processes".
With F5 it's been 4 weeks and almost ready for production release.