Unfortunately still thinks "There are two 'r's in the word "raspberry"." The only one that got it right was the basic version of Gemini "There are actually three "r"s in the word "strawberry". It's a bit tricky because the double "r" sounds like one sound, but there are still two separate letters 'r' next to each other." The paid Gemini advanced had "There are two Rs in the word "strawberry"."
Claude 3.5 Sonnet
161–170 of 287 posts
Re: Claude 3.5 Sonnet
#162Re: Claude 3.5 Sonnet
#163Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…
What I understand is that it's GPT 6 that just went into training, and that GPT 5 is complete and being delayed until after the U.S. election.
The reason the name was changed was because there was a big public scare about gpt-5 taking over and so Altman had to promise not to release gpt-5 soon. So they changed the name to gpt-4o (omni). Which is A) obviously dramatically a different architecture, B) a huge step up in capabilities (most still unreleased) C) very general purpose. Because of A) and B), this should obviously be a new major version (5).
Yes, this is speculation, but it's very obvious speculation to me. It's weird for me that most people not only don't share this view but seem to absolutely hate when I say it.
Re: Claude 3.5 Sonnet
#164Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…
Both GPT-4 and 4o have been completely useless for coding in the past couple of weeks for me - constant errors, and not just your typical LLM inaccuracies but incapable of producing a few lines of self-consistent code e.g. defines variables foo on one line and refers to it as bar on the next, or it misspells it as foox.
Re: Claude 3.5 Sonnet
#165Using this is the first time since GPT-4 where I've been shocked at how good a model is. It's helped by how smooth the 'artifact' UI is for iterating on html pages, but I've been instructing it to make a simple web app one bit of functionality at a time and it's basically perfect (and even quite fast). I'm sure it will be like GPT-4 and the honeymoon period will wear off to reveal big flaws but honestly I'd take this…
i don't think the point of an intern is to have them do this kind of work. to me, it's just a side effect if they accomplish anything at all. if we take this to its logical conclusion, without the kind of basic training that comes from internships, where will we be in 5 years?
Re: Claude 3.5 Sonnet
#166Re: Claude 3.5 Sonnet
#167Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…
Re: Claude 3.5 Sonnet
#168Earlier quoted context omitted.
a Kagi Ultimate subscription gets you access to both (plus others) for $25/mo
This is only via API though. There is a level of magic that Claude.ai and ChatGPT bring to the table that makes it worthwhile.
Re: Claude 3.5 Sonnet
#169Just tried it. This is the first model that immediately gives me the correct answer to my test prompt: "Hi , can you give me an exact solution to pi in python?". All other models I've tried first give an approximation, taking several prompts to come to the correct conclusion: it's impossible.
Re: Claude 3.5 Sonnet
#170Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…
I've had way better success with GPT-4o than claude. I wonder why