Live data from Hacker News

Claude 3.5 Sonnet

anthropic.com

101–110 of 287 posts

Re: Claude 3.5 Sonnet

#103

On a first glance, CS3.5 appears to be slightly faster than gpt-4o (62 vs 49 tok/sec) and slightlhy less capable (78% vs 89% accuracy on our internal reasoning benchmark). When initially launched, gpt-4o had speed of over 100 tok/sec, surprised that speed went down as fast.

Have you tried our prompt generator? https://docs.anthropic.com/en/docs/build-with-claude/prompt-... . We've seen it improve performance.

The benchmark is imagined as zero shot, so no tweaking.

Re: Claude 3.5 Sonnet

#104

Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…

What I understand is that it's GPT 6 that just went into training, and that GPT 5 is complete and being delayed until after the U.S. election.

And after GPT-5's release, what would be the plan for subsequent elections? This seems to be a temporary play in delaying AI regulation if public sentiment further becomes that AI can have a strong influence in the elections.

Re: Claude 3.5 Sonnet

#106

Earlier quoted context omitted.

Have you tried our prompt generator? https://docs.anthropic.com/en/docs/build-with-claude/prompt-... . We've seen it improve performance.

The benchmark is imagined as zero shot, so no tweaking.

Got it, thanks for the feedback!

Re: Claude 3.5 Sonnet

#107

I wish they'd implement branching conversations like in ChatGPT. And convenient message editing, that doesn't paste large chunks of text as an non-editable attachment or break formatting. Seems like such a simple thing to do, relative to developing an AI, yet the minor differences in the UI/UX are what prevents me from using claude a lot more.

I'm actually working on an open source product to solve this. For a long while I've wanted a good "pro" UI that can connect to multiple different llm APIs Convenient editing and branching is one of the items in my roadmap already, what else do you think I could include?

You might be looking for "LLM Web-UI"s. I searched for a while until I found this thread with recommendations:

https://old.reddit.com/r/LocalLLaMA/comments/1847qt6/llm_web...

Re: Claude 3.5 Sonnet

#108

I know it's nitpicking but please anthropic give non business accounts API access.

They don't require it to be a business for API access, they just have an account structure where you can have several users under one 'organization'. Even a family can be an 'organization', nothing special required.

Re: Claude 3.5 Sonnet

#109

I wish they'd implement branching conversations like in ChatGPT. And convenient message editing, that doesn't paste large chunks of text as an non-editable attachment or break formatting. Seems like such a simple thing to do, relative to developing an AI, yet the minor differences in the UI/UX are what prevents me from using claude a lot more.

Pretty much all of the features you mention are already in LibreChat (MIT License). If you don't mind self-hosting, then it has branching, convo search, change models mid-chat, "presets" (save system prompts), and a whole lot more. I've deployed it in my gov agency for months now, and I've had amazing feedback. https://github.com/danny-avila/LibreChat

Re: Claude 3.5 Sonnet

#110

Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…

What I understand is that it's GPT 6 that just went into training, and that GPT 5 is complete and being delayed until after the U.S. election.

This is pure speculation, right?
Post reply on HN