Live data from Hacker News

Claude 3 model family

anthropic.com

11–20 of 723 posts

Re: Claude 3 model family

#11
Now this looks really promising, the only question is if they've taken the constant ridicule by the open LLM community to heart and made it any less ridiculously censored than the previous two.

Re: Claude 3 model family

#13
post #5

[flagged]

GPT-4 was created like 3 years ago internally

the market is evaluating LLMs based on what's actually available. No GPT5 = users go elsewhere.

GPT-4 has little "lock-in" and isn't "good enough" the keep users via inertia.

Re: Claude 3 model family

#14
At this point I wonder how much of the GPT-4 advantage has been OpenAI's pre-training data advantage vs. fundamental advancements in theory or engineering. Has OpenAI mastered deep nuances others are missing? Or is their data set large enough that most test-cases are already a sub-set of their pre-training data?

Re: Claude 3 model family

#16
I don't put a lot of stock on evals. many of the models claiming gpt-4 like benchmark scores feel a lot worse for any of my use-cases. Anyone got any sample output?

Claude isn't available in EU yet, else i'd try it myself. :(

Re: Claude 3 model family

#19
I hate that they require a phone number but this might be the only way to prevent abuse so I'll have to bite the bullet.

> We’ve made meaningful progress in this area: Opus, Sonnet, and Haiku are significantly less likely to refuse to answer prompts that border on the system’s guardrails than previous generations of models.

Finally someone who takes this into account, Gemini and chatGPT is such an obstacle sometimes with their unnecessary refusal because a keyword triggered something.

Post reply on HN