Live data from Hacker News

Claude 3.7 Sonnet and Claude Code

anthropic.com

371–380 of 1001 posts

Re: Claude 3.7 Sonnet and Claude Code

#371

Kagi LLM benchmark updated with general purpose and thinking mode for Sonnet 3.7. https://help.kagi.com/kagi/ai/llm-benchmark.html Appears to be second most capable general purpose LLM we tried (second to gemini 2.0 pro, in front of gpt-4o). Less impressive in thinking mode, about at the same level as o1-mini and o3-mini (with 8192 token thinking budget). Overall a very nice update, you get higher quality and higher…

How did you chose the 8192 token thinking budget? I've often seen Deepseek R1 use way more than that.

Re: Claude 3.7 Sonnet and Claude Code

#372

You can get your HN profile analyzed by it and it's pretty funny :) https://hn-wrapped.kadoa.com/ I'm using this to test the humor of new models.

> You built your own Klaviyo alternative to save €500, but how many hours of development at market rate did that cost? The true Greek economy at work!

ouch (ㅠ﹏ㅠ)

Re: Claude 3.7 Sonnet and Claude Code

#375

You can get your HN profile analyzed by it and it's pretty funny :) https://hn-wrapped.kadoa.com/ I'm using this to test the humor of new models.

> For someone who worked at Reddit, you sure spend a lot of time on HN. It's like leaving Facebook to spend all day on Twitter complaining about social media. Wow, so spot on it hurts!

>Your ideal tech stack is so old it qualifies for social security benefits

>You're the only person who gets excited when someone mentions Trinity Desktop Environment in 2025

> You probably have more opinions about PHP's empty() function than most people have about their entire career choices

Re: Claude 3.7 Sonnet and Claude Code

#376
post #350

Earlier quoted context omitted.

It is API billing like AWS - you pay for what you use. Every time you exit a session we print the cost, and in the middle of a session you can do /cost to see your cost so far that session! You can track costs in a few ways and set spend limits to avoid surprises: https://docs.anthropic.com/en/docs/agents-and-tools/claude-c...

Which is theoretically great, but if anyone can get an Aussie credit card to work, please let me know.

I haven’t had an issue with Aussie cards?

But I still hit limits, I use Claudemind with jetbrains stuff and there is a max of input tokens (j believe), I am ‘tier 2’ but doesn’t look like I can go past this without an enterprise agreement

Re: Claude 3.7 Sonnet and Claude Code

#377
post #341

Earlier quoted context omitted.

If you are open to alternatives, try https://glama.ai/gateway We currently serve ~10bn tokens per day (across all models). OpenAI compatible API. No rate limits. Built in logging and tracing. I work with LLMs every day, so I am always on top of adding models. 3.7 is also already available. https://glama.ai/models/claude-3-7-sonnet-20250219 The gateway is integrated directly into our chat ( https://glama.ai/chat ). So…

Do you have deepseek r1 support? I need it for a current product I’m working on.

They are just selling a frontend wrapper on other people's services, so if someone else offers deepseek, I'm sure they will integrate it.

Re: Claude 3.7 Sonnet and Claude Code

#378
post #150

Claude 3.5 sonnet has been my go to for coding tasks, it’s just so much better than the others. but I’ve tried using the api in production and had to drop it due to daily issues: https://status.anthropic.com/ compare to https://status.openai.com/ any idea when we’ll see some improvements in api availability or will the focus be more on the web version of claude?

Err, if you compare the two consoles you'll see that anthropic is actually slightly better on average than openai's uptime.

click on individual days. you’ll notice that there are daily errors.

Re: Claude 3.7 Sonnet and Claude Code

#379

You can get your HN profile analyzed by it and it's pretty funny :) https://hn-wrapped.kadoa.com/ I'm using this to test the humor of new models.

This thing is hilarious. :)

Roast:

- Your comments have more doom predictions than a Y2K convention in December 1999.

- You've used 'stochastic parrot' so many times, actual parrots are filing for trademark infringement.

- If tech dystopia were an Olympic sport, you'd be bringing home gold medals while explaining how the podium was designed by committee and the medal contains surveillance chips.

Re: Claude 3.7 Sonnet and Claude Code

#380
post #363

Earlier quoted context omitted.

I'm surprised that Gemini 2.0 is first now. I remember that Google models were under performing on kagi benchmarks.

Gemini 2 is really good, and insanely fast.

It is, but in this benchmark gemini scored very poorly in the past.
Post reply on HN