Kagi LLM benchmark updated with general purpose and thinking mode for Sonnet 3.7. https://help.kagi.com/kagi/ai/llm-benchmark.html Appears to be second most capable general purpose LLM we tried (second to gemini 2.0 pro, in front of gpt-4o). Less impressive in thinking mode, about at the same level as o1-mini and o3-mini (with 8192 token thinking budget). Overall a very nice update, you get higher quality and higher…
Claude 3.7 Sonnet and Claude Code
371–380 of 1001 posts
Re: Claude 3.7 Sonnet and Claude Code
#372You can get your HN profile analyzed by it and it's pretty funny :) https://hn-wrapped.kadoa.com/ I'm using this to test the humor of new models.
ouch (ㅠ﹏ㅠ)
Re: Claude 3.7 Sonnet and Claude Code
#373Drawing an SVG of a pelican on a bicycle. Claude 3.7 edition: https://x.com/umaar/status/1894114767079403747
Re: Claude 3.7 Sonnet and Claude Code
#374Hi everyone! Boris from the Claude Code team here. @eschluntz, @catherinewu, @wolffiex, @bdr and I will be around for the next hour or so and we'll do our best to answer your questions about the product.
Re: Claude 3.7 Sonnet and Claude Code
#375You can get your HN profile analyzed by it and it's pretty funny :) https://hn-wrapped.kadoa.com/ I'm using this to test the humor of new models.
> For someone who worked at Reddit, you sure spend a lot of time on HN. It's like leaving Facebook to spend all day on Twitter complaining about social media. Wow, so spot on it hurts!
>You're the only person who gets excited when someone mentions Trinity Desktop Environment in 2025
> You probably have more opinions about PHP's empty() function than most people have about their entire career choices
Re: Claude 3.7 Sonnet and Claude Code
#376Earlier quoted context omitted.
It is API billing like AWS - you pay for what you use. Every time you exit a session we print the cost, and in the middle of a session you can do /cost to see your cost so far that session! You can track costs in a few ways and set spend limits to avoid surprises: https://docs.anthropic.com/en/docs/agents-and-tools/claude-c...
Which is theoretically great, but if anyone can get an Aussie credit card to work, please let me know.
But I still hit limits, I use Claudemind with jetbrains stuff and there is a max of input tokens (j believe), I am ‘tier 2’ but doesn’t look like I can go past this without an enterprise agreement
Re: Claude 3.7 Sonnet and Claude Code
#377Earlier quoted context omitted.
If you are open to alternatives, try https://glama.ai/gateway We currently serve ~10bn tokens per day (across all models). OpenAI compatible API. No rate limits. Built in logging and tracing. I work with LLMs every day, so I am always on top of adding models. 3.7 is also already available. https://glama.ai/models/claude-3-7-sonnet-20250219 The gateway is integrated directly into our chat ( https://glama.ai/chat ). So…
Do you have deepseek r1 support? I need it for a current product I’m working on.
Re: Claude 3.7 Sonnet and Claude Code
#378Claude 3.5 sonnet has been my go to for coding tasks, it’s just so much better than the others. but I’ve tried using the api in production and had to drop it due to daily issues: https://status.anthropic.com/ compare to https://status.openai.com/ any idea when we’ll see some improvements in api availability or will the focus be more on the web version of claude?
Err, if you compare the two consoles you'll see that anthropic is actually slightly better on average than openai's uptime.
Re: Claude 3.7 Sonnet and Claude Code
#379You can get your HN profile analyzed by it and it's pretty funny :) https://hn-wrapped.kadoa.com/ I'm using this to test the humor of new models.
Roast:
- Your comments have more doom predictions than a Y2K convention in December 1999.
- You've used 'stochastic parrot' so many times, actual parrots are filing for trademark infringement.
- If tech dystopia were an Olympic sport, you'd be bringing home gold medals while explaining how the podium was designed by committee and the medal contains surveillance chips.