Live data from Hacker News

Claude 3.7 Sonnet and Claude Code

anthropic.com

601–610 of 1001 posts

Re: Claude 3.7 Sonnet and Claude Code

#601

Earlier quoted context omitted.

The biggest complaint I (and several others) have is that we continuously hit the limit via the UI after even just a few intensive queries. Of course, we can use the console API, but then we lose ability to have things like Projects, etc. Do you foresee these limitations increasing anytime soon? Quick Edit: Just wanted to also say thank you for all your hard work, Claude has been phenomenal.

If you are open to alternatives, try https://glama.ai/gateway We currently serve ~10bn tokens per day (across all models). OpenAI compatible API. No rate limits. Built in logging and tracing. I work with LLMs every day, so I am always on top of adding models. 3.7 is also already available. https://glama.ai/models/claude-3-7-sonnet-20250219 The gateway is integrated directly into our chat ( https://glama.ai/chat ). So…

Just tried it, is there a reason why the webUI is so slow?

Try to delete (close) the panel on the right on a side-by-side view. It took a good second to actually close. Creating one isn't much faster.

This is unbearably slow, to be blurt.

Re: Claude 3.7 Sonnet and Claude Code

#602

Earlier quoted context omitted.

We find that Claude is really good at test driven development, so we often ask Claude to write tests first and then ask Claude to iterate against the tests

Write tests (plural) first, as in write more than one failing test before making it pass?

Time to look up TDD, my friend.

Re: Claude 3.7 Sonnet and Claude Code

#603
post #91

Hi everyone! Boris from the Claude Code team here. @eschluntz, @catherinewu, @wolffiex, @bdr and I will be around for the next hour or so and we'll do our best to answer your questions about the product.

with Claude coder, how does history work? I used it with my account, ran out of credit then switched to a work account but there was no chat history or other saved context of the work that had been done. I logged back in with my account to try copy it but it was gone.

Re: Claude 3.7 Sonnet and Claude Code

#604

Claude 3.7 Sonnet scored 60.4% on the aider polyglot leaderboard [0], WITHOUT USING THINKING. Tied for 3rd place with o3-mini-high. Sonnet 3.7 has the highest non-thinking score, taking that title from Sonnet 3.5. Aider 0.75.0 is out with support for 3.7 Sonnet [1]. Thinking support and thinking benchmark results coming soon. [0] https://aider.chat/docs/leaderboards/ [1] https://aider.chat/HISTORY.html#aider-v0750

Using up to 32k thinking tokens, Sonnet 3.7 set SOTA with a 64.9% score. 65% Sonnet 3.7, 32k thinking 64% R1+Sonnet 3.5 62% o1 high 60% Sonnet 3.7, no thinking 60% o3-mini high 57% R1 52% Sonnet 3.5

Also for $36.83 compared to o1's $186.50

Re: Claude 3.7 Sonnet and Claude Code

#605

Earlier quoted context omitted.

As a growth company, they likely would prefer a larger amount of users even with occasional rate limits, vs smaller pool of power users. As long as capacity is an issue, you can't have both

If people are paying for use, then why can’t you have both?

It takes time to grow capacity to meet growing revenue/usage. As parent is saying, if you are in a growth market at time T with capacity X, you would rather have more people using it even if that means they can each use less.

Re: Claude 3.7 Sonnet and Claude Code

#606
post #200

Earlier quoted context omitted.

Why gatekeep Claude Code, instead of releasing the code for it? It seems like a direct increase in revenue/API sales for your company.

I'm not affiliated with Anthropic, but it seems like doing this will commoditize Claude (the AIaaS). Hosted AI providers are doing all they can to move away from being interchangeable commodities; it's not good for Anthropic's revenue for users to be able to easily swap-out the backend of Cloud Code to a local Olama backend, or a cheaper hosted DeepSeek. Open sourcing Claude Code would make this option 1 or 2 forks/P…

It's not hard to make, its a relatively simple CLI tool so there's no moat. Also, the minified source code is available.

Re: Claude 3.7 Sonnet and Claude Code

#607

This AI race is happening so fast. Seems like it to me anyway. As a software developer/engineer I am worried about my job prospects.. time will tell. I am wondering what will happen to the west coast housing bubbles once software engineers lose their high price tags. I guess the next wave of knowledge workers will move in and take their place?

My guess is that, yes, the software development job market is being massively disrupted, but there are things you can do to come out on top: * Learn more of the entire stack, especially the backend, and devops. * Embrace the increased productivity on offer to ship more products, solo projects, etc * Be highly selective as far as possible in how you spend your productive time: being uber-effective can mean thinking an…

I love, especially the last point.

But, what do you use for agentic assistants?

Re: Claude 3.7 Sonnet and Claude Code

#608

Earlier quoted context omitted.

Neither a statement for or against Grok or Anthropic: I've now just taken to seeing benchmarks as pretty lines or bars on a chart that are in no way reflective of actual ability for my use cases. Claude has consistently scored lower on some benchmarks for me, but when I use it in a real-world codebase, it's consistently been the only one that doesn't veer off course or "feel wrong". The others do. I can't quantify it…

O1 pro is excellent at figuring out complex stuff that Claude misses. It’s my go to mid level debug assistant when Claude spins

Ive found the same but find o3-mini just as good as that. Sonnet is far better as a general model, but when it's an open-ended technical question that isn't just about code, o3-mini figures it out while Sonnet sometimes doesn't. In those cases o3 is less inclined to go with purely the most "obvious" answer when it's wrong.

Re: Claude 3.7 Sonnet and Claude Code

#609

Earlier quoted context omitted.

I really want to try your AI models, but "You must have a valid phone number to use Anthropic's services." is a show-stopper for me. It's the only mainstream AI service that requests this information. After a string of security lapses by many of your competitors, I have zero faith in the ability of a "fast moving" AI-focused company to keep my PII data secure.

It's a phone number. It's probably been bought / sold a few times already. Unless you're on the level of Edward Snowden, I wouldn't worry about it. But maybe your sense of privacy is more valuable than the outcome you'd get from Claude. That's fine too.

It's my phone number... linked to my Google identity... linked to every submitted user prompt... linked to my source code.

There's also been a spate of AI companies rushing to release products and having "oops" moments where they leaked customer chats or whatever.

They're not run like a FAANG, they don't have the same security pedigree, and they generally don't have any real guarantee of privacy.

So yes, my privacy is more valuable.

Conversely: Why is my non-privacy so valuable to Anthropic? Do they plan on selling my data? Maybe not now... but when funding gets a bit tight? Do they plan on selling my information to the likes of Cambridge Analytica? Not just superficial metadata, but also an AI-summarised history of my questions?

The best thing to do would be not to ask. But they are asking.

Why?

Why only them?

Re: Claude 3.7 Sonnet and Claude Code

#610

When you ask: 'How many r's are in strawberry?' Claude 3.7 Sonnet generates a response in a fun and cool way with React code and a preview in Artifacts check out some examples: [1] https://claude.ai/share/d565f5a8-136b-41a4-b365-bfb4f4400df5 [2] https://claude.ai/share/a817ac87-c98b-4ab0-8160-feefd7f798e8

This test has always been so stupid since models work at the token level. Claude 3.5 already 5xs your frontend dev speed but people still say "hurr durr it can't count strawberry" as if that's a useful problem

“Already 5xs”

Even AI marketing doesn’t claim this. Totally baseless claim given how many people report negative experiences trying to use AI.

Post reply on HN