Live data from Hacker News

Claude 3.7 Sonnet and Claude Code

anthropic.com

61–70 of 1001 posts

Re: Claude 3.7 Sonnet and Claude Code

#61

Just like OpenAI or Grok, there is no transparency and no way for self-hosting purposes. Your input and confidential information can be collected for training purposes. I just don't trust those companies when you use their servers. This is not a good approach to LLM democratization.

I wouldn’t assume there’s no way to self host — it just costs a lot more than open weights.

Anthropic claims they don’t train on their inputs. I haven’t seen any reason to disbelieve them.

Re: Claude 3.7 Sonnet and Claude Code

#62
post #10

I've been using O3-mini with reasoning effort set to high in Aider and loving the pricing. This looks as though it'll be about three times as expensive. Curious to see which falls out as most useful for what over the next month!

Aro using o3-mini for editing or just architect in architect-editor mode?

Re: Claude 3.7 Sonnet and Claude Code

#64
post #31
post #26

[flagged]

> Why would my phone number be any of their business? Preventing abuse? It's much harder to create a throwaway phone number than a throwaway email address. > OpenAI does the logical thing. Let's me enter my credit card and I'm good to go. I will stay with them. You'd rather hand over your credit card than your phone number? I think most people would see it the other way around.

Many credit card companies make it easy to generate one-off card numbers/“virtual cards” you can use to subscribe to services that are hard to cancel or otherwise questionable (so you can cancel just the card you used for that company).

Re: Claude 3.7 Sonnet and Claude Code

#65

Just like OpenAI or Grok, there is no transparency and no way for self-hosting purposes. Your input and confidential information can be collected for training purposes. I just don't trust those companies when you use their servers. This is not a good approach to LLM democratization.

[deleted]

Re: Claude 3.7 Sonnet and Claude Code

#67
To me the biggest surprise was seeking grok dominate in all of their published benchmarks. I haven’t seen any benchmarks of it yet (which I take with a giant heap of salt), but it’s still interesting nevertheless.

I’m rooting for Anthropic.

Re: Claude 3.7 Sonnet and Claude Code

#68

> Third, in developing our reasoning models, we’ve optimized somewhat less for math and computer science competition problems, and instead shifted focus towards real-world tasks that better reflect how businesses actually use LLMs. Company: we find that optimizing for LeetCode level programming is not a good use of resources, and we should be training AI less on competition problems. Also Company: we hire SWEs based…

My manager explained to me that LeetCode is proving that you are willing to dance the dance. Same as PhD requirements etc - you probably won't be doing anything related and definitely nothing related to LeetCode, but you display dedication and ability.

I kinda agree that this is probably reason why companies are doing it. I don't like it, but this is besides the matter.

Using Claude other models in interviews probably won't be allowed any time soon, but I do use it the work. So it does make sense.

Re: Claude 3.7 Sonnet and Claude Code

#69
post #35
post #31

Earlier quoted context omitted.

> Why would my phone number be any of their business? Preventing abuse? It's much harder to create a throwaway phone number than a throwaway email address. > OpenAI does the logical thing. Let's me enter my credit card and I'm good to go. I will stay with them. You'd rather hand over your credit card than your phone number? I think most people would see it the other way around.

Credit card is easily changed. Phone number is much more difficult.

You credit card is also easily charged.

Your phone number isn't.

What is a company going to do with your phone number that you're worried about...?

Re: Claude 3.7 Sonnet and Claude Code

#70

> Third, in developing our reasoning models, we’ve optimized somewhat less for math and computer science competition problems, and instead shifted focus towards real-world tasks that better reflect how businesses actually use LLMs. Company: we find that optimizing for LeetCode level programming is not a good use of resources, and we should be training AI less on competition problems. Also Company: we hire SWEs based…

And it's also the reality of hiring practices for most VC-backed and public companies

Some try to do something more like "real-world" tasks, but those end up either being either just toy problems, or long take homes

Personally, I feel the most important things to prioritize when hiring are: is the candidate going to get along with their teammates (colleagues, boss, etc), and do they have the basic skills to relatively quickly learn their jobs once they start?

Post reply on HN