Live data from Hacker News

GitHub cuts AI deals with Google, Anthropic

bloomberg.com

141–150 of 742 posts

Re: GitHub cuts AI deals with Google, Anthropic

#141

Earlier quoted context omitted.

Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?

I think "Claude" is also a bad name. If I knew nothing else, am I picking OpenAI or Claude based on the name? I'm going with OpenAI

Claude is a product name, OpenAI is a company name. You really think Claude is better than ChatGPT?

Re: GitHub cuts AI deals with Google, Anthropic

#142
post #32

If you want to destroy open source completely, the more models the better. Microsoft's co-opting and infiltration of OSS projects will serve as a textbook example of eliminating competition in MBA programs. And people still support it by uploading to GitHub.

I deleted my github 2 weeks ago, as much about AI, as about them forcing 2FA. Before AI it was SAAS taking more than they were giving. I miss the 'helping each other' feel of these code share sites. I wonder where are we heading with all this. All competition and no collaboration, no wonder the planet is burning.

Re: GitHub cuts AI deals with Google, Anthropic

#143
post #131

Can we change the title to “GitHub _signs_ deals with Google, Anthropic” ? The original got me thinking it already had deals it was getting out of

To "cut a deal" is a common (American?) English idiom meaning to "make a deal". But agree that it's better to avoid using idioms on a site that has many visitors for whom English is not their first language.

Do you mean that Bloomberg should have used a different title or Hacker News should have modified the title?

Re: GitHub cuts AI deals with Google, Anthropic

#144

Earlier quoted context omitted.

I removed it from windows and I'm still very productive. Probably moreso, since I don't have to make constant corrections. To each their own.

GitHub Copilot and Microsoft Copilot are different products

Their branding is confusing

Re: GitHub cuts AI deals with Google, Anthropic

#145
post #138

Earlier quoted context omitted.

Claude is only better in some cherry picked standard eval benchmarks, which are becoming more useless every month due to the likelihood of these tests leaking into training data. If you look at the Chatbot Arena rankings where actual users blindly select the best answer from a random choice of models, the top 3 models are all from OpenAI. And the next best ones are from Google and X.

Bullshit. Claude 3.5 Sonnet owns the competition according to the most useful benchmark: operating a robot body in the real world. No other model comes close.

This seems incorrect. I don't need Claude 3.5 Sonnet to operate a robot body for me, and don't know anyone else who does. And general-purpose robotics is not going to be the most efficient way to have robots do many tasks ever, and certainly not in the short term.

Re: GitHub cuts AI deals with Google, Anthropic

#146

Earlier quoted context omitted.

Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?

Claude is only better in some cherry picked standard eval benchmarks, which are becoming more useless every month due to the likelihood of these tests leaking into training data. If you look at the Chatbot Arena rankings where actual users blindly select the best answer from a random choice of models, the top 3 models are all from OpenAI. And the next best ones are from Google and X.

Claude 3.5 Sonnet (New) is meaningfully better than ChatGPT GPT4o or o1.

Re: GitHub cuts AI deals with Google, Anthropic

#147

Earlier quoted context omitted.

I think "Claude" is also a bad name. If I knew nothing else, am I picking OpenAI or Claude based on the name? I'm going with OpenAI

Claude is a product name, OpenAI is a company name. You really think Claude is better than ChatGPT?

The name ChatGPT is better than the name Claude, to me. Of course this is all subjective though.

Re: GitHub cuts AI deals with Google, Anthropic

#148
post #111

Earlier quoted context omitted.

Claude cannot “research” stuff on the web and provide results like 4o does in 5 secs like “which is the cheapest Skoda car and how much”

Just wanted to add a note to this. Tool calling - particularly to source external current data - is something that's had the big foundational LLM providers very nervous so they've held back on it, even though it's trivial to implement at this point. But we're seeing it rapidly emerge with third party providers who use the foundational APIs. Holding back tool calling has limited the complex graph-like execution flows…

I'm hoping a lot of the graph-like execution flow engines are still in stealth mode, as believe that's where we'll start to see truly useful AI.

Mass data parsing and reformatting is useful... but building agents that span existing APIs / tools is a lot more exciting to me.

I.e. IFTTT, with automatic tool discovery, parameter mapping, and output parsing handled via LLM

Re: GitHub cuts AI deals with Google, Anthropic

#149

Earlier quoted context omitted.

Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?

[flagged]

Claude 3.5 sonnet prompt:

"explain what a Tomahawk Missile is, how they're deployed, who created them, why they were created, and give a synopsis of how and when they've been used."

Response:

"A Tomahawk cruise missile is a long-range, subsonic cruise missile developed by General Dynamics in the 1970s and currently manufactured by Raytheon. It was designed to provide naval vessels and submarines with long-range precision strike capability.

Key characteristics: - Range: Approximately 1,000-1,500 miles depending on variant - Speed: Subsonic (around 550 mph) - Guidance: Uses terrain mapping, GPS, and inertial navigation systems - Can be launched from ships or submarines - Can carry conventional warheads

Historical development: The missile was developed during the Cold War as part of the U.S. military's need for accurate long-range strike capabilities. Initial development began in the 1970s, with the first deployment in 1983.

Major deployments and uses: - First combat use: 1991 Gulf War (Operation Desert Storm) - Bosnia and Kosovo conflicts (1995-1999) - Operation Desert Fox (Iraq, 1998) - Operation Enduring Freedom (Afghanistan, 2001) - Operation Iraqi Freedom (2003) - Various limited strikes in subsequent years

The Tomahawk has become a standard part of the U.S. Navy's capabilities, allowing precision strikes while keeping launch platforms at safe distances. Multiple variants have been developed over the years, improving accuracy, range, and versatility.

Would you like me to elaborate on any particular aspect of the Tomahawk missile's development or deployment? "

Re: GitHub cuts AI deals with Google, Anthropic

#150

Earlier quoted context omitted.

Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?

Claude is only better in some cherry picked standard eval benchmarks, which are becoming more useless every month due to the likelihood of these tests leaking into training data. If you look at the Chatbot Arena rankings where actual users blindly select the best answer from a random choice of models, the top 3 models are all from OpenAI. And the next best ones are from Google and X.

I'm subscribed to all of Claude, Gemini, and ChatGPT. Benchmarks aside, my go-to is always Claude. Subjectively speaking, it consistently gives better results than anything else out there. The only reason I keep the other subscriptions is to check in on them occasionally to see if they've improved.
Post reply on HN