Earlier quoted context omitted.
Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?
I think "Claude" is also a bad name. If I knew nothing else, am I picking OpenAI or Claude based on the name? I'm going with OpenAI
GitHub cuts AI deals with Google, Anthropic
141–150 of 742 posts
Re: GitHub cuts AI deals with Google, Anthropic
#142If you want to destroy open source completely, the more models the better. Microsoft's co-opting and infiltration of OSS projects will serve as a textbook example of eliminating competition in MBA programs. And people still support it by uploading to GitHub.
Re: GitHub cuts AI deals with Google, Anthropic
#143Can we change the title to “GitHub _signs_ deals with Google, Anthropic” ? The original got me thinking it already had deals it was getting out of
To "cut a deal" is a common (American?) English idiom meaning to "make a deal". But agree that it's better to avoid using idioms on a site that has many visitors for whom English is not their first language.
Re: GitHub cuts AI deals with Google, Anthropic
#144Re: GitHub cuts AI deals with Google, Anthropic
#145Earlier quoted context omitted.
Claude is only better in some cherry picked standard eval benchmarks, which are becoming more useless every month due to the likelihood of these tests leaking into training data. If you look at the Chatbot Arena rankings where actual users blindly select the best answer from a random choice of models, the top 3 models are all from OpenAI. And the next best ones are from Google and X.
Bullshit. Claude 3.5 Sonnet owns the competition according to the most useful benchmark: operating a robot body in the real world. No other model comes close.
Re: GitHub cuts AI deals with Google, Anthropic
#146Earlier quoted context omitted.
Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?
Claude is only better in some cherry picked standard eval benchmarks, which are becoming more useless every month due to the likelihood of these tests leaking into training data. If you look at the Chatbot Arena rankings where actual users blindly select the best answer from a random choice of models, the top 3 models are all from OpenAI. And the next best ones are from Google and X.
Re: GitHub cuts AI deals with Google, Anthropic
#147Earlier quoted context omitted.
I think "Claude" is also a bad name. If I knew nothing else, am I picking OpenAI or Claude based on the name? I'm going with OpenAI
Claude is a product name, OpenAI is a company name. You really think Claude is better than ChatGPT?
Re: GitHub cuts AI deals with Google, Anthropic
#148Earlier quoted context omitted.
Claude cannot “research” stuff on the web and provide results like 4o does in 5 secs like “which is the cheapest Skoda car and how much”
Just wanted to add a note to this. Tool calling - particularly to source external current data - is something that's had the big foundational LLM providers very nervous so they've held back on it, even though it's trivial to implement at this point. But we're seeing it rapidly emerge with third party providers who use the foundational APIs. Holding back tool calling has limited the complex graph-like execution flows…
Mass data parsing and reformatting is useful... but building agents that span existing APIs / tools is a lot more exciting to me.
I.e. IFTTT, with automatic tool discovery, parameter mapping, and output parsing handled via LLM
Re: GitHub cuts AI deals with Google, Anthropic
#149Earlier quoted context omitted.
Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?
[flagged]
"explain what a Tomahawk Missile is, how they're deployed, who created them, why they were created, and give a synopsis of how and when they've been used."
Response:
"A Tomahawk cruise missile is a long-range, subsonic cruise missile developed by General Dynamics in the 1970s and currently manufactured by Raytheon. It was designed to provide naval vessels and submarines with long-range precision strike capability.
Key characteristics: - Range: Approximately 1,000-1,500 miles depending on variant - Speed: Subsonic (around 550 mph) - Guidance: Uses terrain mapping, GPS, and inertial navigation systems - Can be launched from ships or submarines - Can carry conventional warheads
Historical development: The missile was developed during the Cold War as part of the U.S. military's need for accurate long-range strike capabilities. Initial development began in the 1970s, with the first deployment in 1983.
Major deployments and uses: - First combat use: 1991 Gulf War (Operation Desert Storm) - Bosnia and Kosovo conflicts (1995-1999) - Operation Desert Fox (Iraq, 1998) - Operation Enduring Freedom (Afghanistan, 2001) - Operation Iraqi Freedom (2003) - Various limited strikes in subsequent years
The Tomahawk has become a standard part of the U.S. Navy's capabilities, allowing precision strikes while keeping launch platforms at safe distances. Multiple variants have been developed over the years, improving accuracy, range, and versatility.
Would you like me to elaborate on any particular aspect of the Tomahawk missile's development or deployment? "
Re: GitHub cuts AI deals with Google, Anthropic
#150Earlier quoted context omitted.
Claude is better than OpenAI for most tasks, and yet OpenAI has enormously more users. What is this, if not first mover advantage?
Claude is only better in some cherry picked standard eval benchmarks, which are becoming more useless every month due to the likelihood of these tests leaking into training data. If you look at the Chatbot Arena rankings where actual users blindly select the best answer from a random choice of models, the top 3 models are all from OpenAI. And the next best ones are from Google and X.