Live data from Hacker News

OpenAI says it has evidence DeepSeek used its model to train competitor

ft.com

451–460 of 1001 posts

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#452

This reminds me of the railroads, where once railroads were invented, there was a huge investment boom of eveyrone trying to make money of the railroads, but the competition brought the costs down where the railroads weren’t the people who generally made the money and got the benefit, but the consumers and regular businesses did and competition caused many to fail. AI is probably similar where the Moore’s law and adv…

I saw a thought-provoking post that similarly compared LLM makers to the airlines: https://calpaterson.com/porter.html

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#453
I have no sympathy for OpenAI here. They are (allegedly) a non-profit with open in the title that refuse to open-source their models.

They are now upset at a startup who is more loyal to OpenAI's original mission that OpenAI is today.

Please, give me a break.

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#454
post #95

> “It’s also extremely hard to rally a big talented research team to charge a new hill in the fog together,” he added. “This is the key to driving progress forward.” Well I think DeepSeek releasing it open source and on an MIT license will rally the big talent. The open sourcing of a new technology has always driven progress in the past. The last paragraph too is where OpenAi seems to be focusing their efforts.. > we…

I'm willing to bet ''ban DeepSeek'' voices will start soon. Why compete, when you can just ban?

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#455

The US government likely will favor a large strategic company like OpenAI instead of individual's copyrights, so while ironic, the US government definitely doesn't care. And the US government is also likely itching to reduce the power of Chinese AI companies that could out compete US rivals (similar to the treatment of BYD, TikTok, solar panel manufacturers, network equipment manufacturers, etc), so expect sweeping l…

It's too late for that. That ship sailed a long time ago.

The best language model right now is open source. Let that sink in.

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#457
post #414

Earlier quoted context omitted.

Of course. How else would Americans justify their superiority (and therefore valuations) if a load of foreigners for Christ's sake could just out innovate them? They had to be cheating.

Please don't take HN threads into nationalistic flamewar. It's not what this site is for, and destroys what it is for. https://news.ycombinator.com/newsguidelines.html p.s. yes, that goes both ways - that is, if people are slamming a different country from an opposite direction, we say the same thing (provided we see the post in the first place)

I see where you’re coming from but that comment didn’t strike me as particularly inflammatory.

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#459
So, is this just an example of the first-mover disadvantage (or maybe the problem of producing public goods?). The first AI models were orders of magnitude more expensive to create, but now that they're here we can, with techniques like distillation, replicate them at a fraction of the cost. I am not really literate in the law but weren't patents invented to solve problems like this?

Re: OpenAI says it has evidence DeepSeek used its model to train competitor

#460

I do think that distilling a model from another is much less impressive than distilling one from raw text. However, it is hard to say if it is really illegal or even immoral, perhaps just one step further in the evolution of the space.

It's about as illegal as the billions, if not trillions of IPs that ClosedAI infringed to train their own data without consent. Not that they're alone, and I personally don't mind that AI companies do it, but it's still amusing when they get this annoyed at others doing the same thing to them.

Is the question of training AI on data fair use settled yet? Because if it is not - it looks like fair use to me.
Post reply on HN