> All of this is to say that DeepSeek-V3 is not a unique breakthrough or something that fundamentally changes the economics of LLM’s; it’s an expected point on an ongoing cost reduction curve. What’s different this time is that the company that was first to demonstrate the expected cost reductions was Chinese. Says the CEO whose product [1] costs 15-50x times more. (This is not just the DeepSeek's API, but also 3p pr…
Where does he imply that it's 2x larger than Sonnet?
> Since DeepSeek-V3 is worse than those US frontier models — let’s say by ~2x on the scaling curve, which I think is quite generous to DeepSeek-V3
He says that it's 2x worse. So if it has the ~same quality [1], it would imply it's 2x larger. Unless I misunderstood what he meant there, of course.
[1]: "DeepSeek produced a model close to the performance of US models" -- in his own words.