Earlier quoted context omitted.
While true, enshittification should be mourned every time it's renewed until some real kind of block (regulatory changes to the industry) exists. This is like a company buying your local public pool, and you trying to convince everyone that things won't slowly turn to crap.
This is like if the local pool owner wanted to sell the property, and nobody in the neighborhood made a competing bid. HuggingFace isn't a publicly-owned resource. What we should be seeing is competing businesses trying to undercut Nvidia/HF and offer a better experience. China has alternatives to HF; it's just that in the US, Nvidia is the only business that can get off their ass to sponsor infrastructure like that.…
Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
21–30 of 31 posts
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#22I follow llama.cpp pretty closely as I use either llama.cpp itself or projects that depend on it all the time, and one thing that I don't think gets talked about is the sheer scale of community involvement. It seems like a logistical nightmare, but somehow thousands of different contributors are opening hundreds of PR's every week and getting them merged in to support various hardware or implement a new pattern or al…
I know there's a lot of discussion about AI overwhelming OS maintainers and I can see that. It does seem like llama.cpp is successfully riding that dragon right now though.
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#23Earlier quoted context omitted.
This is like if the local pool owner wanted to sell the property, and nobody in the neighborhood made a competing bid. HuggingFace isn't a publicly-owned resource. What we should be seeing is competing businesses trying to undercut Nvidia/HF and offer a better experience. China has alternatives to HF; it's just that in the US, Nvidia is the only business that can get off their ass to sponsor infrastructure like that.…
As a silver lining, it seems like we're hitting a huge compute limit everywhere we go due to TSMC's reluctance to scale. If demand for intelligence continues to stay high, an opportunity opens up to challenge Nvidia and Nvidia run kernels/models.
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#24Why are we not seeing xcancel alternative links? X threw a fit.
Here's the content of the tweet so you don't have to give twitter any more traffic: Hugging Face has been acquired by NVIDIA It is quite exciting to be a part of this journey! NVIDIA has been an active supporter of the llama.cpp project. For more than a year now, their engineers have actively contributed to the codebase, collaborated with the community and provisioned hardware for development and testing purposes. Th…
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#25Earlier quoted context omitted.
As a silver lining, it seems like we're hitting a huge compute limit everywhere we go due to TSMC's reluctance to scale. If demand for intelligence continues to stay high, an opportunity opens up to challenge Nvidia and Nvidia run kernels/models.
if you were a company on an island threatened by a billion-peopled country and your primary support was a senile corrupt democracy, would you feel like you could just roll the dice on bringing more factories into the world?
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#26Translation: Gervanov's ggml.ai was acquired by Huggingface in Feb 2026, so he is now "excited about the journey" after the Huggingface acquisition by Nvidia. Can we take this as an official statement that Nvidia supports local models? Why would Nvidia increase GPU efficiency for local models? Surely they'll operate like athletes and only establish a new record from time to time when necessary.
> Can we take this as an official statement that Nvidia supports local models? Hedging against a data center bubble burst or at least massive pullback Trivial for them to cut margin on desktop cards and clean up/prop up falling data center sales serving the hungry gamer, blockchain, local AI that's been sitting on its hands the last year or two Nvidia is the PC ecosystem's best answer to Apple
Sorry… what does this mean? I thought most people considered Apple irrelevant because of its low market share.
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#27I follow llama.cpp pretty closely as I use either llama.cpp itself or projects that depend on it all the time, and one thing that I don't think gets talked about is the sheer scale of community involvement. It seems like a logistical nightmare, but somehow thousands of different contributors are opening hundreds of PR's every week and getting them merged in to support various hardware or implement a new pattern or al…
1) Reviews are super fast. Sometimes I get the first review 5 minutes after submitting a PR. Today I had a PR merged with two reviews within 42 minutes. This is incredible work by the whole team. It keeps contributors motivated. The pace is intoxicating. On other projects, I've sometimes been waiting months for a review.
2) Georgi motivates people by giving them responsibility. There's no gatekeeping as with other projects. He happily delegates. You do good work? It's appreciated, and you get the freedom you need to make an impact. That feels awesome.
I've rarely become addicted to an open-source project so quickly.
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#28I follow llama.cpp pretty closely as I use either llama.cpp itself or projects that depend on it all the time, and one thing that I don't think gets talked about is the sheer scale of community involvement. It seems like a logistical nightmare, but somehow thousands of different contributors are opening hundreds of PR's every week and getting them merged in to support various hardware or implement a new pattern or al…
I have been using a PR branch to run GLM 5.3 flash locally so I have been watching the dueling PRs develop to implement it and all the comments and reviews. You're right the scale of it is huge. The speed with which maintainers and community are responding is also super impressive. I know there's a lot of discussion about AI overwhelming OS maintainers and I can see that. It does seem like llama.cpp is successfully r…
Of course, llama.cpp also gets lots of AI slop PRs from first-time contributors, but from what I've seen, maintainers simply ignore the bad submissions and prioritize the good contributions. I think that's probably one of the best ways to handle the slop overwhelm.
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#29Earlier quoted context omitted.
> Can we take this as an official statement that Nvidia supports local models? Hedging against a data center bubble burst or at least massive pullback Trivial for them to cut margin on desktop cards and clean up/prop up falling data center sales serving the hungry gamer, blockchain, local AI that's been sitting on its hands the last year or two Nvidia is the PC ecosystem's best answer to Apple
> Nvidia is the PC ecosystem's best answer to Apple Sorry… what does this mean? I thought most people considered Apple irrelevant because of its low market share.
Re: Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace
#30I follow llama.cpp pretty closely as I use either llama.cpp itself or projects that depend on it all the time, and one thing that I don't think gets talked about is the sheer scale of community involvement. It seems like a logistical nightmare, but somehow thousands of different contributors are opening hundreds of PR's every week and getting them merged in to support various hardware or implement a new pattern or al…
Here's why I think it's working so well, as someone who's participating in llama.cpp development and who has participated in many other open source projects: 1) Reviews are super fast. Sometimes I get the first review 5 minutes after submitting a PR. Today I had a PR merged with two reviews within 42 minutes. This is incredible work by the whole team. It keeps contributors motivated. The pace is intoxicating. On othe…