Google is using them in prod. I think they're so hungry for chips internally that cloud isn't getting much support in selling them.
I would guess that Google's vertexAI managed solution uses TPUs. Also Google uses them internally to train and infer for all their research products.
Who uses Google TPUs for inference in production?
11–20 of 50 posts
Re: Who uses Google TPUs for inference in production?
#12Re: Who uses Google TPUs for inference in production?
#13Google is using them in prod. I think they're so hungry for chips internally that cloud isn't getting much support in selling them.
I would guess that Google's vertexAI managed solution uses TPUs. Also Google uses them internally to train and infer for all their research products.
Re: Who uses Google TPUs for inference in production?
#14[flagged]
Re: Who uses Google TPUs for inference in production?
#15Re: Who uses Google TPUs for inference in production?
#16[flagged]
Re: Who uses Google TPUs for inference in production?
#17Google is using them in prod. I think they're so hungry for chips internally that cloud isn't getting much support in selling them.
https://www.prnewswire.com/news-releases/google-announces-ex...
"Partnership includes important new collaborations on AI safety standards, committing to the highest standards of AI security, and use of TPU v5e accelerators for AI inference "
Re: Who uses Google TPUs for inference in production?
#18We've previously tried and almost always regretted the decision. I think the tech stack needs another 12-18 months to mature (doesn't help that almost all work ex Google is being done in torch).
Google has been doing AI before any other company even thought about it. They are on the 6th generation of TPU hardware.
I don't think there is any maturity issue, just an availability issue because they are all being used internally.
Re: Who uses Google TPUs for inference in production?
#19We've previously tried and almost always regretted the decision. I think the tech stack needs another 12-18 months to mature (doesn't help that almost all work ex Google is being done in torch).
> I think the tech stack needs another 12-18 months to mature Google has been doing AI before any other company even thought about it. They are on the 6th generation of TPU hardware. I don't think there is any maturity issue, just an availability issue because they are all being used internally.
If you aren't internal, the documentation, support, and even just general bug fixing is impossible to get.
Re: Who uses Google TPUs for inference in production?
#20https://pytorch.org/blog/high-performance-llama-2/
> The PyTorch/XLA Team at Google
Meanwhile you have an issue from 5 years ago with 0 support