I have used their gemma 4 31b model through kagi and getting real instantaneous answers is absolutely crazy. A very different feeling and UX. Even if the model is smaller, there is definitely a use case for these. I was wondering if they would put the qwen 27b model, it sounds very interesting to try.
May I ask what you used Gemma 31B for? Last time I try it wasn't bad but then it wasn't particularly good either.
It is great UX when you are in a search results page, but I don't use it in the assistant directly because usually this kind of speed is less relevant there.