Google seems to have anorexia when it comes to model intelligence. They have an internal hard constraint on price per token it seems, and they are trying to squeeze out intelligence with limited compute. I wonder if there is something with their TPU cycles that makes them want to postpone training a new model. My guess is that they have been on the same base model for 6 months and they may have waited for the next ge…
I do think it's still also simultaneously true that they have an actual problem with competing with current frontier progress. It's just that has gone from an existential threat to something they are willing to defer addressing because they see the long game for them sitting at the smaller end.