Earlier quoted context omitted.
Samba is on gen 4 silicon and still lagging, somebody over there is doing something wrong
How are they lagging? They are running faster than anyone else at full precision and with many many fewer chips than Groq. Groq is not real.
Economics and costs are hard to predict. For example, Groq is not using HBM chips. So probably the cards are a lot easier to source.
Its not clear what the capacity of these systems are in terms of total users, or even tokens per second. Then you factor in cost. Then you realize all vendors will match a competitors pricing. Then you realize Groq doesn't sell chips.
¯\_(ツ)_/¯
The only thing you have is the public API to benchmark against: https://artificialanalysis.ai/