Most important, even ignoring latency, is throughput (tokens) per $$$. And according to their own benchmark [1] (famous last words :)), they're quite cost efficient.
[1] https://www.semianalysis.com/p/groq-inference-tokenomics-spe...