|
|
|
|
|
by spwa4
11 days ago
|
|
... and it was not SOTA at the time of release. Gemini 3.1 Pro was previewed on 19 February 2026. It just barely beat GPT-5.3 and was roughly equal with Sonnet 4.6. Better according to some, worse according to others. And that lasted about 10 days (until GPT-5.4 came out which was also not a huge jump). And these are large averages. On coding or terminal Gemini 3.1 Pro was not close to Opus 4.6. Also it only matches the GLM 5 open model, more or less. Last time Google had a "everybody agrees" SOTA model was Gemini 3 in November 2025, it beat GPT-5.1, it really was better and held it for a month. Currently Google's best model doesn't match open models, in any category (performance, price or speed), in fact the last 4 open weight champions all beat Google's best model. Here's a plot of the evolution: https://www.reddit.com/r/LocalLLaMA/comments/1v20g29/kimik3_... |
|
I have no opinion on whether this is true, but "It's been 6 months since Google had the best model in the world" seems a rather weak criticism!
It's anyways been clear for a long time that people are finding value at all sorts of different model sizes and price points, and that Pareto frontier and cost-to-complete task are more important than who benchamaxxed who.
If a model as strong as Gemini 3.6 Flash(!) had been released a year ago, then everyone would be falling over themselves calling it AGI - it is extremely capable, and free usage in the chat app is essentially unlimited.