|
|
|
|
|
by basch
41 days ago
|
|
Did you look at the charts in the article? It out-performed every model that wasnt a max/ultrafrontier of some sort, except for the one that the article was extolling the virtues of, including grok high. you could make a good argument that deepseek is a better value, but gemini flash is when bundled is already pretty accessible. nowhere did i claim that flash was better than fable or 5.5xhigh. |
|
I don't care about someone else's charts, i care about my own lived experiences. Benchmarks can be gamed to get to the top of charts. When I pay for a service I care about how it performs in my test cases, not about which tops some random charts.
Read my comment again please. I think I was pretty clear with detailed examples on where Gemini sucks and where it's good at.
>nowhere did i claim that flash was better than fable or 5.5xhigh.
And nowhere did I claim that. I said even basic GPT and Grok are better than Gemini Flash at reasoning tasks. Again, read my comment again, I have already explained why with examples.