|
|
|
|
|
by llmslave
11 days ago
|
|
I keep saying this and people dont believe me, but I have b2b saas systems with actual agents running around the clock, and the performance/stability of the flash model is higher than most other models. Meaning, its predictable with tool calls, wont spin off a million tools/do weird behavior, its reasonable. Even sonnet in a real world decision making scenario is not reliable, or will reason so long its incredibly expensive. The benchmarks arent catching all the value, and most people have never actually ran an ai agent in a real context that matters |
|