|
|
|
|
|
by XCSme
7 days ago
|
|
In my tests, 3.6 Flash is NOT more token efficient, so it actually ends up costing more than 3.5 Flash, even with the output price reduction. EDIT:
It less less verbose in final output though, but it reasons more. I assume the optimization comes when you have long-running tasks with many tool calls, and by reasoning more, it reduces the number of tool calls needed. |
|