|
|
|
|
|
by Fabricio20
36 days ago
|
|
One thing I see noone asking, is this not a case of optimization? Hidden reasoning means they dont need to process the output of all that, it stays internal within the model. Less cost for them -> less cost for us (even if they benefit mroe), compared to streaming all of those reasoning tokens out? |
|
[1] https://blog.cryptographyengineering.com/2026/05/29/fooling-...
Edit: other comments under this post seem to indicate that thinking tokens are cached on the server side as well? I'm a bit confused.