|
|
|
|
|
by kamranjon
1 day ago
|
|
There is a really interesting startup in Prague that is doing just that. They fine-tuned Qwen 3.6 27b to have 46% fewer reasoning tokens while maintaining most of the performance characteristics. I'm interested to see if they continue down this path of optimizing reasoning for other models. https://bottlecapai.com/post/thinkingcap-qwen3-6-27b/ |
|