Hacker News new | ask | show | jobs
by kamranjon 10 hours ago
I haven't spun up thinkingcap yet but I'm aware of it and am intending to try it out soon. How did you find it?
1 comments

Not tested much but it is not noticeably worse than the underlying Qwen 3.6 27B in Q4_K_M, which in my experience is kind of a first for a Qwen fine-tune of this nature. They are almost always worse.

I think it does use fewer tokens while reasoning, which is potentially useful. I need to do more testing, because any performance advantage over the 27B is useful for me on an M1 Max.