|
|
|
|
|
by andy99
13 days ago
|
|
Qwen 3.5 to 3.6 was a big jump for the same size, e.g. 29 to 32 on artificial analysis intelligence for the 35BA3B models. Although I don’t think anyone has released a better model of that size since. I would love to see something like a 90B A6B model that is optimized for 128GB machines e.g. strix halo, I haven’t seen anything really targeting the combination of RAM and compute these machines have, but I’m biased because I have one. |
|
Qwen 3.6 27b 8b quant 16b kv cache is already pretty good on the Strix.