Hacker News new | ask | show | jobs
by julianlam 43 days ago
Of course.

Qwen 3.6 35B-A3B on a Framework 13 with 32GB of memory.

Running llama.cpp, 15 tokens per second. Outputs code and text faster than I can parse.