Hacker News new | ask | show | jobs
by denn-gubsky 12 hours ago
I'm running qwen3.6, gpt-oss and embeddinggemma with Ollama on Ryzen 7 8700G + 96 GB DDR5 with 12-14 tokens/sec. Consider this as a floor. On Strix Halo it will run 4-6 times faster depending on memory bandwidth. For my local task 14 tok/s is quite enough for the price I paid.
1 comments

So no GPU? Just a strong processor? I am considering building a 'in my house' system. Timing of this post is very fortuitous.