Hacker News new | ask | show | jobs
by dankwizard 404 days ago
I've given up trying to locally use LLMs on AMD
1 comments

Basically anything llama.cpp (Vulkan backend) should work out of the box w/o much fuss (LM Studio, Ollama, etc).

The HIP backend can have a big prefill speed boost on some architectures (high-end RDNA3 for example). For everything else, I keep notes here: https://llm-tracker.info/howto/AMD-GPUs