Hacker News new | ask | show | jobs
by mfro 28 days ago
I've done some basic testing of the CoreAI framework (using Apple's official 'llm-runner' and officially supported .coreai converted models) and seen no noticable performance increase between standard MLX or GGUF with llama.cpp. I'd love to see some thorough benchmarks from someone though.
1 comments

The idea is that it uses a lot less power than the GPU.