Hacker News new | ask | show | jobs
by bigyabai 20 days ago
It's really not that hard. If Apple Silicon was competitive for inference, then Apple would not be shelving them: https://9to5mac.com/2026/03/02/some-apple-ai-servers-are-rep...

They're not usable for deployment. They're perfectly fine for "enthusiast" low-end usage with 10-30B models, but the same goes for almost every dGPU made in the last 10 years. Your Mac Studio cannot run frontier LLMs at an interactive speed, even Apple has given up on using it as an inference backbone.