Hacker News new | ask | show | jobs
by ttoinou 2 days ago
We could make LLM inference 100x cheaper to run at home efficiently, but that solution might need to be updated every 1-2 years, whereas current GPU are useful for various others tasks and last longer
1 comments

> We could make LLM inference 100x cheaper to run at home efficiently

do you genuinely think that's going to happen?