Y
Hacker News
new
|
ask
|
show
|
jobs
by
zozbot234
96 days ago
You can already run inference on ordinary hardware but if you want workable throughput you're limited to small models, and these have very poor world-knowledge.