Y
Hacker News
new
|
ask
|
show
|
jobs
by
3eb7988a1663
5 days ago
Exactly the kind of breakdown I was hoping to see. Thanks.
1 comments
pbgcp2026
5 days ago
And we can quickly see that the real problem is not the models, but HW to run them. You can build whole Enterprise on Gemma 4 31B full precision without significant problems. If you can afford not to lobotomise it by quantisation.
link