| HN Mirror

Y	Hacker News new \| ask \| show \| jobs

by philipkglass 78 days ago

Do you have plans to do a follow-up model release with quantization aware training as was done for Gemma 3?

Having 4 bit QAT versions of the larger models would be great for people who only have 16 or 24 GB of VRAM.