Hacker News new | ask | show | jobs
Sinkhorn: Make LLMs even smaller through quantisation while maintaining accuracy (github.com)
4 points by ilitirit 299 days ago
1 comments