| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by anorwell 357 days ago
	I think your example reflects well on oss-20b, not poorly. It (may) show that they've been successful in separating reasoning from knowledge. You don't _want_ your small reasoning model to waste weights memorizing minutiae.