| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by dhampi 34 days ago
	Well, with thinking models, it’s not that simple. The probability distribution is next token. But if a model thinks to produce an answer, you can have a high confidence next token even if MCMC sampling the model’s thinking chain would reveal that the real probability distribution had low confidence.