Remix clone Hacker News

new | show | ask | jobs Github

	▲	dhampi an hour ago
		Well, with thinking models, it’s not that simple. The probability distribution is next token. But if a model thinks to produce an answer, you can have a high confidence next token even if MCMC sampling the model’s thinking chain would reveal that the real probability distribution had low confidence.