Remix clone Hacker News

new | show | ask | jobs Github

	▲	vintermann 15 hours ago
		Interesting, but couldn't a model "cheat" in this task by being very good at telling model outputs apart? How far do you get with a classifier simply trained to distinguish models by their output? It seems to me many models - maybe by design - have a recognizable style which would be much easier to detect than evaluating the factual quality of answers.
	▲	nestorD 9 hours ago \| parent [-]
		In theory, yes! If this metric ever becomes a widely used standard, one would have to start accounting for that... But, in practice, when asking a model to pick the best answer they see a single question / answers pair and focus on determining what they think is best.