Remix.run Logo
▲ zwaps 4 hours ago

No mention of calibration. Is it just another llm finetune?

▲kflansburg 4 hours ago | parent [-]

> Our post-training utilizes label-smoothed cross-entropy for valid schema outputs paired with a Brier loss to refine probability calibration.