| ▲ | zwaps 4 hours ago | |
No mention of calibration. Is it just another llm finetune? | ||
| ▲ | kflansburg 4 hours ago | parent [-] | |
> Our post-training utilizes label-smoothed cross-entropy for valid schema outputs paired with a Brier loss to refine probability calibration. | ||