logoalt Hacker News

zwaps • today at 5:08 PM • 1 reply • view on HN

No mention of calibration. Is it just another llm finetune?


Replies

kflansburg • today at 5:11 PM

> Our post-training utilizes label-smoothed cross-entropy for valid schema outputs paired with a Brier loss to refine probability calibration.