
Sharpening Tax in Post-Training
Post-training sharpens base model behaviors at the cost of solution coverage, introducing a quantifiable "Sharpening Tax"; a posterior-tempered group sampler reduces this tax while boosting accuracy.
Published Oct 1, 2026 · 0 citations · ▲ 102 on Hugging Face · Code ★ 25
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.


























