
On the Geometry of On-Policy Distillation
On-policy distillation updates occupy a sparse, low-dimensional parameter subspace that is functionally sufficient and geometrically distinct from supervised fine-tuning and reinforcement learning.
Published Jun 5, 2026 · 0 citations · ▲ 75 on Hugging Face
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.

