
Iterative Policy Refinement through Semantic Rollout Analysis
A closed-loop framework iteratively refines structured imitation-learning policies via LLM analysis of rollout tables, improving performance by up to 15% and cutting compute 75%.
Published Oct 1, 2026 · 0 citations
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.


