Learning from Runtime Feedback through Failure-Bank Self-Evolution for Vision-Language-Action Models
FailBank turns runtime shield feedback into persistent VLA policy updates via failure-bank self-evolution, raising success rates up to 25.4 points and cutting policy-induced cost up to 35.6%.
Published Sep 30, 2026 · 0 citations · ▲ 15 on Hugging Face · Code ★ 2
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.







