
Flowing Faster to Coordinate: One-Step Online Multi-Agent Flow Policies
OMAF proposes a one-step flow policy framework for online multi-agent reinforcement learning, achieving up to 3.4x higher returns and 10.5x sample efficiency over baselines.
Published Oct 1, 2026 · 0 citations
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.










