
Dream4ACT: A Shared Visual Action Interface for Multi-Embodiment Video-Action Modeling
Dream4ACT introduces action views to unify cross-embodiment joint actions as shared visual representations, enabling joint video-action modeling and 88.98% RoboTwin 2.0 success with training-free multiview recovery.
Published Sep 30, 2026 · 0 citations · ▲ 8 on Hugging Face
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.


























