AnyMo introduces OmniHuMo dataset with 5,000 hours of multimodal motion data and proposes a masked modeling framework for scalable any-modality conditional motion synthesis with flexible spatial and stylistic control.
Realtime-VLA FLASH uses a draft model and parallel verification to replace most full diffusion-based VLA inference rounds with faster speculative ones, cutting average latency 3.04x to 19.1 ms.
AtomWorld-Mem restores latent hidden dynamical states from atomistic snapshots via multi-scale memory to improve long-horizon kinetic Monte Carlo evolution and transfer across unseen alloys.
Knowledge-graph paths provide intermediate supervision for self-evolving search agents, improving question validity via relational context and solver rewards via waypoint coverage, boosting multi-hop QA across benchmarks.