Good Papers

Showing Robot manipulation Show all papers

76%Highly rated
?Highly ratedVote to see the score

EmbodiedSmith: Scaling Embodied Data through Recursive Self-Improvement Flywheel in Simulation

EmbodiedSmith unifies asset, scene, and task generation in a recursive self-improvement loop to scale embodied simulation data, improving generation success and robot policy generalization across diverse embodiments and physics.

Yikai Qin, Yifei Deng, Mingjian Liang, Wenxuan Song and 12 more

Published Oct 6, 2026 · ▲ 7 on Hugging Face

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 5/10
strict 0/5
83%Must read
?Must readVote to see the score

Execution-Aligned Progressive Noise for Consistent Asynchronous Replanning in Generative Robot Policies

Execution-Aligned Progressive Noise structures noise across chunks and time to maintain consistent generative states during asynchronous replanning, achieving up to 96.7% real-robot success.

Di Wu, Ping Liu, Xuhua Chen, He Zheng and 2 more

Published Oct 5, 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 2/5
78%Highly rated
?Highly ratedVote to see the score

Conditional Trajectory Peaks: Single-Pass Multimodal Policies over Action Chunks

Conditional Trajectory Peaks predicts multimodal action-chunk candidates in a single pass, achieving 97.25% LIBERO success and 3× faster inference while preserving diverse behaviors.

Di Wu, Rongtian Shen, Ping Liu, Xuhua Chen and 3 more

Published Oct 5, 2026 · ▲ 1 on Hugging Face

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 3/5
medium 5/10
strict 3/5
80%Highly rated
?Highly ratedVote to see the score

InterMimicGen: Scaling Humanoid Loco-Manipulation through Self-Evolving Motion Imitation

InterMimicGen retargets human motion capture to humanoids and self-evolves tracking data via iterative simulation-verified augmentation to scale dexterous loco-manipulation.

Yucheng Zhang, Sirui Xu, Jinhong Li, Liuyu Bian and 7 more

Published Oct 5, 2026 · ▲ 4 on Hugging Face

100% Readers1 of 1 upvoted
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 1/5
76%Highly rated
?Highly ratedVote to see the score

RobotUse: Allocating Computation, Context, and Decisions

RobotUse organizes robot computation, context, and decisions around revisable physical actions via visual target selection and persistent playbooks, achieving 45% RoboLab success and real-world learning.

Junhoo Lee, Injun Baek, Seungyeon Kim, Suhyun Jeon and 3 more

Published Oct 4, 2026 · ▲ 14 on Hugging Face · Code ★ 4

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 5/10
strict 1/5
90%Must read
?Must readVote to see the score

World Action Learning via Interaction-Centric Spectral Latent Guidance

WING distills interaction-centric latent actions from egocentric video and uses cross-embodiment spectral low-frequency guidance to transfer them to robot policies, achieving high success rates on LIBERO, RoboTwin, RoboCasa, and real-world tasks.

Zhiming Liu, Yikun Miao, Ying Chen, Hongrui Yin and 6 more

Published Oct 2, 2026 · 0 citations · ▲ 20 on Hugging Face

100% Readers1 of 1 upvoted
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

89%Must read
?Must readVote to see the score

EyeRobot 2.0: Active Gaze for Precise Manipulation without Wrist Cameras

EyeRobot 2.0 uses active gaze and fixation-relative frames to enable precise bimanual manipulation with only a single stereo camera, outperforming passive stereo by 40% in real-world trials and doubling ego-plus-wrist success under occlusion.

Kush Hari, Justin H. Kerr, Nidhya Shivakumar, Samarth Mahapatra and 6 more

Published Oct 2, 2026 · 0 citations · ▲ 10 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

89%Must read
?Must readVote to see the score

PointWAM: 3D World Action Modeling for Dexterous Robotic Manipulation

PointWAM forecasts 3D point trajectories of scenes and hands in a shared space-time frame to guide dexterous robot manipulation, improving DexJoCo success by 56.9 points with video pre-training.

Chunghyun Park, Beomjun Kim, Seungcheol Park, 권희승 and 4 more

Published Oct 2, 2026 · 0 citations · ▲ 42 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

78%Highly rated
?Highly ratedVote to see the score

InterEvolve: Test-Time Evolution of Reward Programs for Humanoid Loco-Manipulation

InterEvolve evolves reward programs at test time via an LLM agent and numerical optimizer to compose a humanoid controller's skills for novel loco-manipulation tasks. The approach releases latent controller competence through object-aware forward-backward models, generating novel strategies that tra

Zhuo Lin, Sirui Xu, Liuyu Bian, Yu-Xiong Wang and 1 more

Published Oct 1, 2026 · 0 citations · ▲ 61 on Hugging Face

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

86%Must read
?Must readVote to see the score

OpenRUA: Robot-Use Agents Are Zero-Shot Visuomotor Policies

OpenRUA gives off-the-shelf coding agents only ROS 2 terminal access to act as zero-shot visuomotor policies, achieving 99% on CaP-Bench and 87% on LIBERO-PRO without bespoke harnesses or training, and reveals emergent perception and closed-loop control behaviors.

Zhaoyang Chu, Earl T. Barr, Claire Le Goues, Peter W. O’Hearn and 3 more

Published Oct 1, 2026 · 0 citations · ▲ 5 on Hugging Face · Code ★ 3

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

78%Highly rated
?Highly ratedVote to see the score

Magic-W0: A Structured World-Action Foundation Model for Physical Intelligence

Magic-W0 is a structured world-action foundation model that couples 3D geometry, motion, and future semantics with action generation via layer-aligned interaction, achieving top simulated scores and strong real-robot adaptation.

Xuhua Chen, 尹貞漢, Yuan Zhang, Lingfeng Zhang and 14 more

Published Sep 30, 2026 · 0 citations · Code ★ 2

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

74%Highly rated
?Highly ratedVote to see the score

Beyond the Current Scene: Event-Referential Grasping with Active View Selection

BeyondSCe enables zero-shot event-referential grasping via active view selection, achieving 76% and 77% success on visible and occluded targets versus 40% and 55% baselines.

Hyunjoon Lee, Haebeom Jung, Eunsung Cha, Daeun Lee and 3 more

Published Sep 30, 2026 · 0 citations · ▲ 51 on Hugging Face · Code ★ 8

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

78%Highly rated
?Highly ratedVote to see the score

Dream4ACT: A Shared Visual Action Interface for Multi-Embodiment Video-Action Modeling

Dream4ACT introduces action views to unify cross-embodiment joint actions as shared visual representations, enabling joint video-action modeling and 88.98% RoboTwin 2.0 success with training-free multiview recovery.

Xiangyu Zhu, Jin Xu, Yue Guo, Xin Wu and 5 more

Published Sep 30, 2026 · 0 citations · ▲ 8 on Hugging Face

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

91%Must read

MotorMind: Scaffolding General Vision Language Models for Zero-Shot Robot Manipulation

MotorMind connects general vision-language models to deterministic robot control via mid-level actions and feedback loops, achieving 66.7% zero-shot success on LIBERO-PRO and 95% on real robots without task-specific training or external tools.

Bingxuan Li, Siqi Song, Yizhuo Wu, Jiarui Yao and 2 more

Published Sep 29, 2026 · 0 citations · ▲ 90 on Hugging Face

– ReadersNo votes yet
18/20 AI panelreviewers recommend it

Only vote on papers you've read. Sign in with GitHub to vote.

88%Must read
?Must readVote to see the score

Agent Priors-guided Policy Learning

Agent Priors-guided Policy Learning embeds structural priors in skill interfaces to enable compositional and out-of-distribution skill generalization.

Puming (Oscar) Jiang, Tao Hu, Haozhe Du, Yibo Li and 3 more

Published Sep 28, 2026 · 0 citations · ▲ 84 on Hugging Face · Code ★ 3

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 5/5
medium 9/10
strict 1/5
89%Must read
?Must readVote to see the score

MM-ABC: Towards Generalist Mobile Manipulation via Seeing, Coordinating and Imagining

MM-ABC is a mobile manipulation foundation model combining multi-level vision features, future imagination supervision, and masked joint attention to coordinate arm-base actions, achieving up to 99.1% success across benchmarks and 83% in real-world tasks.

Qiwei Liang, Guangyu Chen, Shaolong Zhu, Zikuan Xiao and 5 more

Published Sep 28, 2026 · 0 citations · ▲ 5 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 5/5
medium 9/10
strict 2/5
45%Niche pick
?Niche pickVote to see the score

Semantics-to-Contact: A Stagewise Framework for Robust Contact-Rich Manipulation

Guanren Qiao, Ruixiang Ouyang, Sheng Xu, Yueci Deng and 3 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
67%Highly rated
?Highly ratedVote to see the score

Hydra-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control

Jinhao Zhang, Zhexuan Zhou, Huizhe Li, Yichen Lai and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
2/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 2 of 20 reviewers recommend it
lenient 2/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

UniRAP: Towards Unified Part-level Physical Affordance Reasoning and Actionable Perception

Linfei Li, Ruining Hu, Lin Zhang, Zhong Wang and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

InfCoiL: Coordinated Planner-Controller Learning for Closed-Loop Physics-Based Human-Object Interaction

Yude Zou, Yifei Yao, Yiwei Hao, Hanqing Wang and 1 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Towards Scalable Egocentric HOI for Humanoids: Benchmarking Whole-Body Dexterous Interaction with Tactile Prediction

Zhenyu Wei, Haoyang Luo, Chixuan Zhang, Guo Chen and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

LIBERO-PeRM: Benchmarking Personalized Robotic Manipulation

Zhixu Li, Keqian Tang, Litian Gong, Jingyu Yao and 6 more

Atlanta Poster Session 5, Fri, Dec 11, 10:00 AM–1:00 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Scaling Whole-Body Loco-Manipulation through Compositional Data Synthesis

Yufei Zhu, Runyi Yu, Yinhuai Wang, Aoru Xue and 5 more

Atlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
67%Highly rated
?Highly ratedVote to see the score

DexOPE: 6D Object Pose Estimation in Dexterous Manipulation

Ke Wu, Zhiwei Yang, Xiangting Meng, Zicheng Zhang and 4 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
2/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 2 of 20 reviewers recommend it
lenient 2/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

CSO: Refining Robotic Policies via Skill Distribution Alignment and Skill-Grained Optimization

Zhiyuan Xiang, Xiang Deng, Wanjie Tao, Xu Chen and 2 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

From Sample to Subset Construction: Coverage-Aware Curation of Robot Demonstrations

Shuo Liu, Minxin Lai, Wuyang Zhang, Yu Zhang and 3 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Learning Active Perception and Manipulation via Spatio-temporal Visual Memory

Enshen Zhou, Mengzhen Liu, Yibo Li, Yanjun Ding and 5 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

ToS: Tree-of-Skill Reinforcement Learning driven by Behavior Tree for Long-Horizon Manipulation

Xinglin Chen, Yishuai Cai, Yunxin Mao, Minglong Li and 5 more

Paris Poster Session 3, Thu, Dec 10, 12:30 PM–2:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

PACE: Progress Actively-internalized Conditioning Execution for Long-horizon Manipulation

Fei Ni, Zhuo Chen, Xianze Yao, Zerui Chen and 5 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

IDEAL: Interaction Dynamics and Force-Aware Learning for Dual-Humanoid Collaborative Manipulation

Zhaoyang Li, Chiyu Zhang, Chaoyue Li, Xiao Zhang and 2 more

Paris Poster Session 6, Fri, Dec 11, 2:30 PM–4:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

OmniDex: Scaling Dexterous Hand Grasping to Diverse Cluttered Scenes

Naiyu Fang, Zhongjin Luo, Yuxin Mo, Siyuan Huang and 6 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

GeoMIND: A Benchmark for Spatial Understanding in Robotic Manipulation

Jinghe Wang, Xinrui Cao, Duo Wu, Chenghao Gu and 4 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

OntoPlan: An Ontology-Grounded Scene Representation and Agentic Framework for Scalable Robot Task Planning

Hyeongwoo Nam, Woongje Cho, Juwon Kim, Jongeun Choi

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
71%Highly rated
?Highly ratedVote to see the score

Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry

IDP uses conditional expert geometry to implicitly enforce training-time drifting corrections, yielding a one-step generative policy that improves robot control speed while maintaining valid action manifolds.

Zemin Yang, Yaoyu He, Yiming Zhong, Yuhao Zhang and 4 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 3/5
medium 4/10
strict 0/5
76%Highly rated
?Highly ratedVote to see the score

Masked Visual Actions for Unified World Modeling

Masked Visual Actions expresses robot and object motion as revealed pixel trajectories to unify forward dynamics, planning, and inverse modeling in video world models with minimal finetuning.

Hadi Alzayer, Wenlong Huang, Haonan Chen, Christopher Luey and 7 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 9 on Hugging Face · Code ★ 110

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 1/5
71%Highly rated
?Highly ratedVote to see the score

DeformGen: Dynamics-Based Topology Augmentation for Deformable Manipulation Policy Learning

DeformGen uses dynamics-based topological augmentation to generate diverse deformable object states and warp trajectories for improved manipulation policy learning.

Zili Lin, Wenyao Zhang, Yuyang Zhang, Zekun Qi and 8 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 4/5
medium 2/10
strict 1/5
80%Must read
?Must readVote to see the score

HandEdit: A Unified Benchmark for Egocentric Human-to-Robot Dexterous Hand Image Editing

HandEdit provides a 200M-instance benchmark and dataset for transforming egocentric human hands into diverse dexterous robot embodiments via image editing. It evaluates 11 baselines across hand-only and hand-arm tracks with embodiment-aware metrics.

Zhenjie Yang, Xingyu Jiao, Guopeng Zhong, Shuzhe Yang and 17 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 1/5
72%Highly rated
?Highly ratedVote to see the score

MyoChallenge 2025: A New Benchmark for Human Athletic Intelligence

MyoChallenge 2025 benchmarks musculoskeletal sports control via simulated table tennis and soccer tasks, advancing agile motor algorithms across 70 teams.

Cheryl Wang, Chun Kwang Tan, Balint Hodossy, Shirui Lyu and 20 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 5/5
medium 2/10
strict 1/5
74%Highly rated
?Highly ratedVote to see the score

JODA: Composable Joint Dynamics for Articulated Objects

JODA parameterizes articulated joint dynamics as structured three-channel fields via PCHIP interpolation, inferring and refining them from visual observations using vision-language models for controllable simulation.

Tianhong Gao, Cheng Yu, Yinghao Xu, Mengyu Chu

Atlanta Poster Session 4, Thu, Dec 10, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 0/5
72%Highly rated
?Highly ratedVote to see the score

PO-PDDL: Learning Symbolic POMDPs from Visual Demonstrations for Robot Planning Under Uncertainty

PO-PDDL learns symbolic POMDPs from robot videos to enable belief-space planning under partial observability and stochasticity, outperforming prior methods with lower planning cost.

Wenjing Tang, Jin Xuanjin, Yuan Liu, RenMing Huang and 2 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 5/5
medium 3/10
strict 0/5
86%Must read
?Must readVote to see the score

SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework

SUGAR converts human videos into humanoid loco-manipulation skills via automated priors, physics refinement, and policy distillation, scaling with video data and enabling zero-shot real-world transfer.

Tianshu Wu, Xiangqi Kong, Yue Chen, Qize Yu and 4 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 2/5
80%Must read
?Must readVote to see the score

RigidFormer: Learning Rigid Dynamics using Transformers

RigidFormer is a transformer that learns mesh-free rigid-body dynamics via object-level anchors and differentiable Kabsch projection, outperforming mesh-based baselines with faster inference and scalability to 200+ objects.

Zhiyang Dou, Minghao Guo, Haixu Wu, Doug Roble and 2 more

Atlanta Poster Session 3, Thu, Dec 10, 10:00 AM–1:00 PM, Hall C1 · Published 2026 · ▲ 13 on Hugging Face · Code ★ 93

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 1/5
80%Must read
?Must readVote to see the score

Point Tracking Improves World Action Models

JOPAT predicts tracks and pixels via diffusion transformers to learn dynamics robust to occlusion and appearance variation, improving long-horizon robot policy performance.

Jiarui Guan, Wenshuai Zhao, Yue Pei, Ziliang Chen and 2 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 1/5
89%Must read
?Must readVote to see the score

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning

SOLE-R1 is a video-language reasoning model providing dense progress rewards for online robot reinforcement learning, enabling zero-shot unseen manipulation without ground-truth rewards and outperforming prior vision-language rewarders with less reward hacking.

Philip Schroeder, Thomas Weng, Karl Schmeckpeper, Eric Rosen and 2 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 1 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 4/5
76%Highly rated
?Highly ratedVote to see the score

Learning Options for Compositional Motor Control with Adapter Banks

A shared recurrent core with residual adapter banks learns compositional motor skills via emergent low-rank dynamics, cutting generalization error versus multitask baselines by up to an order of magnitude.

Sreejan Kumar, Marcelo G Mattar, Lea Duncker

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 5/10
strict 0/5
91%Must read
?Must readVote to see the score

Feedback World Model Enables Precise Guidance of Diffusion Policy

A feedback world model updates predictions online with real observations to correct errors, reducing prediction error by up to 76.4% and improving out-of-distribution policy success by 30%.

Tuo An, Jindou Jia, Gen Li, Jingliang Li and 7 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
17/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 17 of 20 reviewers recommend it
lenient 5/5
medium 10/10
strict 2/5
83%Must read
?Must readVote to see the score

FLASH: Efficient Visuomotor Policy via Sparse Sampling

FLASH represents continuous robot actions with sparse Legendre polynomials and history-anchored flow matching to achieve real-time, accurate, and rapid visuomotor control.

Jiaqi Bai, Jindou Jia, Yuxuan Hu, Gen Li and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 2/5
89%Must read
?Must readVote to see the score

PointZero: 3D Point Track Completion for Learning Transferable 3D Dynamics

PointZero predicts full 3D point tracks from sparse tracks and RGB-D to learn transferable dynamics without robot actions, outperforming baselines on dynamics and manipulation tasks.

Bardienus Duisterhof, Kaifeng Zhang, Adam Hung, Bowen Wen and 4 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 5/5
medium 9/10
strict 2/5
72%Highly rated
?Highly ratedVote to see the score

ForceFlow: Learning to Feel and Act via Contact-Driven Flow Matching

ForceFlow uses force-aware flow matching with asymmetric multimodal fusion and vision-to-force handover to achieve robust contact-rich manipulation with 37% higher success and stronger zero-shot generalization.

Shuoheng Zhang, Yifu Yuan, Hongyao Tang, YAN ZHENG and 6 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 5/5
medium 3/10
strict 0/5
78%Highly rated
?Highly ratedVote to see the score

Reverse to Advance: Teleoperation-Cost Effective Hard Policy Learning from Reversed Easy Tasks

It proposes reversing easy-task trajectories to train hard-task policies via automated collection, hierarchical refinement, and iterative learning, achieving higher success with less teleoperation.

Qiyuan Qiao, Ge Yuan, Can Wang, Dong Xu

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

Runtime Monitoring of Perception-Based Autonomous Systems via Embedding Temporal Logic

ETL monitors perception-based autonomous systems directly in learned embedding spaces via distance-based temporal logic predicates, enabling reliable specification of high-level visual behaviors with conformal calibration.

Parv Kapoor, Abigail Hammer, Ashish Kapoor, Karen Leung and 1 more

Atlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 0/5
83%Must read
?Must readVote to see the score

Action Images: End-to-End Policy Learning via Multiview Video Generation

Action Images formulates robot policy learning as multiview video generation using interpretable pixel-grounded action images, enabling zero-shot control without separate policy heads and improving video-action joint generation.

Haoyu Zhen, Zixian Gao, Qiao Sun, yilin zhao and 6 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026 · ▲ 8 on Hugging Face

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
72%Highly rated
?Highly ratedVote to see the score

NDPP-Grasp: Non-Differentiable Physical Plausibility Constraint-Guided Task-Oriented Dexterous Grasp Generation

NDPP-Grasp injects non-differentiable physical plausibility guidance into grasp diffusion denoising, improving task-oriented dexterous grasp physical plausibility while preserving alignment.

Qiuchi Xiang, Haoxuan Qu, Hossein Rahmani, Jun Liu

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 5/5
medium 3/10
strict 0/5
72%Highly rated
?Highly ratedVote to see the score

EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields

EA-WM projects robotic actions into camera-aligned visual fields and fuses them via event-aware attention to preserve spatial geometry and interaction dynamics, achieving state-of-the-art results on WorldArena.

Zhaoyang Yang, Yurun Jin, Lizhe Qi, Cong Huang and 1 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 4/5
medium 4/10
strict 0/5
78%Highly rated
?Highly ratedVote to see the score

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation

GTP-FA decouples grasping from planning with failure attribution to diagnose failures and optimize both modules, substantially improving robotic manipulation success across diverse policy learners.

Jiahao Xu, Peiyuan Wang, Hanzhuo Zhang, Zihao Yu and 7 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 0/5
83%Must read
?Must readVote to see the score

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning

LaST-R1 uses reinforcement learning with adaptive latent reasoning to optimize robotic action policies, achieving 99.9% success on LIBERO and up to 22.5% real-world gains.

Hao Chen, Zhonghao Yan, Jiaming Liu, Nuowei Han and 6 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 3 on Hugging Face · Code ★ 122

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
89%Must read
?Must readVote to see the score

EgoTac: In-the-wild Tactile Prediction from Egocentric Vision

EgoTac predicts tactile signals from egocentric videos using 5.7M image-tactile pairs, achieving under 0.06N force error and outperforming contact estimators.

Wenkang Zhang, Chengbo Yuan, Zicheng Zhang, Zhengxue Cheng and 1 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 3/5
83%Must read
?Must readVote to see the score

FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation

FlowHOI generates semantically grounded hand-object interaction sequences via two-stage flow matching, achieving 1.7x higher simulation success and 40x faster inference than diffusion baselines.

Huajian Zeng, Lingyun Chen, Jiaqi Yang, Yuantai Zhang and 3 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
89%Must read
?Must readVote to see the score

DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation

DynaFLIP pre-trains dynamics-aware visual encoders via image-language-3D flow alignment, boosting robot manipulation generalization by up to 22.5%.

Jusuk Lee, Seungjae Lee, Jonghun Shin, Hoseong Jung and 5 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 9 on Hugging Face

– ReadersNo votes yet
16/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 16 of 20 reviewers recommend it
lenient 5/5
medium 9/10
strict 2/5