Good Papers

Showing papers from Beijing University of Posts and Telecommunications Show all papers

72%Highly rated
?Highly ratedVote to see the score

On the Geometry of On-Policy Distillation

On-policy distillation updates occupy a sparse, low-dimensional parameter subspace that is functionally sufficient and geometrically distinct from supervised fine-tuning and reinforcement learning.

Zhennan Shen, Yanshu Li, Qingyu Yin, Chak Tou Leong and 5 more

Published Jun 5, 2026 · 0 citations · ▲ 75 on Hugging Face

– ReadersNo votes yet. 1 from authors or colleagues not counted
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 3/5
medium 4/10
strict 1/5
83%Must read
?Must readVote to see the score

SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection

SAHG detects LLM-driven social bots by applying direction-dependent hyperbolic curvature and dual-channel feature fusion, achieving top accuracy and F1 across three benchmarks.

Hanning Lu, Yingguang Yang, Jinwei Su, Yang; Liu and 7 more

Published May 28, 2026 · 0 citations

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
67%Highly rated
?Highly ratedVote to see the score

Learning Continuously Evolving Spatio-Temporal Explanations for Traffic Flow Forecasting

Cuiying Huo, Baoxu Wang, Lin Wu, Yu Mei and 3 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
2/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 2 of 20 reviewers recommend it
lenient 2/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Mixture-of-Top-$k$ Attention: Efficient Attention as Scalable Fast Weights

Qishuai Wen, Zhiyuan Huang, meng xianghan, Wei He and 1 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
67%Highly rated
?Highly ratedVote to see the score

Learning to Synergize Textual and Visual Prompts for Fine-Grained Traffic Element Detection in HD Maps

Xiaoyang Bi, Haowen Guo, Caoshengzhe Xue, Siyuan Li and 3 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
2/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 2 of 20 reviewers recommend it
lenient 2/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Concentrated Gradients Amplify Forgetting: Dominant-direction Projection for Continual Multimodal Learning

Chengxiang Huang, Haopeng Zhang, Yuzhe Han, Rui Dai and 3 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Behavioral Probes for Information Flow in LLM Swarms

Junhui Chen, Yu Zhang, Lantian Li, Dong Wang

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Patch4Patch: Restoring Structural Connectivity in Patch-based Vision Encoders

Yaqi Zhang, Shuntian Yao, Niantai Qu, Runguo Chen and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Efficient Collaborative LLM Fine-Tuning over Heterogeneous Mobile Devices via Many Backbones to One Side-Network Tuning

Xingke Yang, Liang Li, Sicong Li, Liwei Guan and 5 more

Atlanta Poster Session 5, Fri, Dec 11, 10:00 AM–1:00 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Learning Data-free Universal Adversarial Perturbation with Hybrid Priors and Gradient-Guided Sharpness Regularization

Zhi Tan, Jiazheng Cui, Wenwen Zhang, Yiran Liu and 3 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Anymotion: A Dataset, Benchmark, and Baseline for Controllable Human Motion Editing

Haiyang Yan, Jianxin Sun, Yuhan Wu, libin wang and 4 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

GEAR-Align: Grounding-Evidence-Aware Gradient Routing for Multimodal Alignment

Yu Yongkang, Haobo Wang, Meng Chen, Han Fang and 6 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Personalized Safety in Federated Fine-Tuning of Large Language Models

Tianzhe Xiao, Gaozhuo Liu, Yichen Li, Haozhao Wang and 3 more

Paris Poster Session 4, Thu, Dec 10, 5:30 PM–7:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

TreemapMix: Dirichlet-Controlled Multi-Image Augmentation for Probability and Ordinal Supervision

Ejafa Bassam, Konstantin Garbers, Yingsheng Geng, Dalton Jens and 2 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

CrossSteer:Cross-Modal Safety Steering for Audio-Language Models

Houde Dong, Weifei Jin, Yuxin Cao, Wei Song and 2 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

GeoG2U-Bench: When Does Generation Help Understanding in Ultra-High-Resolution Remote Sensing?

Fengxiang Wang, Yueying Li, Mingshuo Chen, Boya Miao and 11 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

DynEdit: Dynamic Entropy-Guided Sequential Editing for Large Language Models

Jinhu Fu, Yan Bai, Yihang Lou, Li Sun and 3 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Modality-Depth Routing for Visual Reasoning in VLM Post-Training

Yiming Ren, Yiran Xu, Chufan Shi, Yu Qiao and 2 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Heterogeneity-aware Distillation for Federated Continual Learning

Gaozhuo Liu, Yichen Li, Xiuying Wang, Yulong Li and 3 more

Paris Poster Session 1, Wed, Dec 9, 12:30 PM–2:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Rehearsal-Free Statistical Prototype Regularization for Federated Incremental Learning

Xiuying Wang, Yichen Li, Jiahua Cheng, Xiwei Liu and 3 more

Paris Poster Session 3, Thu, Dec 10, 12:30 PM–2:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

LINC: Decoupling Local Consequence Scoring from Hidden Matching in Constructive Neural Routing

LINC explicitly computes local routing consequences to score actions via shared linear comparison and context modulation, improving neural routing baselines especially at larger scales.

ShaoFeng Qin, Li Wang

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 4/5
medium 5/10
strict 0/5
78%Highly rated
?Highly ratedVote to see the score

NS-VLA: Towards Neuro-Symbolic Vision-Language-Action Models

NS-VLA introduces neuro-symbolic encoding and hierarchical optimization to improve robotic manipulation generalization and exploration over prior VLA methods.

Ziyue Zhu, Shangyang Wu, Shuai Zhao, Zhao ZhiQiu and 7 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 0/5
71%Highly rated
?Highly ratedVote to see the score

RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism

RotMoLE adds a rotation gate to MoE-LoRA experts that rotates rather than merely scaling selected experts, improving specialization and performance on multi-task and multilingual benchmarks.

Mengyang Sun, MaoChuan Dou, Tao Feng, Dan Zhang and 4 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
6/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 6 of 20 reviewers recommend it
lenient 4/5
medium 2/10
strict 0/5
83%Must read
?Must readVote to see the score

Beyond Appearance Shifts: Task-Semantic Action Calibration for VLA Models

BAS-VLA calibrates frozen VLA actions via breaking-centered calibration and selective preservation gating to suppress stale-task drift and separate semantics. It achieves 98% clean success, 0% under target swaps, and 70% under style shifts versus 42%.

Shuaijun Liu, Feiyang You, Chengyu Wu, Shuyang Hao and 5 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 4/5
medium 9/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating

CaC advances video reward models via hierarchical spatiotemporal concentrating, improving fine-grained anomaly accuracy by 25.7% and reducing generated-video anomalies by 11.7%.

Jiyuan Wang, Huan Ouyang, Chunyu Lin, Dewen Fan and 14 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 4/5
medium 4/10
strict 1/5
88%Must read
?Must readVote to see the score

RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations

RepoMirage evaluates code agents via repository perturbations, revealing severe repository context reasoning gaps and exploration drift, while RepoAnchor improves performance through structure-first scaffolding.

Hanyu Li, Yichi Zhang, Speed Zhu, Hang Su and 2 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 5/5
medium 9/10
strict 1/5
76%Highly rated
?Highly ratedVote to see the score

Search More, Think Less: Rethinking Long-Horizon Agentic Search for Efficiency and Generalization

SMTL replaces sequential reasoning with parallel evidence acquisition for efficient long-horizon agentic search, achieving state-of-the-art results on multiple benchmarks with far fewer reasoning steps.

Chengjun Yu, Shu XU, Jiaqi Wu, Qianben Chen and 20 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 23 on Hugging Face · Code ★ 3

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 6/10
strict 0/5
71%Highly rated
?Highly ratedVote to see the score

FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation

FaithfulFaces improves identity-preserving video generation via pose-shared identity alignment and achieves state-of-the-art consistency across pose changes and occlusions.

Yuanzhi Wang, Xuhua Ren, Jiaxiang Cheng, bing ma and 6 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 1 on Hugging Face

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 5/5
medium 2/10
strict 0/5
70%Highly rated
?Highly ratedVote to see the score

UxSID: Semantic-Aware User Interests Modeling for Ultra-Long Sequence

UxSID captures target-aware preferences via semantic-group shared interest memory and dual-level attention, achieving state-of-the-art results and 0.337% revenue lift.

Hongwei Zhang, qiqiang zhong, Jiangxia Cao, Junfeng Shu and 7 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
5/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 5 of 20 reviewers recommend it
lenient 4/5
medium 1/10
strict 0/5
70%Highly rated
?Highly ratedVote to see the score

Structure-Semantic Co-optimized Latent Diffusion Model for Fast Visual Anagram Synthesis

S2CO-Anagram applies structure-semantic co-optimization to adversarially distilled latent diffusion for faster, higher-resolution visual anagrams with improved visual harmony and semantic fidelity.

Xiang Gao, Yunpeng Jia

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
5/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 5 of 20 reviewers recommend it
lenient 3/5
medium 2/10
strict 0/5
86%Must read
?Must readVote to see the score

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective

EvoMemBench benchmarks LLM agent memory via self-evolving scope and content axes, finding no universal memory method and that long-context baselines remain competitive.

Yuyao Wang, Zhongjian Zhang, Mo Chi, Kaichi Yu and 6 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 1 on Hugging Face

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 3/5
86%Must read
?Must readVote to see the score

PAMod: Modeling Cyclical Shifts via Phase-Amplitude Modulation for Non-stationary Time Series Forecasting

PAMod models cyclical non-stationary shifts via phase-amplitude modulation in normalized space to achieve state-of-the-art forecasting with lower cost and broad plug-and-play gains.

Yingbo Zhou, Yutong Ye, Shuhao Li, Rui Qian and 4 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 1/5
83%Must read
?Must readVote to see the score

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning

LaST-R1 uses reinforcement learning with adaptive latent reasoning to optimize robotic action policies, achieving 99.9% success on LIBERO and up to 22.5% real-world gains.

Hao Chen, Zhonghao Yan, Jiaming Liu, Nuowei Han and 6 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 3 on Hugging Face · Code ★ 122

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
80%Must read
?Must readVote to see the score

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving

CoWorld-VLA embeds multi-expert world tokens into vision-language-action models and couples diffusion planning with scene context to generate continuous ego trajectories, improving autonomous driving performance.

Jingqi Wang, minqing huang, Zihan Liang, Yujiao Xiang and 6 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 1/5
86%Must read
?Must readVote to see the score

Counterfactual Rollout Replay: Forkable Environments as Free Process Rewards for Software Engineering Agents

Counterfactual Rollout Replay uses forkable environments to compute step-level return contrasts without human process labels, improving 14B SWE agent pass@1 by 5.0 points over outcome-only reinforcement learning.

Yuanhao li, Hongbo Wang, Xuhong Chen, Yiming Cao and 1 more

Atlanta Poster Session 1, Wed, Dec 9, 10:00 AM–1:00 PM, Hall C1 · Published 2026

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 1/5