Good Papers

Showing papers from Institute of Information Engineering, Chinese Academy of Sciences Show all papers

80%Highly rated
?Highly ratedVote to see the score

Co-Evolving Policy Distillation

Co-Evolving Policy Distillation co-trains experts via bidirectional online policy distillation during RLVR to avoid divergence and absorption gaps, integrating multi-modal reasoning to surpass domain-specific experts.

Naibin Gu, Chenxu Yang, Qingyi Si, Chuanyu Qin and 6 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 67 on Hugging Face

100% Readers1 of 1 upvoted
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 2/5
medium 7/10
strict 1/5
45%Niche pick
?Niche pickVote to see the score

ToPA: Block-wise Toeplitz Adaptation for Expressive and Efficient Fine-Tuning

Sicong Li, Qianqian Xu, Zhiyong Yang, Zitai Wang and 3 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Iterative Nonlinear Computation Underlying Abstract Reasoning

Zitian Gao, Yilong Chen, Yihao Xiao, Xinyu Yang and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

BootstrapAgent: Turning Repository Setup into Reusable Agent Knowledge

Sihan Fu, Oucheng Liu, Shiyuan Wang, Jin Shi and 1 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

SkillCIR: Intent-Guided Skill Composition for Training-Free Composed Image Retrieval

Yuanmin Tang, Lin Li, Yang Du, Yuan Gao and 7 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
78%Highly rated
?Highly ratedVote to see the score

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

VeriTrip benchmarks travel planning agents via evidence-grounded reasoning over noisy multimodal web corpora, revealing a retrieval-reasoning trade-off that erodes instruction retention.

Yuting Xu, Jiayi Tian, Jian Liang, Xin Xiong and 3 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 5/5
medium 5/10
strict 1/5
76%Highly rated

Self-Distilled RLVR

RLSD combines RLVR and self-distillation, using token-level policy differences for update magnitudes and environmental feedback for directions, improving convergence and stability.

Chenxu Yang, Chuanyu Qin, Qingyi Si, Minghui Chen and 6 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 145 on Hugging Face

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 3/5
medium 6/10
strict 1/5
80%Must read
?Must readVote to see the score

Harnessing Streaming Video in the Wild

Streaming-Train-248K and Streaming Harness adapt VLMs to real-time video streams with proactive interaction, 12-hour memory, and sub-second latency.

Dingyu Yao, Shuhuan Gu, Qingyi Si, Junhao Zhou and 7 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 5/5
medium 5/10
strict 2/5
80%Must read
?Must readVote to see the score

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding

GUI-SD uses on-policy self-distillation with privileged visual contexts and entropy-guided distillation for GUI grounding, outperforming GRPO methods in accuracy and efficiency.

Yan Zhang, Daiqing Wu, Huawen Shen, Can Ma and 1 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 5 on Hugging Face

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 0/5
83%Must read
?Must readVote to see the score

NP-LoRA: Null Space Projection for Subject-Style LoRA Fusion

NP-LoRA fuses subject and style LoRAs via null-space projection to reduce subspace interference, improving controllable generation without retraining.

Chuheng Chen, Xiaofei Zhou, Geyuan Zhang, Yong Huang and 1 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
78%Highly rated
?Highly ratedVote to see the score

Mitigating Overthinking in Large Reasoning Language Models via Reasoning Path Deviation Monitoring

A reasoning path deviation metric detects high-entropy overthinking tokens to dynamically terminate redundant reasoning, improving performance and efficiency over existing early-exit methods.

Weixin Guan, Liang Li, Jiapeng Liu, Bing Li and 5 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents

SPIRAL uses sequential planning and reflective agents in a closed loop to generate long-horizon action-conditioned videos with iterative refinement and self-evolving post-training.

Yu Yang, Yue Liao, Jianbiao Mei, Baisen Wang and 7 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 4/5
medium 5/10
strict 0/5