Good Papers

Showing papers from Dept. of Computer Sci. & Eng., Shanghai Jiao Tong University Show all papers

45%Niche pick
?Niche pickVote to see the score

Despa: Resolving Spatial Collapse in VLMs via Depth-Grounded Geometry

Yujing Lou, Pingyi Chen, Shen Cao, Lubin Fan and 3 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

From Views to Worlds: Active Exploration over 3D Worlds for Vision-Language Models

Qijian Tian, Jiayu Ying, Ke Fan, Lizhuang Ma and 1 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
70%Highly rated
?Highly ratedVote to see the score

DiPO: Disentangled Perplexity Policy Optimization for Fine-grained Exploration-Exploitation Trade-Off

DiPO disentangles perplexity into exploration and exploitation subspaces to enable fine-grained trade-offs, improving LLM reasoning and function calling via stable perplexity-guided policy optimization.

Xiaofan Li, Ming Yang, Zhiyuan Ma, Shichao Ma and 8 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 13 on Hugging Face

– ReadersNo votes yet
4/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 4 of 20 reviewers recommend it
lenient 2/5
medium 2/10
strict 0/5
76%Highly rated
?Highly ratedVote to see the score

VicEdit: Learning to Edit Videos from Visual In-Context Examples

VicEdit enables visual in-context video editing via multi-modal guidance and achieves state-of-the-art results on instruction and visual reference tasks.

Yuji Wang, Teng Hu, Yuheng Chen, Ran Yi and 5 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 4/10
strict 2/5