Good Papers

Showing papers from Nanjing university Show all papers

57%Worth a look
?Worth a lookVote to see the score

TopoRefine: Topology-Aware Correspondence and Residual Refinement for Training-Free Subject-Consistent Generation

Zhanxin Gao, Zexin Ti, Chen Zhao, Beier Zhu and 4 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

TIDE: Trajectory-Aware Watermark Propagation for Text-to-Image Diffusion Models

Yihan Meng, Suping Xu, Yanfeng Wu, Chongjun Wang and 1 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

The Price of Locality in Label Privacy: Optimal Rates for Classification and Regression

Zongrui Zou, Mina Dalirrooyfard, Jingcheng Liu, Jalaj Upadhyay

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
76%Highly rated
?Highly ratedVote to see the score

Automated Reformulation of Robust Optimization via Memory-Augmented Large Language Models

AutoREM is a tuning-free memory-augmented framework that automates robust optimization reformulation via experience memory and improves accuracy across models.

Jinbiao Chen, Shuang Jin, Guoyun Zhang, Junyu Zhang and 2 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 5/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

ViDiC: Video Difference Captioning

ViDiC introduces a video difference captioning task and ViDiC-1K benchmark that reveals large multimodal models struggle with fine-grained comparative video perception.

Jiangtao Wu, Shihao Li, Zhaozhou Bian, Jialu Chen and 6 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 29 on Hugging Face · Code ★ 15

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 5/5
medium 3/10
strict 1/5
86%Must read
?Must readVote to see the score

World–Value–Action Model: Implicit Planning for Vision–Language–Action Systems

WAV introduces a latent-space planning framework for vision-language-action models that predicts future states and evaluates trajectory values to enable efficient long-horizon decision-making.

Runze Li, Hongyin Zhang, Junxi Jin, Qixin Zeng and 4 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 1/5
76%Highly rated
?Highly ratedVote to see the score

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models

Hierarchical Token GRPO integrates sketch-then-paint stages and hierarchical credit assignment into diffusion multi-modal LLM reinforcement learning, substantially improving image quality and benchmark scores.

Siqi Luo, Jianghan Shen, Yi Xin, Huayu Zheng and 8 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 6/10
strict 0/5
83%Must read
?Must readVote to see the score

LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning

LatentOmni replaces text chain-of-thought with interleaved audio-visual latent reasoning states to preserve dense sensory signals, improving joint reasoning over explicit text baselines.

Yifan Dai, zhenhua wu, Bohan Zeng, Daili Hua and 17 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 44 on Hugging Face · Code ★ 24

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 4/5
medium 9/10
strict 0/5
88%Must read
?Must readVote to see the score

WebNavigator: Global Web Navigation via Interaction Graph Retrieval

WebNavigator overcomes topological blindness via interaction graphs to turn web navigation into deterministic retrieval and pathfinding, doubling multi-site success on WebArena.

Xuanwang Zhang, Yuteng Han, Jinnan Qi, Xinyu Liu and 3 more

Atlanta Poster Session 1, Wed, Dec 9, 10:00 AM–1:00 PM, Hall C1 · Published 2026

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 2/5