Good Papers

Showing papers from Tencent Hunyuan Show all papers

57%Worth a look
?Worth a lookVote to see the score

MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation

Zhiqi Li, Wenhuan Li, Tengfei Wang, Zhenwei Wang and 7 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

ReGen: Agentic Video World Modeling with Synergized Reasoning and Generation

Chenguo Lin, Yu Tang, Weiqiao Zheng, Enhua Jiang and 10 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Toward Embodied World Agents via Embodied-Planning Dataset and Interactive World Models

Xiaokun Feng, Junshu Tang, zeyi lin, Ling-Hao Chen and 3 more

Paris Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

HYDRA: Representation Harmonized Tokenization for Multimodal Generation and Understanding

Xuerui Qiu, Yutao Cui, Guozhen Zhang, Junzhe Li and 8 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
86%Must read
?Must readVote to see the score

DD-Ranking: Rethinking the Evaluation of Dataset Distillation

DD-Ranking reveals dataset distillation gains come from extra evaluation techniques rather than image quality, proposing fair metrics to assess true synthetic dataset value.

Zekai Li, Xinhao Zhong, Samir Khaki, Zhiyuan Liang and 36 more

Atlanta Poster Session 1, Wed, Dec 9, 10:00 AM–1:00 PM, Hall C1 · Published 2026

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 1/5
88%Must read
?Must readVote to see the score

Evidence-RL: Towards Evidence-intensive Visual Reasoning

Counterfactual Evidence Disentanglement (CED) audits vision-language model grounding by comparing evidence-region and non-evidence-region support drops inside GRPO, improving visual reasoning across benchmarks without inference overhead or evidence annotations.

Haojie Huang, Xinlei Yu, Chengming Xu, Zhangquan Chen and 5 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 16 on Hugging Face · Code ★ 4

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 4/5
medium 8/10
strict 3/5
72%Highly rated
?Highly ratedVote to see the score

Tango3D: Towards Alignment for Global and Local 2D-3D Correspondence

Tango3D unifies global retrieval and dense pixel-to-point correspondence via shared 2D-3D alignment with progressive training.

Zebin He, Mingxin Yang, Shuhui Yang, Hanxiao Sun and 3 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 4/5
medium 4/10
strict 0/5
76%Highly rated
?Highly ratedVote to see the score

Baton: Explicit Semantic Blueprints for Joint Video-Audio Generation

Baton introduces explicit semantic blueprints via a multimodal planner and relative positional encoding for synchronized joint video-audio generation.

Shuyuan Tu, Qi Tian, Zihan Yang, Yue Wu and 8 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 6/10
strict 0/5
80%Must read
?Must readVote to see the score

PRECISE: SDE-Consistent Stochastic Sampling for RL Post-Training of Flow-Matching Models

PRECISE introduces an SDE-consistent stochastic sampler balancing exploration and stability for RL post-training of flow-matching models, enabling faster, more stable reward optimization with significantly reduced training time.

Bo Peng, Tao Huang, Weijie Kong, Junzhe Li and 6 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 3/5
medium 8/10
strict 1/5
76%Highly rated
?Highly ratedVote to see the score

Exemplar2VQA: A Scalable Exemplar-Driven Visual Question Answering Generation Framework via Multi-Agent Coding

Exemplar2VQA uses multi-agent coding with geometric libraries to generate scalable 3D spatial question-answer pairs that improve MLLM spatial reasoning across indoor, outdoor, and mixed benchmarks.

Jiayu Ying, Qijian Tian, Ruijie Xu, Xinnan Zhu and 3 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 5/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

Hydra-X: Native Unified Multimodal Models with Holistic Visual Tokenizers

Hydra-X unifies image and video tokenization in one vision transformer via causal temporal attention and hierarchical compression, achieving strong unified understanding and generation performance.

Guozhen Zhang, Xuerui Qiu, Yutao Cui, Tianhui Song and 10 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 31 on Hugging Face

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 3/5
medium 5/10
strict 1/5
76%Highly rated
?Highly ratedVote to see the score

Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

Flow-DPPO replaces PPO ratio clipping with exact KL divergence constraints for flow matching models, improving reward, stability, and multi-objective alignment.

Bowen Ping, Xiangxin Zhou, Penghui Qi, Minnan Luo and 2 more

Sydney Poster Session 4, Wed, Dec 9, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026 · ▲ 42 on Hugging Face

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 6/10
strict 0/5