Good Papers

Showing papers from NVIDIA Research Show all papers

45%Niche pick
?Niche pickVote to see the score

Full-Duplex Speech-Motion Model for Dyadic Interaction

Koki Nagano, Hongyu Liu, Wookie Park, Tianye Li and 7 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

SCULPT: Advancing Masked Discrete Diffusion for High-Resolution Image Synthesis.

Shufan Li, Greg Heinrich, Hanrong Ye, Yonggan Fu and 3 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
83%Must read
?Must readVote to see the score

Grounded-Exo2Ego: Structured Semantic Grounding for Robust Exocentric-to-Egocentric Video Generation

Grounded-Exo2Ego couples geometric anchoring with semantic grounding and camera relocalization to robustly generate egocentric video from exocentric inputs, outperforming prior methods on EgoExo4D.

Shengze Wang, Michael Stengel, Tianye Li, Wookie Park and 4 more

Atlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5
80%Must read
?Must readVote to see the score

Continuous Diffusion Scales Competitively with Discrete Diffusion for Language

RePlaid, a continuous diffusion language model aligned with modern discrete architectures, achieves scaling laws rivaling discrete diffusion and sets a continuous diffusion perplexity record of 22.1 on OpenWebText.

Zhihan Yang, Wei Guo, Shuibai Zhang, Subham Sahoo and 4 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 3/5
medium 7/10
strict 2/5
69%Highly rated
?Highly ratedVote to see the score

DiLaDiff: Distilled Latent-augmented Diffusion for Language Modeling

DiLaDiff proposes a latent-augmented masked diffusion language model with consistency distillation that improves quality and accelerates inference by generating continuous latents in negligible time.

Jean-Marie Lemercier, Tomas Geffner, Morteza Mardani, Karsten Kreis and 2 more

Paris Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
3/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 3 of 20 reviewers recommend it
lenient 2/5
medium 1/10
strict 0/5
71%Highly rated
?Highly ratedVote to see the score

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video

World from Motion generates dynamic 3D Gaussian reconstructions from monocular video via generative video modeling to fix artifacts and fill missing regions, achieving state-of-the-art 4D reconstruction.

Liyuan Zhu, Shengyu Huang, Amrita Mazumdar, Tianye Li and 5 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 3/5
medium 3/10
strict 1/5
88%Must read
?Must readVote to see the score

Fast 4D Mesh Generation by Spatio-Temporal Attention Chains

Spatio-Temporal Attention Chains accelerate training-free 4D mesh generation 13x to 9 seconds via latent temporal correspondences, improving quality, scaling to longer videos, and enabling tracking and camera estimation.

Dvir Samuel, Yuval Atzmon, Gal Chechik, Yoni Kasten

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026 · ▲ 11 on Hugging Face

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 2/5