Good Papers

Showing papers from UT-Austin Show all papers

88%Must read
?Must readVote to see the score

MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation

MURPHY extends GRPO to multi-turn code generation via feedback-conditioned rollout trees with retrospective credit assignment, achieving up to 6% absolute pass@1 gains over prior methods.

Chanakya Ekbote, Vijay Lingam, Sujay Sanghavi, Luke Huan and 3 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 5/5
medium 9/10
strict 1/5
83%Must read
?Must readVote to see the score

Enabling approximate joint sampling in diffusion LMs

A lightweight sampler layer on frozen diffusion LMs approximates joint token sampling, yielding MAUVE 0.87 versus 0.31 when unmasking four tokens per step.

Parikshit Bansal, Sujay Sanghavi

Atlanta Poster Session 4, Thu, Dec 10, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 2/5
76%Highly rated
?Highly ratedVote to see the score

Token Time Continuous Diffusion for Language Modeling

Token time continuous diffusion deterministically maps continuous noise to tokens with per-token times, outperforming discrete diffusion models at high speedups.

Parikshit Bansal, Sujay Sanghavi

Atlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026 · ▲ 4 on Hugging Face

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 4/5
medium 6/10
strict 0/5