Good Papers

Showing papers from Allen Institute for Artificial Intelligence Show all papers

78%Highly rated
?Highly ratedVote to see the score

Varying Shades of Wrong: Aligning LLMs with Wrong Answers Only

LLMs distinguish degrees of wrongness among incorrect answers, and alignment with such preferences yields less wrong answers and better calibration.

Jihan Yao, Wenxuan Ding, Shangbin Feng, Lucy Lu Wang and 1 more

Published Oct 14, 2024 · 0 citations · Code ★ 10

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 4/5
medium 5/10
strict 2/5
70%Highly rated
?Highly ratedVote to see the score

P3Sum: Preserving Author’s Perspective in News Summarization with Diffusion Language Models

P3Sum uses diffusion language models to preserve authorial perspective in news summarization, outperforming autoregressive baselines on perspective fidelity.

Yuhan Liu, Shangbin Feng, Xiaochuang Han, Vidhisha Balachandran and 3 more

Published 2024 · 2 citations

– ReadersNo votes yet
5/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 5 of 20 reviewers recommend it
lenient 3/5
medium 2/10
strict 0/5
72%Highly rated
?Highly ratedVote to see the score

LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition

LoraHub dynamically combines existing LoRA modules without extra parameters or gradients to generalize to unseen tasks with few examples, trading some accuracy for much lower inference token costs versus in-context learning.

Chengsong Huang, Qian Liu, Lin, Bill Yuchen, Tianyu Pang and 2 more

Published Jul 25, 2023 · 7 citations · ▲ 34 on Hugging Face · Code ★ 668

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 5/5
medium 3/10
strict 0/5
78%Highly rated
?Highly ratedVote to see the score

BotPercent: Estimating Bot Populations in Twitter Communities

BotPercent estimates community-specific Twitter bot populations by calibrating detection models across social contexts, revealing heterogeneous spatial-temporal bot distributions and achieving state-of-the-art community-level detection accuracy.

Zhaoxuan Tan, Shangbin Feng, Melanie Sclar, Herun Wan and 3 more

Published 2023 · 16 citations

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 0/5
86%Must read
?Must readVote to see the score

MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction

MolmoMotion predicts goal-conditioned 3D point trajectories from visual history and language, outperforming baselines on PointMotionBench and improving robot manipulation and video synthesis.

Jianing Zhang, Chenhao Zheng, Yajun Yang, Rustin Soraki and 6 more

Atlanta Poster Session 5, Fri, Dec 11, 10:00 AM–1:00 PM, Hall C1 · Published 2026 · ▲ 53 on Hugging Face · Code ★ 151

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 2/5
80%Must read
?Must readVote to see the score

Vision-Language Grounding as Bidirectional Concept Correspondence

Grounding is formulated as bidirectional concept correspondence to recover all image-text span correspondences without prespecified phrases via ConCor-1, improving F1 by 48% and 29% over baselines.

Jieyu Zhang, Ziqi Gao, Luke Zettlemoyer, Ranjay Krishna

Atlanta Poster Session 3, Thu, Dec 10, 10:00 AM–1:00 PM, Hall C1 · Published 2026 · ▲ 6 on Hugging Face · Code ★ 8

– ReadersNo votes yet
12/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 12 of 20 reviewers recommend it
lenient 4/5
medium 8/10
strict 0/5