Good Papers

Showing papers from MIT Media Lab (US); USI Lugano (CH) Show all papers

86%Must read
?Must readVote to see the score

LLM Alignment--Utility Asymmetry under Semantic-Preserving Transformations

Synthetic semantic-preserving transformations reveal alignment-utility asymmetry: LLMs retain task utility on shifted inputs but suffer sharp alignment failures, with harmful rates surging over 40 points despite minimal capability loss.

Mohan Li, Chengyu Yu, Francesco Sovrano, Marc Langheinrich and 1 more

Sydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 2/5