80%Must read
?Must readVote to see the score

Sparse Layers are Critical to Scaling Looped Language Models
Looped-MoE models scale better than standard transformers via routing divergence that recovers expressivity, and loop boundaries enable efficient early exits with minimal quality loss.
Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026
– ReadersNo votes yet
12/20 AI panelreviewers recommend it
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.
AI panel: 12 of 20 reviewers recommend it
lenient 4/5
medium 7/10
strict 1/5