92%Must read
?Must readVote to see the score
General Agent Evaluation
A systematic comparison of general agent architectures finds backbone choice dominates performance while architecture shifts results up to 12pp, and open models suffer generality sinks.
Paris Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026 · ▲ 14 on Hugging Face · Code ★ 76
– ReadersNo votes yet
19/20 AI panelreviewers recommend it
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.
AI panel: 19 of 20 reviewers recommend it
lenient 5/5
medium 10/10
strict 4/5