Good Papers

Showing Multi-agent LLM systems Show all papers

88%Must read
?Must readVote to see the score

Dynamic Harness Search: Building Multi-Agent Systems Per-Query via Prediction

SHIFT predicts harness utility via a local LLM to search multi-agent structures per query, achieving ~80% mean accuracy across benchmarks while reducing execution tokens by 32%.

Som Sagar, Shasha Li, Hejie Cui, Ransalu Senanayake and 1 more

Published Oct 2, 2026 · ▲ 10 on Hugging Face

– ReadersNo votes yet
15/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 15 of 20 reviewers recommend it
lenient 4/5
medium 8/10
strict 3/5
83%Must read
?Must readVote to see the score

Science Utopia? Closed-Loop LLM Simulation of Academic Research Ecosystems

SciUtopia is a closed-loop LLM simulation framework modeling entire academic ecosystems; it finds resubmission amplifies reviewer burden, cautious exploration balances impact and diversity, and inequality can emerge without cumulative funding advantage.

Yiqiao Jin, Yiyang Wang, Lucheng Fu, Bing He and 7 more

Published Oct 1, 2026 · 0 citations · ▲ 40 on Hugging Face · Code ★ 29

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 2/5
74%Highly rated
?Highly ratedVote to see the score

Understanding Issues, Causes and Solutions in Open-Source LLM-based Multi-Agent Systems

Open-source LLM multi-agent systems face orchestration and execution issues mostly caused by workflow, tool integration, and memory problems, primarily solved by workflow optimization.

Asad Ur Rehman, Syed Mohammad Kashif, Ruiyin Li, Peng Liang and 2 more

Published Oct 1, 2026 · 0 citations

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 5/5
medium 3/10
strict 1/5
72%Highly rated
?Highly ratedVote to see the score

Auditing Action Settlement in LLM Agent Environments: Order, Progress, and Replay

A typed settlement contract audits concurrent LLM agent actions, showing joint policies complete 59% more six-agent doorway tasks than conservative rejection, with exact replay of 156 checkpoints and rejection of 1,332 corruptions.

Haotian Chen, Bowen Ye, Yuning Zhang, Jingkun Yu

Published Oct 1, 2026 · 0 citations

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 3/5
medium 3/10
strict 2/5
78%Highly rated
?Highly ratedVote to see the score

Raven: The Harness of Harnesses for Composable Agentic Intelligence

Raven autonomously constructs modular agent harnesses and orchestrates them across domains via a multi-agent ecosystem, significantly outperforming state-of-the-art systems on complex long-horizon tasks.

EverMind AI

Published Sep 27, 2026 · 0 citations · ▲ 566 on Hugging Face · Code ★ 5,252

– ReadersNo votes yet
11/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 11 of 20 reviewers recommend it
lenient 5/5
medium 6/10
strict 0/5
83%Must read
?Must readVote to see the score

Receiver-Conditioned Latent Communication gives 94% CacheBack

CacheBack uses receiver-conditioned filtering of sender KV caches via attention weights to cut transferred state by 75%, boosting multi-agent accuracy by 14.7 points and reducing latency 3.2x versus text.

Maximillian Rossi, Prajwal Raghunath, Haoqing Xuan, Yusen Zhang and 1 more

Published Sep 25, 2026 · 0 citations · ▲ 11 on Hugging Face · Code ★ 6

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 8/10
strict 0/5
74%Highly rated
?Highly ratedVote to see the score

ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration

ARIS is an open-source autonomous research harness using cross-model adversarial collaboration to coordinate ML workflows and verify experimental claims.

Ruofeng Yang, Yongcan Li, Shuai Li

Published May 4, 2026 · 1 citation · ▲ 154 on Hugging Face · Code ★ 17,051

– ReadersNo votes yet
9/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 9 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 0/5
86%Must read

Recursive Multi-Agent Systems

RecursiveMAS scales multi-agent collaboration through recursive latent-space computation via RecursiveLink and inner-outer loop co-optimization, improving accuracy by 8.3% with 1.2-2.4x speedup and 34.6%-75.6% token reduction over baselines.

Jiaru Zou, Rui Pan, Ruizhong Qiu, Pan Lu and 7 more

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published Apr 28, 2026 · 0 citations · ▲ 239 on Hugging Face · Code ★ 961

– ReadersNo votes yet
14/20 AI panelreviewers recommend it

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 14 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 2/5
76%Highly rated
?Highly ratedVote to see the score

Kimi K2.5: Visual Agentic Intelligence

Kimi K2.5 is an open-source multimodal agentic model using joint text-vision optimization and Agent Swarm to achieve state-of-the-art agentic, coding, vision, and reasoning results with up to 4.5x lower latency.

Kimi Team, Tongtong Bai, Yifan Bai, Yiping Bao and 36 more

Published Feb 2, 2026 · 2 citations · ▲ 277 on Hugging Face · Code ★ 2,313

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 5/5
medium 4/10
strict 1/5
71%Highly rated
?Highly ratedVote to see the score

PaperBanana: Automating Academic Illustration for AI Scientists

PaperBanana automates publication-ready academic illustrations via agentic VLM and image generation, outperforming baselines on a 292-case benchmark.

Dawei Zhu, Meng, Rui, Yale Song, Xiyu Wei and 3 more

Published Jan 30, 2026 · 1 citation · ▲ 230 on Hugging Face · Code ★ 7,132

– ReadersNo votes yet
7/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 7 of 20 reviewers recommend it
lenient 4/5
medium 3/10
strict 0/5
70%Highly rated
?Highly ratedVote to see the score

TradingAgents: Multi-Agents LLM Financial Trading Framework

TradingAgents proposes a multi-agent LLM framework with specialized trading roles and collaborative dynamics, outperforming baselines on cumulative returns, Sharpe ratio, and drawdown.

Xiao, Yijia, Edward W. Sun, Luo, Di, Wei Wang

Published Dec 28, 2024 · 6 citations · ▲ 150 on Hugging Face · Code ★ 110,044

– ReadersNo votes yet
5/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 5 of 20 reviewers recommend it
lenient 4/5
medium 1/10
strict 0/5
71%Highly rated
?Highly ratedVote to see the score

Very Large-Scale Multi-Agent Simulation in AgentScope

AgentScope introduces actor-based distributed infrastructure, flexible environments, and automated agent generation to enable scalable, efficient multi-agent simulations.

Xuchen Pan, Gao, Dawei, Yuexiang Xie, Chen, Yushuo and 5 more

Published Jul 25, 2024 · 1 citation · ▲ 47 on Hugging Face · Code ★ 32,834

– ReadersNo votes yet
6/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 6 of 20 reviewers recommend it
lenient 4/5
medium 2/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

SkillOrchestra: Learning to Route Agents via Skill Transfer

Jiayu Wang, Yifei Ming, Zixuan Ke, Shafiq Joty and 2 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
45%Niche pick
?Niche pickVote to see the score

Open-Ended Scientific Discovery and the Social Dynamics of Evolving Agent Networks

Tennison Liu, Silas Ruhrberg Estévez, Rob Davis, Ryan M Sheridan and 2 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Risk-Calibrated Context Selection for Healthcare Multi-Agent Handoff

Manu Agrawal, Sumit Mukherjee

Atlanta Poster Session 6, Fri, Dec 11, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Social Interaction Breaks Replicate Independence in Controlled-Clone LLM Agents

Yandan Zheng, Haoran Luo, Anh Tuan Luu

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 1/5
45%Niche pick
?Niche pickVote to see the score

Belief Engine: Configurable Stance Dynamics for Multi-Agent LLM Deliberation

Joshua C Yang, Maurice Flechtner, Damian Dailisan, Michiel Bakker

Paris Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

TritonTune: LLM-Guided Multi-Agent Optimization of GPU Kernel Configurations

Shukai Duan, Mihai Capotă, Guixiang Ma, Hou-Jen Ko and 4 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

UMAS: System-Level Uncertainty Quantification for Multi-Agent LLM Systems

Hanwen Li, Jinhao Duan, Xiaoshuang Shi, Yue Zhang and 3 more

Atlanta Poster Session 2, Wed, Dec 9, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
57%Worth a look
?Worth a lookVote to see the score

Qubrio: High-Performance Quantum Compilation via Multi-Agent LLM Collaboration

Jixuan Ruan, Zhuo Cui, Zhengding Hu, Zhongkai Yu and 9 more

Atlanta Poster Session 4, Thu, Dec 10, 4:30 PM–7:30 PM, Hall C1 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
Show 20 more papers