Good Papers

Showing papers from nuaa Show all papers

57%Worth a look
?Worth a lookVote to see the score

Choosing Before Acting: Comparative Value Estimation for Long-Horizon Tool-Use Agents

Yu Li, Zheng Zhang, Xin Liu, shengtian yang and 2 more

Sydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
1/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 1 of 20 reviewers recommend it
lenient 1/5
medium 0/10
strict 0/5
71%Highly rated
?Highly ratedVote to see the score

PhGPO: Pheromone-Guided Policy Optimization for Long-Horizon Tool Planning

PhGPO learns reusable tool-transition patterns from past trajectories via pheromone guidance to improve long-horizon tool planning.

Yu Li, Guangfeng Cai, shengtian yang, Han Luo and 4 more

Sydney Poster Session 6, Thu, Dec 10, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
6/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 6 of 20 reviewers recommend it
lenient 4/5
medium 2/10
strict 0/5
83%Must read
?Must readVote to see the score

AgentBrew: Offline Tool-Use Agent Learning from Raw Real-World Trajectories

AgentBrew learns tool-use policies offline from raw real-world trajectories via retrospective task inference and PMI-based credit assignment, improving Qwen3-32B by +8.7 accuracy over larger baselines.

Zhiyi Lyu, Yewen Li, Longtao Zheng, shengtian yang and 6 more

Sydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
13/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 13 of 20 reviewers recommend it
lenient 5/5
medium 7/10
strict 1/5