Good Papers

Showing papers from Beihang University Show all papers

45%Niche pick
?Niche pickVote to see the score

On the Bias of Group-Based Advantage Estimation

Fengkai Yang, Zherui Chen, Xiaohan Wang, Xiaodong Lu and 8 more

Paris Poster Session 4, Thu, Dec 10, 5:30 PM–7:30 PM, Paris Poster Hall · Published 2026

– ReadersNo votes yet
0/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 0 of 20 reviewers recommend it
lenient 0/5
medium 0/10
strict 0/5
76%Highly rated
?Highly ratedVote to see the score

Heterogeneous Agent Collaborative Reinforcement Learning

HACRL enables heterogeneous agents to share verified rollouts during collaborative on-policy training and execute independently at inference, with HACPO improving all agents by 3.6% over baselines at half the rollout cost.

Zhixia Zhang, Zixuan Huang, Gonxun Li, Huaiyang Wang and 8 more

Paris Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026 · ▲ 110 on Hugging Face

– ReadersNo votes yet
10/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 10 of 20 reviewers recommend it
lenient 3/5
medium 7/10
strict 0/5