45%Niche pick?Niche pickVote to see the scoreNeurIPS 2026PekingU California, BerkeleyMeituanBeihangInstitute of automation, ChineseFairness & biasOn the Bias of Group-Based Advantage EstimationFengkai Yang, Zherui Chen, Xiaohan Wang, Xiaodong Lu and 8 moreParis Poster Session 4, Thu, Dec 10, 5:30 PM–7:30 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet0/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 0 of 20 reviewers recommend itlenient 0/5medium 0/10strict 0/5
76%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026BeihangBeijing University of AeronauticSchool of Computer Science and EAppleBytedanceMulti-agent RLHeterogeneous Agent Collaborative Reinforcement LearningHACRL enables heterogeneous agents to share verified rollouts during collaborative on-policy training and execute independently at inference, with HACPO improving all agents by 3.6% over baselines at half the rollout cost.Zhixia Zhang, Zixuan Huang, Gonxun Li, Huaiyang Wang and 8 moreParis Poster Session 5, Fri, Dec 11, 11:30 AM–1:30 PM, Paris Poster Hall · Published 2026 · ▲ 110 on Hugging Face– ReadersNo votes yet10/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 10 of 20 reviewers recommend itlenient 3/5medium 7/10strict 0/5