57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026U StellenboschInstaDeepU Cape Town & InstaDeepInstaDeep LtdInstadeep LtdMulti-agent RLCoordinating Hundreds of RL Agents through Scalable Inference-Time SearchDaniel Rajaonarivonivelomanantsoa, Oussama Hidaoui, Refiloe Shabe, Noah De Nicola and 12 moreSydney Poster Session 1, Tue, Dec 8, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026Cape Institute for Safe AIU Cape Town & InstaDeepDeepMindPrivacyA Generative Model of Contextual Integrity: Appropriate vs. Inappropriate Information SharingOmer Ebead, Juan Formanek, Joel LeiboSydney Poster Session 5, Thu, Dec 10, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
80%Must read?Must readVote to see the scoreNeurIPS 2026InstaDeepInstaDeep LtdU StellenboschInstadeep LtdU Cape Town & InstaDeepDeep RLSelf-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy OptimisationCPPO is an on-policy contrastive RL method deriving advantages from contrastive Q-values via PPO without rewards or replay buffers, outperforming prior CRL baselines in 14 of 18 tasks and matching or exceeding hand-crafted-reward PPO in 12 of 18.Asim Osman, Sasha Abramowitz, Mark Bergh, Ulrich Armel Mbou Sob and 12 moreSydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet12/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 12 of 20 reviewers recommend itlenient 4/5medium 5/10strict 3/5