57%Worth a look?Worth a lookVote to see the scoreNeurIPS 2026AmazonAmazon RoboticsDukeOffline RLA$^2$IQL: Adaptive Asymmetric Implicit Q Learning for Automated Warehouse ConsolidationGuangyi Liu, Andrea Angiuli, Mirko Ristivojevic, Joseph W Durham and 2 moreAtlanta Poster Session 1, Wed, Dec 9, 10:00 AM–1:00 PM, Hall C1 · Published 2026– ReadersNo votes yet1/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 1 of 20 reviewers recommend itlenient 1/5medium 0/10strict 0/5
71%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026Imperial College LondonAmazon RoboticsBoston University / Broad InstitExplorationNon-Asymptotic Best Policy Identification Guarantees in Online Reinforcement LearningNavigate and Stop achieves first non-asymptotic best-policy identification guarantees in online tabular reinforcement learning, with sample complexity depending on MDP connectivity, characteristic-time curvature, and instance-dependent quantities.Joseph Lazzaro, Alessio Russo, Aldo PacchianoSydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet6/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 6 of 20 reviewers recommend itlenient 2/5medium 3/10strict 1/5