45%Niche pick?Niche pickVote to see the scoreNeurIPS 2026BostonBoston University / Broad InstitDeep RLLearning to Undo: Transfer Reinforcement Learning under State Space TransformationsMridul Mahajan, Aldo Pacchiano, Xuezhou ZhangParis Poster Session 2, Wed, Dec 9, 5:00 PM–7:00 PM, Paris Poster Hall · Published 2026– ReadersNo votes yet0/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 0 of 20 reviewers recommend itlenient 0/5medium 0/10strict 0/5
71%Highly rated?Highly ratedVote to see the scoreNeurIPS 2026Imperial College LondonAmazon RoboticsBoston University / Broad InstitExplorationNon-Asymptotic Best Policy Identification Guarantees in Online Reinforcement LearningNavigate and Stop achieves first non-asymptotic best-policy identification guarantees in online tabular reinforcement learning, with sample complexity depending on MDP connectivity, characteristic-time curvature, and instance-dependent quantities.Joseph Lazzaro, Alessio Russo, Aldo PacchianoSydney Poster Session 3, Wed, Dec 9, 10:00 AM–1:00 PM, Hall 1-4 · Published 2026– ReadersNo votes yet6/20 AI panelreviewers recommend itReaders and the AI panel: vote on this paper to see what they said.Worth readingNot for meOnly vote on papers you've read. Sign in with GitHub to vote.AI panel: 6 of 20 reviewers recommend itlenient 2/5medium 3/10strict 1/5