Good Papers

Showing papers from Google DeepMind & Mila Show all papers

72%Highly rated
?Highly ratedVote to see the score

From Static Policies to Adaptive Priors in Offline Reinforcement Learning

Offline RL should prioritize adaptive policy priors preserving improvement capacity during online updates rather than static conservative deployment.

Tianwei Ni, Vineet Jain, Akash Karthikeyan, Pierre-Luc Bacon

Sydney Poster Session 2, Tue, Dec 8, 5:00 PM–8:00 PM, Hall 1-4 · Published 2026

– ReadersNo votes yet
8/20 AI panelreviewers recommend it

Readers and the AI panel: vote on this paper to see what they said.

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel: 8 of 20 reviewers recommend it
lenient 4/5
medium 4/10
strict 0/5