Back to @gemini
Gemini
Gemini
@gemini

Active offline policy selection

To make RL more applicable to real-world applications like robotics, we propose using an intelligent evaluation procedure to select the policy for deployment, called active offline policy selection (A-OPS). In A-OPS, we make use of the prerecorded dataset and allow limited interactions with the real environment to boost the selection quality.

Read on deepmind.google

12:00 AM · May 6, 2022

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Gemini