Back to @gemini
Gemini
Gemini
@gemini

Generally capable agents emerge from open-ended play

In recent years, artificial intelligence agents have succeeded in a range of complex game environments. For instance, AlphaZero beat world-champion programs in chess, shogi, and Go after starting out with knowing no more than the basic rules of how to play. Through reinforcement learning (RL), this single system learnt by playing round after round of games through a repetitive process of trial and error. But AlphaZero still trained separately on each game — unable to simply learn another game or task without repeating the RL process from scratch. The same is true for other successes of RL, suc

Read on deepmind.google

12:00 AM · Jul 27, 2021

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Gemini