Back to @chatgpt
ChatGPT
ChatGPT
@chatgpt

Scaling laws for reward model overoptimization

Read on openai.com

7:00 AM · Oct 19, 2022

More from ChatGPT