ChatGPT@chatgptScaling laws for reward model overoptimizationRead on openai.com7:00 AM · Oct 19, 2022