ChatGPT@chatgptVariance reduction for policy gradient with action-dependent factorized baselinesRead on openai.com7:00 AM · Mar 20, 2018