ChatGPT@chatgptEquivalence between policy gradients and soft Q-learningRead on openai.com7:00 AM · Apr 21, 2017