default search action
"Posterior Value Functions: Hindsight Baselines for Policy Gradient Methods."
Chris Nota, Philip S. Thomas, Bruno C. da Silva (2021)
- Chris Nota, Philip S. Thomas, Bruno C. da Silva:
Posterior Value Functions: Hindsight Baselines for Policy Gradient Methods. ICML 2021: 8238-8247
manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.