Official implementation of the AAAI 2021 paper Deep Bayesian Quadrature Policy Optimization.
-
Updated
Feb 17, 2021 - Python
Official implementation of the AAAI 2021 paper Deep Bayesian Quadrature Policy Optimization.
Benchmarking the Natural Gradient in Policy Gradient Methods and Evolution Strategies
Model-free policy gradient algorithm for LQR
My solutions to the labs from this bootcamp:
Deep Reinforcement Learning from mathematical foundations to PyTorch implementations: rigorous proofs, step-by-step derivations of objectives and gradient estimators, and reproducible experiments on Gymnasium and MuJoCo benchmarks spanning policy gradients, actor-critic methods, value-based learning, and continuous control.
To associate your repository with the natural-policy-gradient topic, visit your repo's landing page and select "manage topics."