Logo
Explore Help
Register Sign In
wassname/ray
Watch 1
Star 0
Fork 0
mirror of https://github.com/wassname/ray.git synced 2026-08-05 13:21:03 +08:00
Code Issues Packages Projects Releases Wiki Activity
Files
2e41a29c8fe3405fee1e56d9bc82b54ede0e4b62
ray/rllib/agents/pg
T
History
Sumanth RatnaandGitHub 9da7bdcc8e Use master for links to docs in source (#10866)
2020-09-19 00:30:45 -07:00
..
tests
ci: Redo format.sh --all script & backfill lint fixes (#9956)
2020-08-07 16:49:49 -07:00
__init__.py
[RLlib] Examples folder restructuring (models) part 1 (#8353)
2020-05-08 08:20:18 +02:00
pg_tf_policy.py
[RLlib] PPO, APPO, and DD-PPO code cleanup. (#10420)
2020-09-02 14:03:01 +02:00
pg_torch_policy.py
[RLlib] First attempt at cleaning up algo code in RLlib: PG. (#10115)
2020-08-20 17:05:57 +02:00
pg.py
Use master for links to docs in source (#10866)
2020-09-19 00:30:45 -07:00
README.md
[RLlib] First attempt at cleaning up algo code in RLlib: PG. (#10115)
2020-08-20 17:05:57 +02:00
utils.py
[RLlib] First attempt at cleaning up algo code in RLlib: PG. (#10115)
2020-08-20 17:05:57 +02:00

README.md

Policy Gradient (PG)

An implementation of a vanilla policy gradient algorithm for TensorFlow and PyTorch.

Detailed Documentation

Implementation

Reference in New Issue View Git Blame Copy Permalink
Powered by Gitea Version: 1.27.1 Page: 50ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API