Home Projects trpo
trpo

trpo

by pat-coady · GitHub

Trust Region Policy Optimization with TensorFlow and OpenAI Gym

# machine-learning# mujoco# policy-gradient# reinforcement-learning
View on GitHub
⭐ Stars
363
🍴 Forks
106
📜 License
MIT
Commercial use OK
📅 Created
2017
🔄 Last commit
6 yr ago
🏷️ Category
machine-learning
💻 Language
You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
trpo — GitHub preview card
📈 Star history
364363
2026-07-202026-07-21
📄 About

Trust Region Policy Optimization with TensorFlow and OpenAI Gym

Frequently asked questions

What is trpo?

Trust Region Policy Optimization with TensorFlow and OpenAI Gym

Is trpo open source?

trpo is an open-source project. It is released under the MIT license.

Is trpo free?

Yes. trpo is free and open source — you can use, modify and self-host it.

🏅 Maintainer of this project?
OpenSourceAI badge — trpo

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![OpenSourceAI](https://opensourceai.tech/badge.php?tool=pat-coady-trpo)](https://opensourceai.tech/project/pat-coady-trpo.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.