EasyPPO

EasyPPO is a straightforward and user-friendly implementation of Proximal Policy Optimization (PPO), offering integration with Weights & Biases (WandB) logging. This repository is designed for simplicity and serves as an ideal starting point for researchers new to PPO or those seeking a comprehensible implementation.

Overview

Originating from The Movement Lab at Stanford, EasyPPO was initially created for methods in character animation using mujoco. However, it has been successfully used for robotic arm control in other physics engines (pybullet). The only prerequisite is a compatible gym environment. Its primary features include:

A clean, minimalist PPO implementation suitable for both beginners and those seeking a simplified codebase.
Integrated WandB logging for effortless performance tracking and visualization during training.

Getting Started

To get started with this implementation, follow these steps:

Clone the repository
Install the required python version + dependencies
1. python=3.11.*
2. gymnasium=0.29.1
3. mujoco=2.3.7
4. pytorch
5. wandb
Run the example script to train the model (train.py)
Load and run the trained model (vis.py)

Example command: python train.py --config configs/humanoid-v5_cfg.cpy

Name		Name	Last commit message	Last commit date
Latest commit History 9 Commits
algs		algs
configs		configs
.gitignore		.gitignore
LICENSE		LICENSE
README.md		README.md
train.py		train.py
vis.py		vis.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

EasyPPO

Overview

Getting Started

About

Uh oh!

Releases

Packages

Uh oh!

Uh oh!

Contributors

Uh oh!

Languages

Folders and files

Latest commit

History

Repository files navigation

EasyPPO

Overview

Getting Started

About

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Packages 0

Uh oh!

Uh oh!

Contributors

Uh oh!

Languages

Packages