Generalized Advantage Estimation (GAE) calculator computing temporal difference advantages and returns
-
Updated
Sep 30, 2026 - Python
Generalized Advantage Estimation (GAE) calculator computing temporal difference advantages and returns
Generalized Advantage Estimation (GAE) calculator computing temporal difference advantages and returns
Exact-oracle conformance tests for turn-level credit assignment in agentic RL (GRPO, RLOO, GAE, GiGPO; verl, TRL, OpenRLHF).
Tensorflow implementation of Proximal Policy Optimization (Reinforcement Learning) and its common optimizations. Features Tensorboard integration and lots of sample runs on custom, classical and robotics oriented environments.
To associate your repository with the advantage-estimation topic, visit your repo's landing page and select "manage topics."