PyTorch implementations of algorithms from "Reinforcement Learning: An Introduction by Sutton and Barto", along with various RL research papers.
-
Updated
Aug 14, 2025 - Python
PyTorch implementations of algorithms from "Reinforcement Learning: An Introduction by Sutton and Barto", along with various RL research papers.
Lightweight, GPU-accelerated reinforcement learning for Isaac Lab, mjlab, and Gymnasium.
A C# Unity mod connected through a named pipe with Python for training a Reinforcement Learning agent to fight Hollow Knight Hornet Protector
End-to-end RL trading framework with PPO agent, self-attention neural network, custom Gym environment, and advanced backtesting.
A Complete Collection of Deep RL Famous Algorithms implemented in Gymnasium most Popular environments
🚦 Next-generation AI Traffic Management System with real-time computer vision, reinforcement learning optimization, emergency vehicle detection, and immersive 3D visualization
This repository is dedicated to implementing algorithms "From Scratch". It goes beyond simple API calls, diving deep into the underlying logic of everything from basic training to cutting-edge techniques like DeepSeek-R1.
A legged locomotion project
This is a project for PPO S&P 500 trading
An exploratory 2-week project into Reinforcement Learning and the PPO algorithm
Reinforcement learning–based controller for balancing an inverted pendulum using Proximal Policy Optimization (PPO). Supports configurable mass, length, and gravity settings (Earth, lunar, microgravity) with automated training logs, reward visualization, and performance analysis.
This repository implements a Proximal Policy Optimization (PPO) agent that learns to play Super Mario Bros using TensorFlow/Keras and OpenAI Gym. Features CNNs for vision, Actor-Critic architecture, and parallel environments. Train your own Mario master or run a pre-trained one!
Autonomous driving system using PPO-based Reinforcement Learning and CARLA Simulator for lane following and navigation.
RL agent that packs 3d boxes into a container
2D orbital rocket sim with PPO in PyTorch. Models thrust, drag, gravity, fuel; agent learns efficient ascent. Includes telemetry & visualization
What is Proximal Policy Optimization (PPO) leaving behind : a study of PPO's exploitation mechanism.
Smart Pricing for NANCY using Multi-Agent Reinforcement Learning and Reverse Auction Theory. Smart_Pricing_MARL_NANCY is an open-source, EU-co-funded Smart Pricing Module (SPM) developed for the NANCY project. It leverages Multi-Agent Reinforcement Learning (MARL) and Reverse Auction Theory to calculate optimal pricing strategies.
Autonomous Microgrid Balancer using PPO RL with Adversarial Training for resilience under High-Impact Low-Probability (HILP) disturbances.
This Legal Document Analyzer is a proof-of-concept NLP project demonstrating the potential of transformers for legal document summarization.
AI-powered production line optimization using reinforcement learning (PPO).
To associate your repository with the ppo-algorithm topic, visit your repo's landing page and select "manage topics."