ppo

39 Projekte teilen dieses GitHub-Topic

ppo — easy-rl ★14.5kppocleanrl — ★10.2kReinforcement-learning-with-tensorflow — ★9.5kdeep-reinforcement-learning — ★5.2kReinforcement-Learning — ★4.7kDeep-reinforcement-learning-with-pytorch — ★4.6kElegantRL — ★4.4kpytorch-a2c-ppo-acktr-gail — ★3.9kDeep_reinforcement_learning_Course — ★3.9kDRL-Pytorch — ★3.4kminimalRL — ★3.2kPPO-PyTorch — ★2.4kSLM-Lab — ★1.4kSuper-mario-bros-PPO-pytorch — ★1.3kPyTorch-RL — ★1.3kPPO-for-Beginners — ★1.3kLearningHumanoidWalking — ★1.2kpurejaxrl — ★1.1kDeepRL-Tutorials — ★1.1kxuance — ★1.1kHands-On-Reinforcement-Learning-With-Python — ★860autonomous-learning-library — ★655pytorch-DRL — ★617DeepRL-TensorFlow2 — ★600Open-AgentRL — ★595Awesome-RL-for-Video-Generation — ★574Autonomous-Driving-in-Carla-using-Deep-Reinforcement-Learning — ★562Deep-Reinforcement-Learning-With-Python — ★475rllte — ★470rl-collision-avoidance — ★457LLM-RLHF-Tuning — ★453machin — ★418rad — ★415Reinforcement_Learning — ★412lagom — ★378evorl — ★343rlgraph — ★323TF2-RL — ★316RL_Matrix — ★310cleanrl★ 10.2kReinforcement-learning-w…★ 9.5kdeep-reinforcement-learn…★ 5.2kReinforcement-Learning★ 4.7kDeep-reinforcement-learn…★ 4.6kElegantRL★ 4.4kpytorch-a2c-ppo-acktr-ga…★ 3.9kDeep_reinforcement_learn…★ 3.9kDRL-Pytorch★ 3.4kminimalRL★ 3.2kPPO-PyTorch★ 2.4kSLM-Lab★ 1.4kSuper-mario-bros-PPO-pyt…★ 1.3kPyTorch-RL★ 1.3kPPO-for-Beginners★ 1.3kLearningHumanoidWalking★ 1.2kpurejaxrl★ 1.1kDeepRL-Tutorials★ 1.1kxuance★ 1.1kHands-On-Reinforcement-L…★ 860autonomous-learning-libr…★ 655pytorch-DRL★ 617DeepRL-TensorFlow2★ 600Open-AgentRL★ 595Awesome-RL-for-Video-Gen…★ 574Autonomous-Driving-in-Ca…★ 562Deep-Reinforcement-Learn…★ 475rllte★ 470rl-collision-avoidance★ 457LLM-RLHF-Tuning★ 453machin★ 418rad★ 415Reinforcement_Learning★ 412lagom★ 378evorl★ 343rlgraph★ 323TF2-RL★ 316RL_Matrix★ 310

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
easy-rl
强化学习中文教程(蘑菇书🍄),在线阅读地址:https://datawhalechina.github.io/easy-rl/
★ 14.5k
cleanrl
High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly…
★ 10.2k
Reinforcement-learning-with-tensorflow
Simple Reinforcement learning tutorials, 莫烦Python 中文AI教学
★ 9.5k
deep-reinforcement-learning
Repo for the Deep Reinforcement Learning Nanodegree program
★ 5.2k
Reinforcement-Learning
Learn Deep Reinforcement Learning in 60 days! Lectures & Code in Python. Reinforcement Learning + Deep…
★ 4.7k
Deep-reinforcement-learning-with-pytorch
PyTorch implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....
★ 4.6k
ElegantRL
Massively Parallel Deep Reinforcement Learning. 🔥
★ 4.4k
pytorch-a2c-ppo-acktr-gail
PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable…
★ 3.9k
Deep_reinforcement_learning_Course
Implementations from the free course Deep Reinforcement Learning with Tensorflow and PyTorch
★ 3.9k
DRL-Pytorch
Clean, Robust, and Unified PyTorch implementation of popular Deep Reinforcement Learning (DRL) algorithms…
★ 3.4k
minimalRL
Implementations of basic RL algorithms with minimal lines of codes! (pytorch based)
★ 3.2k
PPO-PyTorch
Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch
★ 2.4k
SLM-Lab
Modular Deep Reinforcement Learning framework in PyTorch. Companion library of the book "Foundations of Deep…
★ 1.4k
Super-mario-bros-PPO-pytorch
Proximal Policy Optimization (PPO) algorithm for Super Mario Bros
★ 1.3k
PyTorch-RL
PyTorch implementation of Deep Reinforcement Learning: Policy Gradient methods (TRPO, PPO, A2C) and…
★ 1.3k
PPO-for-Beginners
A simple and well styled PPO implementation. Based on my Medium series:…
★ 1.3k
LearningHumanoidWalking
Training a humanoid robot for locomotion using Reinforcement Learning
★ 1.2k
purejaxrl
Really Fast End-to-End Jax RL Implementations
★ 1.1k
DeepRL-Tutorials
Contains high quality implementations of Deep Reinforcement Learning algorithms written in PyTorch
★ 1.1k
xuance
XuanCe: A Comprehensive and Unified Deep Reinforcement Learning Library
★ 1.1k
Hands-On-Reinforcement-Learning-With-Python
Master Reinforcement and Deep Reinforcement Learning using OpenAI Gym and TensorFlow
★ 860
autonomous-learning-library
A PyTorch library for building deep reinforcement learning agents.
★ 655
pytorch-DRL
PyTorch implementations of various Deep Reinforcement Learning (DRL) algorithms for both single agent and…
★ 617
DeepRL-TensorFlow2
🐋 Simple implementations of various popular Deep Reinforcement Learning algorithms using TensorFlow2
★ 600
Open-AgentRL
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
★ 595
Awesome-RL-for-Video-Generation
A curated list of papers on reinforcement learning for video generation
★ 574
Autonomous-Driving-in-Carla-using-Deep-Reinforcement-Learning
Deep Reinforcement Learning (PPO) in Autonomous Driving (Carla) [from scratch]
★ 562
Deep-Reinforcement-Learning-With-Python
Master classic RL, deep RL, distributional RL, inverse RL, and more using OpenAI Gym and TensorFlow with…
★ 475
rllte
Long-Term Evolution Project of Reinforcement Learning
★ 470
rl-collision-avoidance
Implementation of the paper "Towards Optimally Decentralized Multi-Robot Collision Avoidance via Deep…
★ 457
LLM-RLHF-Tuning
LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)
★ 453
machin
Reinforcement learning library(framework) designed for PyTorch, implements DQN, DDPG, A2C, PPO, SAC, MADDPG,…
★ 418
rad
RAD: Reinforcement Learning with Augmented Data
★ 415
Reinforcement_Learning
Reinforcement learning tutorials
★ 412
lagom
lagom: A PyTorch infrastructure for rapid prototyping of reinforcement learning algorithms.
★ 378
evorl
EvoRL is a fully GPU-accelerated framework for Evolutionary Reinforcement Learning, implemented with JAX. It…
★ 343
rlgraph
RLgraph: Modular computation graphs for deep reinforcement learning
★ 323
TF2-RL
Reinforcement learning algorithms implemented for Tensorflow 2.0+ [DQN, DDPG, AE-DDPG, SAC, PPO, Primal-Dual…
★ 316
RL_Matrix
Deep Reinforcement Learning in C#
★ 310
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.