#reinforcement-learning (30 Repositories)
Ranked open-source repositories tagged with #reinforcement-learning, scored by pull request acceptance likelihood and maintainer engagement velocity.
72.7%
37.1h
30 repositories tagged #reinforcement-learning
upkie/upkie
Open-source wheeled biped robots
utilForever/baba-is-auto
Baba Is You simulator using C++ with some reinforcement learning
unilabsim/UniLab
UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms
enactic/openarm
A fully open-source humanoid arm for physical AI research and deployment in contact-rich environments.
hsahovic/poke-env
Poke-env: Python Interface for Pokemon Showdown Bots
bayesianbandits/bayesianbandits
A Pythonic microframework for multi-armed bandit problems
ComputationalPsychiatry/pyhgf
PyHGF: A neural network library for predictive coding
inclusionAI/Awex
A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from training to inference in RL workflows
NVlabs/Sana
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
verl-project/verl-omni
Multimodal RL training framework for diffusion & omni models
OpenPipe/ART
Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!
hanruihua/ir-sim
A Python-based lightweight robot simulator designed for navigation, control, and learning
RobotControlStack/robot-control-stack
A lean, ROS-free sim-to-real framework for training and deploying Vision-Language-Action (VLA) models and RL agents. Native MuJoCo Gymnasium wrappers with synchronous execution for Franka, UR5e, xArm, SO101 and YAM.
kvcache-ai/Mooncake
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
RLinf/RLinf
RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
X-GenGroup/Flow-Factory
A unified framework for easy reinforcement learning in Flow-Matching models
Tencent-Hunyuan/UniRL
UniRL is a Framework for Unified Multimodal Model Reinforcement Learning
nissymori/mahjax
A GPU-Accelerated Mahjong Simulator for RL in JAX
hud-evals/hud-python
RL environments + evals for AI agents. Define once, train anything.
ray-project/ray
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
areal-project/AReaL
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
google-deepmind/open_spiel
OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games.
pytorch/rl
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
NVIDIA-NeMo/Gym
Evaluate and improve models and agents using environments
allenai/ScienceWorld
ScienceWorld is a text-based virtual environment centered around accomplishing tasks from the standardized elementary science curriculum.
wandb/wandb
The AI developer platform. Use Weights & Biases to train and fine-tune models, and manage models from experimentation to production.
InternLM/xtuner
A Next-Generation Training Engine Built for Ultra-Large MoE Models
inclusionAI/AReno
An easy-to-use, fast toolkit to scale up RL post-training on a single node.
NVlabs/ProtoMotions
ProtoMotions is a GPU-accelerated simulation and learning framework for training physically simulated digital humans and humanoid robots.