#rl (17 Repositories)
Ranked open-source repositories tagged with #rl, scored by pull request acceptance likelihood and maintainer engagement velocity.
28.0%
13.1h
17 repositories tagged #rl
tensorlakeai/tensorlake
Tensorlake is a serverless runtime for sandboxes and deploying background agentic applications
utilForever/baba-is-auto
Baba Is You simulator using C++ with some reinforcement learning
OpenPipe/ART
Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!
kubernetes-sigs/agent-sandbox
agent-sandbox enables easy management of isolated, stateful, singleton workloads, ideal for use cases like AI agent runtimes and reinforcement learning (RL).
hud-evals/hud-python
RL environments + evals for AI agents. Define once, train anything.
areal-project/AReaL
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
inclusionAI/AReno
An easy-to-use, fast toolkit to scale up RL post-training on a single node.
Goekdeniz-Guelmez/MLX-LoRA-Studio
A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.
NVlabs/alpagym
AlpaGym is a reinforcement-learning framework for end-to-end autonomous-driving policies.
jiangxinke/Harness-RL
Agentic RAG R1 Framework via Reinforcement Learning
LeCAR-Lab/model-based-diffusion
Official implementation for the paper "Model-based Diffusion for Trajectory Optimization". Model-based diffusion (MBD) is a novel diffusion-based trajectory optimization framework that employs a dynamics model to run the reverse denoising process to generate high-quality trajectories.
gbionics/amp-rsl-rl
🔁 AMP-RSL-RL: Adversarial Motion Priors for robotic RL (PPO + motion imitation)
pathak22/noreward-rl
[ICML 2017] TensorFlow code for Curiosity-driven Exploration for Deep Reinforcement Learning
Purewhiter/mobilegym
MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research · 浏览器里运行的安卓模拟器 · Browser-hosted Android Simulator · Verifiable Evaluation · Scalable Online RL Training (EMNLP 2026)
opennars/opennars
OpenNARS for Research 3.0+
DLR-RM/rl-baselines3-zoo
A training framework for Stable Baselines3 reinforcement learning agents, with hyperparameter optimization and pre-trained agents included.
tingaicompass/AI-Compass
“AI-Compass”将为社区指引在 AI 技术海洋中航行的方向,无论你是初学者还是进阶开发者,都能在这里找到通往 AI 各大方向的路径。旨在帮助开发者系统性地了解 AI 的核心概念、主流技术、前沿趋势,并通过实践掌握从理论到落地的全过程。