All tags
Catalogue tag

#reinforcement-learning

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

18 records
Tagged “reinforcement-learning”Ranked by health index
PyPI · Go · crates.io
94Excellenthealth index
wandb/wandb
The AI developer platform. Use Weights & Biases to train and fine-tune models, and manage models from experimentation to production.
Python · Go★ 11.2K↓ 23M/moJul 20, 2026
MITJul 20, 2026 · metrics 1.13.0
Maven · PyPI
90Excellenthealth index
ray-project/ray
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
Python · C++★ 43.3KJul 17, 2026
Apache-2.0Jul 17, 2026 · metrics 1.13.0
PyPI
85Excellenthealth index
Farama-Foundation/Gymnasium
A standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)
Python★ 12.2K↓ 10.1M/moJul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
PyPI
85Excellenthealth index
google-deepmind/optax
Optax is a gradient processing and optimization library for JAX.
Python★ 2,301Jul 18, 2026
Apache-2.0Jul 18, 2026 · metrics 1.13.0
crates.io · PyPI
82Goodhealth index
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
Python★ 30.3K↓ 274M/moJul 14, 2026
Apache-2.0Jul 14, 2026 · metrics 1.13.0
PyPI · npm
82Goodhealth index
unslothai/unsloth
Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.
Python · TypeScript★ 68.5K↓ 2.3M/moJul 20, 2026
Apache-2.0Jul 20, 2026 · metrics 1.13.0
PyPI · npm
79Goodhealth index
NVIDIA-NeMo/Gym
Evaluate and improve models and agents using environments
Python · MDX★ 1,055↓ 406.4K/moJul 18, 2026
Apache-2.0Jul 18, 2026 · metrics 1.13.0
PyPI
79Goodhealth index
pytorch/rl
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
Python★ 3,491Jul 15, 2026
MITJul 15, 2026 · metrics 1.13.0
PyPI
78Goodhealth index
AgileRL/AgileRL
Streamlining reinforcement learning with RLOps. State-of-the-art RL algorithms and tools, with 10x faster training through evolutionary hyperparameter optimization.
Python★ 938↓ 1,021/moJul 20, 2026
Apache-2.0Jul 20, 2026 · metrics 1.13.0
PyPI
78Goodhealth index
mujocolab/mjlab
Isaac Lab API, powered by MuJoCo-Warp, for RL and robotics research
Python★ 2,709Jul 22, 2026
Apache-2.0Jul 22, 2026 · metrics 1.13.0
PyPI
77Goodhealth index
munich-quantum-toolkit/predictor
MQT Predictor - A Tool for Automatic Device Selection with Device-Specific Circuit Compilation for Quantum Computing
Python★ 86↓ 285/moJul 21, 2026
MITJul 21, 2026 · metrics 1.13.0
npm · PyPI
76Goodhealth index
trycua/cua
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
HTML · Python · Rust★ 20.5K↓ 153/moJul 22, 2026
MITJul 22, 2026 · metrics 1.13.0
npm
75Goodhealth index
enactic/openarm
A fully open-source humanoid arm for physical AI research and deployment in contact-rich environments.
MDX · TypeScript★ 2,718↓ 0/moJul 14, 2026
Apache-2.0Jul 14, 2026 · metrics 1.13.0
PyPI
71Goodhealth index
ttktjmt/mjswan
MuJoco Simulation on Web Assembly with Neural netwroks
Python · TypeScript★ 309Jul 15, 2026
Apache-2.0Jul 15, 2026 · metrics 1.13.0
PyPI · crates.io
55Moderatehealth index
tsilva/SuperMarioBros-Nes-turbo
🚀 Blazing fast SuperMarioBros-Nes environment for RL research 🍄
Python · Rust★ 0↓ 2,748/moJul 16, 2026
MITJul 16, 2026 · metrics 1.13.0
npm
49At riskhealth index
ruvnet/agentdb
Vector memory that gets smarter every time your agent uses it.
TypeScript · HTML · JavaScript★ 80↓ 649.1K/moJul 23, 2026
MITJul 23, 2026 · metrics 1.13.0
PyPI
39At riskhealth index
waybarrios/crystal
CRYSTAL: Beyond Final Answers: Benchmark for Transparent Multimodal Reasoning Evaluation | arXiv 2603.13099
Python★ 2Jul 17, 2026
No licenseJul 17, 2026 · metrics 1.13.0