Todas las etiquetas
Etiqueta del catálogo

#rl

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

10 registros
Con la etiqueta «rl»Ordenado por índice de salud
PyPI
96Excepcionalíndice de salud
Farama-Foundation/Gymnasium
A standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)
Python★ 12.4K↓ 5.7M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
PyPI
94Excepcionalíndice de salud
PrimeIntellect-ai/verifiers
Our library for RL environments + evals
Python★ 4565↓ 378.1K/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
PyPI · npm · crates.io
94Excepcionalíndice de salud
tensorlakeai/tensorlake
Tensorlake is a serverless runtime for sandboxes and deploying background agentic applications
Python · Rust · TypeScript★ 976↓ 153K/mes31 jul 2026
Apache-2.031 jul 2026 · métricas 2.10.0
PyPI · npm
93Excepcionalíndice de salud
NVIDIA-NeMo/Gym
Evaluate and improve models and agents using environments
Python · MDX★ 1055↓ 406.4K/mes18 jul 2026
Apache-2.018 jul 2026 · métricas 2.10.0
PyPI
92Excelenteíndice de salud
JudgmentLabs/judgeval
The Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.
Python★ 1057↓ 233.1K/mes13 ago 2026
Apache-2.013 ago 2026 · métricas 2.10.0
PyPI
91Excelenteíndice de salud
pytorch/rl
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
Python★ 3548↓ 84.8K/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
PyPI
90Excelenteíndice de salud
Farama-Foundation/PettingZoo
A standard API for multi-agent reinforcement learning environments, with popular reference environments and related utilities
Python★ 348913 ago 2026
MIT13 ago 2026 · métricas 2.10.0
PyPI
75Buenoíndice de salud
freesolo-co/flash
LoRA post-training for open-weight models: SFT, GRPO, and on-policy distillation. Describe a run in TOML; Flash allocates a GPU, trains, streams checkpoints, and serves the adapter.
Python★ 2↓ 5834/mes29 ago 2026
Apache-2.029 ago 2026 · métricas 2.10.0
crates.io
60Moderadoíndice de salud
rust-control/rusty_mujoco
A safe, low-level Rust binding for MuJoCo physics simulator
Rust★ 9↓ 2162/mes17 jul 2026
Apache-2.017 jul 2026 · métricas 2.10.0
npm
50Moderadoíndice de salud
ruvnet/agentdb
Vector memory that gets smarter every time your agent uses it.
TypeScript · HTML · JavaScript★ 80↓ 649.1K/mes23 jul 2026
MIT23 jul 2026 · métricas 2.10.0