Maven · PyPI100Винятковийіндекс здоров'я

ray-project/rayRay is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
Python · C++★ 43.4K5 серп. 2026 р.
PyPI · Go · crates.io100Винятковийіндекс здоров'я

wandb/wandbThe AI developer platform. Use Weights & Biases to train and fine-tune models, and manage models from experimentation to production.
Python · Go★ 11.2K↓ 26.6M/міс27 серп. 2026 р.
PyPI98Винятковийіндекс здоров'я

kvcache-ai/MooncakeMooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
C++ · Python★ 6 25612 серп. 2026 р.
PyPI · crates.io98Винятковийіндекс здоров'я

sgl-project/sglangSGLang is a high-performance serving framework for large language models and multimodal models.
Python★ 31.3K↓ 144M/міс4 серп. 2026 р.
PyPI97Винятковийіндекс здоров'я
Python★ 2 30118 лип. 2026 р.
PyPI96Винятковийіндекс здоров'я

Farama-Foundation/GymnasiumA standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)
Python★ 12.4K↓ 5.7M/міс27 серп. 2026 р.
PyPI95Винятковийіндекс здоров'я

google-deepmind/open_spielOpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games.
C++ · Python★ 5 44028 серп. 2026 р.
PyPI94Винятковийіндекс здоров'я
Python★ 4 565↓ 378.1K/міс28 серп. 2026 р.
PyPI · npm94Винятковийіндекс здоров'я

unslothai/unslothUnsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek-V4, GLM and other models.
Python · TypeScript★ 69.6K4 серп. 2026 р.
PyPI93Винятковийіндекс здоров'я
AgileRL/AgileRLStreamlining reinforcement learning with RLOps. State-of-the-art RL algorithms and tools, with 10x faster training through evolutionary hyperparameter optimization.
Python★ 938↓ 1 021/міс20 лип. 2026 р.
PyPI · npm93Винятковийіндекс здоров'я
Python · MDX★ 1 055↓ 406.4K/міс18 лип. 2026 р.
PyPI93Винятковийіндекс здоров'я
mujocolab/mjlabIsaac Lab API, powered by MuJoCo-Warp, for RL and robotics research
Python★ 2 70922 лип. 2026 р.
npm93Винятковийіндекс здоров'я

proffesor-for-testing/agentic-qeAgentic QE Fleet is an open-source AI-powered QA/QE platform designed for use with Coding Agents (works best with Claude Code) featuring specialized agents and skills to support testing activities for a product at any stage of the SDLC. Free to use, fork, build, and contribute. Based on the Agentic QE Framework created by Dragan Spiridonov.
TypeScript★ 474↓ 56.4K/міс5 вер. 2026 р.
PyPI · npm93Винятковийіндекс здоров'я

trycua/cuaScale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
HTML · Rust · Python★ 20.9K↓ 26.8K/міс5 серп. 2026 р.
PyPI92Відміннийіндекс здоров'я

JudgmentLabs/judgevalThe Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.
Python★ 1 057↓ 233.1K/міс13 серп. 2026 р.
PyPI92Відміннийіндекс здоров'я
Python★ 86↓ 285/міс21 лип. 2026 р.
PyPI91Відміннийіндекс здоров'я

DLR-RM/stable-baselines3PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.
Python★ 13.7K12 серп. 2026 р.
PyPI91Відміннийіндекс здоров'я
C++ · IDL★ 2 44113 серп. 2026 р.
PyPI91Відміннийіндекс здоров'я

pytorch/rlA modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
Python★ 3 548↓ 84.8K/міс5 вер. 2026 р.
PyPI90Відміннийіндекс здоров'я

Farama-Foundation/PettingZooA standard API for multi-agent reinforcement learning environments, with popular reference environments and related utilities
Python★ 3 48913 серп. 2026 р.
npm90Відміннийіндекс здоров'я

enactic/openarmA fully open-source humanoid arm for physical AI research and deployment in contact-rich environments.
MDX · TypeScript★ 2 9245 вер. 2026 р.
PyPI86Відміннийіндекс здоров'я

google-deepmind/dm_controlGoogle DeepMind's software stack for physics-based simulation and Reinforcement Learning environments, using MuJoCo.
Python★ 4 66312 серп. 2026 р.
PyPI84Відміннийіндекс здоров'я
Python · TypeScript★ 30915 лип. 2026 р.
PyPI81Відміннийіндекс здоров'я

desy-ml/cheetahFast and differentiable particle accelerator optics simulation for reinforcement learning and optimisation applications.
Python · Jupyter Notebook★ 69↓ 2 065/міс16 серп. 2026 р.
PyPI80Відміннийіндекс здоров'я

ugr-sail/sinergymGym environment for building simulation and control using reinforcement learning
Python★ 23415 серп. 2026 р.
PyPI75Добрийіндекс здоров'я

freesolo-co/flashLoRA post-training for open-weight models: SFT, GRPO, and on-policy distillation. Describe a run in TOML; Flash allocates a GPU, trains, streams checkpoints, and serves the adapter.
Python★ 2↓ 5 834/міс29 серп. 2026 р.
PyPI73Добрийіндекс здоров'я

miskibin/py-draughtsFastest Python draughts/checkers library — bitboards, 8 variants, alpha-beta engine, web UI. ~200x faster than pydraughts.
Python★ 16↓ 673/міс17 серп. 2026 р.
PyPI71Добрийіндекс здоров'я

araffin/sbxSBX: Stable Baselines Jax (SB3 + Jax) RL algorithms
Python★ 608↓ 1 960/міс6 вер. 2026 р.
npm · PyPI67Добрийіндекс здоров'я
alex-jb/orallexa-ai-trading-agentSelf-tuning multi-agent AI trading system. 8-source signal fusion, Bull/Bear/Judge debate on Claude Opus 4.7, Kelly + ATR position sizing. Python · Kalshi + Polymarket adapters.
Python · TypeScript★ 6026 лип. 2026 р.

JuliaPOMDP/POMDPs.jlMDPs and POMDPs in Julia - An interface for defining, solving, and simulating fully and partially observable Markov decision processes on discrete and continuous spaces.
Julia★ 7648 серп. 2026 р.