All tags
Catalogue tag

#grpo

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

6 records
Tagged “grpo”Ranked by health index
PyPI
99Exceptionalhealth index
huggingface/trl
Train transformer language models with reinforcement learning.
Python★ 19.1K↓ 4.1M/moAug 24, 2026
Apache-2.0Aug 24, 2026 · metrics 2.10.0
PyPI
94Exceptionalhealth index
PrimeIntellect-ai/verifiers
Our library for RL environments + evals
Python★ 4,565↓ 378.1K/moAug 28, 2026
MITAug 28, 2026 · metrics 2.10.0
PyPI
94Exceptionalhealth index
modelscope/ms-swift
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
Python★ 15.4K↓ 111K/moAug 27, 2026
Apache-2.0Aug 27, 2026 · metrics 2.10.0
PyPI
92Excellenthealth index
JudgmentLabs/judgeval
The Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.
Python★ 1,057↓ 233.1K/moAug 13, 2026
Apache-2.0Aug 13, 2026 · metrics 2.10.0
PyPI
75Goodhealth index
freesolo-co/flash
LoRA post-training for open-weight models: SFT, GRPO, and on-policy distillation. Describe a run in TOML; Flash allocates a GPU, trains, streams checkpoints, and serves the adapter.
Python★ 2↓ 5,834/moAug 29, 2026
Apache-2.0Aug 29, 2026 · metrics 2.10.0
PyPI
34At Riskhealth index
waybarrios/crystal
CRYSTAL: Beyond Final Answers: Benchmark for Transparent Multimodal Reasoning Evaluation | arXiv 2603.13099
Python★ 2Jul 17, 2026
No licenseJul 17, 2026 · metrics 2.10.0