PyPI99卓越健康指数huggingface/trlTrain transformer language models with reinforcement learning.Python★ 19.1K↓ 4.1M/月2026年8月24日Apache-2.02026年8月24日 · 指标 2.10.0
PyPI94卓越健康指数modelscope/ms-swiftUse PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).Python★ 15.4K↓ 111K/月2026年8月27日Apache-2.02026年8月27日 · 指标 2.10.0
PyPI86优秀健康指数MakazhanAlpamys/SoupSoup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.Python★ 742026年7月15日Apache-2.02026年7月15日 · 指标 2.10.0
PyPI75良好健康指数freesolo-co/flashLoRA post-training for open-weight models: SFT, GRPO, and on-policy distillation. Describe a run in TOML; Flash allocates a GPU, trains, streams checkpoints, and serves the adapter.Python★ 2↓ 5,834/月2026年8月29日Apache-2.02026年8月29日 · 指标 2.10.0
npm · PyPI71良好健康指数invergent-ai/surogateTraining/Fine-tuning at the speed of lightC++ · Python · Cuda★ 806↓ 0/月2026年7月21日Apache-2.02026年7月21日 · 指标 2.10.0