All tags
Catalogue tag

#regression-testing

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

14 records
Tagged “regression-testing”Ranked by health index
PyPI
94Exceptionalhealth index
reframe-hpc/reframe
A powerful Python framework for writing and running portable regression tests and benchmarks for HPC systems.
Python★ 283↓ 15.3K/moSep 4, 2026
BSD-3-ClauseSep 4, 2026 · metrics 2.10.0
PyPI · npm
89Excellenthealth index
UiPath/coder_eval
Test that your Claude Code skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, activation checks, A/B experiments, CI gates.
Python · TypeScript★ 116↓ 13.9K/moAug 19, 2026
Apache-2.0Aug 19, 2026 · metrics 2.10.0
PyPI · npm
88Excellenthealth index
hidai25/eval-view
Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.
Python★ 124↓ 2,207/moJul 26, 2026
Apache-2.0Jul 26, 2026 · metrics 2.10.0
crates.io
83Excellenthealth index
typst-community/tytanic
A test runner for typst projects.
Rust★ 129↓ 1,209/moAug 3, 2026
Apache-2.0Aug 3, 2026 · metrics 2.10.0
npm
78Goodhealth index
BenSheridanEdwards/StyleProof
Catch every CSS change before it ships — review PRs and certify refactors by the browser's computed styles, not pixels. Works with any styling system.
JavaScript · TypeScript★ 2↓ 6,417/moAug 22, 2026
MITAug 22, 2026 · metrics 2.10.0
PyPI
78Goodhealth index
attenlabs/hotato
Find what broke in your agent calls. Pin it so it never ships again. Local voice-agent call forensics and regression guards.
Python★ 1↓ 1,680/moAug 22, 2026
MITAug 22, 2026 · metrics 2.10.0
PyPI
77Goodhealth index
rtl-buddy/rtl_buddy
RTL Buddy is a Python CLI for Verilog and SystemVerilog regression testing, Verilator and VCS simulation workflows, coverage, and YAML-based RTL verification automation.
Python★ 3↓ 3,758/moAug 15, 2026
BSD-3-ClauseAug 15, 2026 · metrics 2.10.0
Go
67Goodhealth index
kabirnarang39/skillci
CI for Claude Skills — lint, eval, and regression-test SKILL.md files across a model matrix, with a self-growing eval loop that turns uncovered regressions into permanent test cases.
Go★ 8Sep 2, 2026
Apache-2.0Sep 2, 2026 · metrics 2.10.0
Maven
65Goodhealth index
arextest/arex-agent-java
Lightweight Java agent for traffic capture and replay, enhancing testing and debugging.
Java★ 558Aug 9, 2026
Apache-2.0Aug 9, 2026 · metrics 2.10.0
npm
62Moderatehealth index
Medhovarsh/forkmind
🧠 Local-first LLM state branching & debugging. Capture, visualize, branch, and regression-test LLM calls as a DAG. Free & local via Ollama, any OpenAI-compatible API, MCP for agents. Zero config.
JavaScript★ 2↓ 2,178/moJul 24, 2026
MITJul 24, 2026 · metrics 2.10.0
NuGet
60Moderatehealth index
AnthonyLloyd/CsCheck
Random testing library for C#
C#★ 210Aug 3, 2026
Apache-2.0Aug 3, 2026 · metrics 2.10.0
Go
59Moderatehealth index
boringSQL/regresql
Catch broken queries and performance regressions before production. SQL regression testing, EXPLAIN plan baselines, and CI/CD integration for PostgreSQL.
Go★ 393Aug 22, 2026
BSD-2-ClauseAug 22, 2026 · metrics 2.10.0
RubyGems
59Moderatehealth index
justi/ruby_llm-contract
Validate and retry LLM outputs for ruby_llm. Describe the JSON response you expect, fall back to a stronger model when the cheaper one fails the rules, and gate CI on regressions — all as one contract object per step.
Ruby★ 35Aug 22, 2026
MITAug 22, 2026 · metrics 2.10.0
npm
41Weakhealth index
Aegis-Runner/AegisRunner
AegisRunner crawls any website from a single URL, discovers every page, form, and interaction, then uses AI to generate a complete Playwright test suite. Also runs accessibility, SEO, security, and performance audits on every page — no recording, no scripting, no setup.
JavaScript★ 0↓ 8,920/moJul 25, 2026
MITJul 25, 2026 · metrics 2.10.0