Усі теги
Тег каталогу

#agent-testing

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

5 записів
З тегом «agent-testing»Упорядковано за індексом здоров'я
npm · PyPI
94Винятковийіндекс здоров'я
langwatch/scenario
Agentic testing for agentic codebases
Python · TypeScript★ 942↓ 50K/міс5 серп. 2026 р.
Apache-2.05 серп. 2026 р. · метрики 2.10.0
PyPI · npm
89Відміннийіндекс здоров'я
UiPath/coder_eval
Test that your Claude Code skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, activation checks, A/B experiments, CI gates.
Python · TypeScript★ 116↓ 13.9K/міс19 серп. 2026 р.
Apache-2.019 серп. 2026 р. · метрики 2.10.0
npm
80Відміннийіндекс здоров'я
reticlehq/reticle
AI agents can generate code, but they still struggle to understand what they build. Reticle gives them runtime perception of web applications.
TypeScript · JavaScript★ 191↓ 14.3K/міс23 лип. 2026 р.
Власна ліцензія23 лип. 2026 р. · метрики 2.10.0
PyPI · npm · crates.io
65Добрийіндекс здоров'я
automators-com/flowproof
Agents are starting to run real business processes. flowproof tests them like anything else: record one real run, then replay it on every commit with zero LLM calls, asserting which tools were called, with which arguments, in which order, and which were not.
Rust★ 5↓ 5 937/міс4 серп. 2026 р.
Apache-2.04 серп. 2026 р. · метрики 2.10.0
PyPI
59Помірнийіндекс здоров'я
mrwersa/agentverity
Your agent test passed. Would it pass again? Checks whether AI agent test results are repeatable and varied enough to trust as a regression baseline. Reads Promptfoo and DeepEval runs you already have.
Python★ 15 серп. 2026 р.
Apache-2.05 серп. 2026 р. · метрики 2.10.0