Усі теги
Тег каталогу

#ai-safety

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

42 записи
З тегом «ai-safety»Упорядковано за індексом здоров'я
npm · Go · crates.io
96Винятковийіндекс здоров'я
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 6 138↓ 14.7K/міс28 серп. 2026 р.
MIT28 серп. 2026 р. · метрики 2.10.0
PyPI · npm
94Винятковийіндекс здоров'я
sattyamjjain/agent-audit-kit
Static scanner for MCP-connected AI agent pipelines — 271 rules across 12 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10/10, GitHub Action, SARIF, public CVE-to-rule ledger.
Python★ 13↓ 2 808/міс2 серп. 2026 р.
MIT2 серп. 2026 р. · метрики 2.10.0
PyPI · RubyGems
90Відміннийіндекс здоров'я
OWASP/www-project-agent-memory-guard
OWASP Foundation web repository
Python★ 155↓ 2 902/міс25 серп. 2026 р.
Apache-2.025 серп. 2026 р. · метрики 2.10.0
PyPI · Go · npm
90Відміннийіндекс здоров'я
cordum-io/cordum
The action firewall for AI agents. Enforce policy and human approval before risky tool calls, shell commands, workflows, and production changes, with auditable evidence.
Go · TypeScript · Python★ 494↓ 17/міс9 серп. 2026 р.
Власна ліцензія9 серп. 2026 р. · метрики 2.10.0
npm · PyPI
86Відміннийіндекс здоров'я
issdandavis/SCBE-AETHERMOORE
Geometric AI governance and evaluation framework with a 14-layer security pipeline, semantic projection, and reproducible benchmark lanes.
Python · TypeScript★ 6↓ 4 607/міс26 серп. 2026 р.
MIT26 серп. 2026 р. · метрики 2.10.0
npm
84Відміннийіндекс здоров'я
JKHeadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 77↓ 157.9K/міс5 вер. 2026 р.
MIT5 вер. 2026 р. · метрики 2.10.0
npm
80Відміннийіндекс здоров'я
lua-ai-global/governance
Zero-dependency TypeScript SDK for AI agent governance: policy enforcement, injection detection, tamper-evident audit, and standards mapping (EU AI Act, OWASP, NIST, ISO 42001)
TypeScript★ 25↓ 3 545/міс26 лип. 2026 р.
MIT26 лип. 2026 р. · метрики 2.10.0
npm
78Добрийіндекс здоров'я
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9 089/міс22 лип. 2026 р.
Власна ліцензія22 лип. 2026 р. · метрики 2.10.0
npm
77Добрийіндекс здоров'я
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4 611/міс19 лип. 2026 р.
MIT19 лип. 2026 р. · метрики 2.10.0
PyPI
77Добрийіндекс здоров'я
alizahidraja/isnad
Grade every agent, scraper and model in a claim's chain — provenance, trust scoring and audit evidence for LLM pipelines
Python★ 37↓ 4 204/міс29 серп. 2026 р.
Apache-2.029 серп. 2026 р. · метрики 2.10.0
npm
75Добрийіндекс здоров'я
bookedsolidtech/rea
Zero-trust governance layer for Claude Code. Policy-enforced MCP gateway with autonomy controls, middleware chain, audit log, HALT kill-switch, and prompt-injection defense.
TypeScript · Shell★ 0↓ 1 950/міс27 лип. 2026 р.
MIT27 лип. 2026 р. · метрики 2.10.0
PyPI · npm
75Добрийіндекс здоров'я
gautamvarmadatla/mcpsafetywarden
MCP servers expose tools with no information about what they actually do at runtime. mcpsafetywarden sits between your agent and any MCP server, profiling tool behavior, blocking destructive calls, and running active security audits before you trust them in a workflow.
Python★ 9↓ 2 923/міс1 серп. 2026 р.
Власна ліцензія1 серп. 2026 р. · метрики 2.10.0
npm
73Добрийіндекс здоров'я
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3 459/міс19 лип. 2026 р.
Apache-2.019 лип. 2026 р. · метрики 2.10.0
PyPI · npm
73Добрийіндекс здоров'я
aiexponenthq/license-compliance-checker
License Compliance Checker — Multi-ecosystem license + AI model scanner for EU AI Act Article 53 GPAI compliance. SBOM, SARIF, training-data risk. Apache 2.0.
Python · TypeScript★ 1↓ 258/міс27 лип. 2026 р.
Apache-2.027 лип. 2026 р. · метрики 2.10.0
PyPI
73Добрийіндекс здоров'я
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 628/міс22 серп. 2026 р.
AGPL-3.022 серп. 2026 р. · метрики 2.10.0
PyPI
73Добрийіндекс здоров'я
b7n0de/proofbundle
Offline cryptographic receipts for AI evaluation results — Ed25519 + RFC 6962 Merkle + optional SD-JWT. Integrity, not truth
Python★ 2↓ 6 574/міс23 лип. 2026 р.
MIT23 лип. 2026 р. · метрики 2.10.0
Go
73Добрийіндекс здоров'я
gumieri/nenya
A lightweight, highly secure AI API Gateway/Proxy written in Go. Acts as transparent middleware between local AI coding clients (OpenCode/Pi/Cursor) and upstream LLM providers (Gemini, DeepSeek, Zhipu z.ai).
Go★ 2524 лип. 2026 р.
Apache-2.024 лип. 2026 р. · метрики 2.10.0
PyPI
73Добрийіндекс здоров'я
philpaz/recusal
Deterministic governance for Claude and MCP tool calls. Pin approved capabilities, detect drift, and refuse unsafe or unapproved actions before execution. No model in the decision path.
Python★ 3↓ 2 854/міс27 лип. 2026 р.
Apache-2.027 лип. 2026 р. · метрики 2.10.0
npm
71Добрийіндекс здоров'я
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2 900/міс17 лип. 2026 р.
Власна ліцензія17 лип. 2026 р. · метрики 2.10.0
PyPI
69Добрийіндекс здоров'я
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9 632/міс16 лип. 2026 р.
Apache-2.016 лип. 2026 р. · метрики 2.10.0
PyPI
69Добрийіндекс здоров'я
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1 793/міс17 лип. 2026 р.
Apache-2.017 лип. 2026 р. · метрики 2.10.0
PyPI
69Добрийіндекс здоров'я
fathom-lab/styxx
Verification for the agent era. Your coding agent's PR summary cannot lie about its diff - one CI line. Plus the research: the first map of which AI minds can read each other. Every claim machine-verified against committed receipts, negatives included. pip install styxx
Python★ 14↓ 2 231/міс19 серп. 2026 р.
MIT19 серп. 2026 р. · метрики 2.10.0
PyPI
67Добрийіндекс здоров'я
CognitiveThoughtEngine/constitutional-agent-governance
The WHY layer for AI agents: six constitutional gates + 12 hard constraints, plus cross-session risk composition — catches agents that pass every individual gate but accumulate risk across a sequence (stateless engines can't). pip install constitutional-agent
Python★ 0↓ 331/міс27 лип. 2026 р.
MIT27 лип. 2026 р. · метрики 2.10.0
Go
67Добрийіндекс здоров'я
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 819 лип. 2026 р.
MIT19 лип. 2026 р. · метрики 2.10.0
crates.io
63Помірнийіндекс здоров'я
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 5↓ 159/міс6 вер. 2026 р.
Apache-2.06 вер. 2026 р. · метрики 2.10.0
PyPI
62Помірнийіндекс здоров'я
sunglasses-dev/sunglasses
Sunglasses for AI agents. Protection layer + neighborhood watch.
Python★ 4↓ 2 911/міс18 серп. 2026 р.
MIT18 серп. 2026 р. · метрики 2.10.0
Go
62Помірнийіндекс здоров'я
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 2221 лип. 2026 р.
Apache-2.021 лип. 2026 р. · метрики 2.10.0
PyPI
62Помірнийіндекс здоров'я
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5 683/міс16 лип. 2026 р.
Власна ліцензія16 лип. 2026 р. · метрики 2.10.0
PyPI
60Помірнийіндекс здоров'я
YugantM/hvtracker
AI Agent Trust Registry — independent, evidence-based trust scores for 300+ open-source AI agents. Runtime-trust calibrated (MCP, dependencies, provenance drift), with side-by-side comparison and embeddable live badges. Ranked by verifiable signals, not hype.
HTML★ 526 лип. 2026 р.
MIT26 лип. 2026 р. · метрики 2.10.0
PyPI
60Помірнийіндекс здоров'я
haqaliz/belay
The agent harness: sandbox any agent, verify each step by replaying it against real state, and keep a deterministic trace.
Python★ 0↓ 2 022/міс20 серп. 2026 р.
Apache-2.020 серп. 2026 р. · метрики 2.10.0