Усі теги
Тег каталогу

#ai-safety

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

21 запис
З тегом «ai-safety»Упорядковано за індексом здоров'я
npm · Go · crates.io
86Відміннийіндекс здоров'я
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 4 869↓ 12.3K/міс20 лип. 2026 р.
MIT20 лип. 2026 р. · метрики 1.13.0
npm
70Добрийіндекс здоров'я
jkheadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 72↓ 87.7K/міс14 лип. 2026 р.
MIT14 лип. 2026 р. · метрики 1.13.0
npm
66Помірнийіндекс здоров'я
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9 089/міс22 лип. 2026 р.
Власна ліцензія22 лип. 2026 р. · метрики 1.13.0
npm
65Помірнийіндекс здоров'я
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4 611/міс19 лип. 2026 р.
MIT19 лип. 2026 р. · метрики 1.13.0
npm
65Помірнийіндекс здоров'я
emiliaprotocol/emilia-protocol
Receipt Required for AI agents — no receipt, no irreversible action. An open, offline-verifiable authorization-receipt protocol: a named human signs the exact high-risk action (payment, deploy, delete, permissions) before it runs, and anyone can verify it later, trusting no one. Apache-2.0 · formally verified · IETF Internet-Drafts.
JavaScript★ 3↓ 934/міс15 лип. 2026 р.
Apache-2.015 лип. 2026 р. · метрики 1.13.0
npm
64Помірнийіндекс здоров'я
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3 459/міс19 лип. 2026 р.
Apache-2.019 лип. 2026 р. · метрики 1.13.0
npm
64Помірнийіндекс здоров'я
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2 900/міс17 лип. 2026 р.
Власна ліцензія17 лип. 2026 р. · метрики 1.13.0
PyPI
64Помірнийіндекс здоров'я
b7n0de/proofbundle
Offline cryptographic receipts for AI evaluation results — Ed25519 + RFC 6962 Merkle + optional SD-JWT. Integrity, not truth
Python★ 2↓ 6 574/міс23 лип. 2026 р.
MIT23 лип. 2026 р. · метрики 1.13.0
PyPI
62Помірнийіндекс здоров'я
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9 632/міс16 лип. 2026 р.
Apache-2.016 лип. 2026 р. · метрики 1.13.0
PyPI
62Помірнийіндекс здоров'я
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1 793/міс17 лип. 2026 р.
Apache-2.017 лип. 2026 р. · метрики 1.13.0
PyPI
61Помірнийіндекс здоров'я
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 8 503/міс14 лип. 2026 р.
AGPL-3.014 лип. 2026 р. · метрики 1.13.0
Go
61Помірнийіндекс здоров'я
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 819 лип. 2026 р.
MIT19 лип. 2026 р. · метрики 1.13.0
PyPI
59Помірнийіндекс здоров'я
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5 683/міс16 лип. 2026 р.
Власна ліцензія16 лип. 2026 р. · метрики 1.13.0
PyPI
58Помірнийіндекс здоров'я
provael/provael
Provael — red-team open Vision-Language-Action (VLA) robot policies in simulation and report an Attack Success Rate (ASR). Prove it. Prevail.
Python★ 3↓ 2 044/міс19 лип. 2026 р.
Apache-2.019 лип. 2026 р. · метрики 1.13.0
Go
57Помірнийіндекс здоров'я
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 2221 лип. 2026 р.
Apache-2.021 лип. 2026 р. · метрики 1.13.0
crates.io
55Помірнийіндекс здоров'я
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 4↓ 96/міс15 лип. 2026 р.
Apache-2.015 лип. 2026 р. · метрики 1.13.0
PyPI
55Помірнийіндекс здоров'я
rozmiard/sclite
Schema-backed contract layer for governed AI/security workflows: intent, policy decisions, approvals, execution receipts, evidence, and replayable audit traces.
Python★ 218 лип. 2026 р.
MIT18 лип. 2026 р. · метрики 1.13.0
npm
54Помірнийіндекс здоров'я
agledger-ai/sdk
AGLedger™ SDK — Accountability and audit infrastructure for agentic systems. Patent Pending.
TypeScript★ 0↓ 2 001/міс15 лип. 2026 р.
Власна ліцензія15 лип. 2026 р. · метрики 1.13.0
Go
45У зоні ризикуіндекс здоров'я
nlink-jp/mcp-guardian
MCP governance proxy - zero-dependency single binary (Go)
Go★ 019 лип. 2026 р.
MIT19 лип. 2026 р. · метрики 1.13.0
PyPI
42У зоні ризикуіндекс здоров'я
vdalal/agentx-security-sdk
Runtime action firewall for AI agents: the keyless Shield. Blocks catastrophic tool calls (DROP TABLE, SSRF, secret exfil) in-process, no key. MIT.
Python★ 0↓ 5 745/міс19 лип. 2026 р.
MIT19 лип. 2026 р. · метрики 1.13.0
npm
32У зоні ризикуіндекс здоров'я
jyswee/a2a-trustgate
Compliance for AI systems in production — screen every AI-agent action through a 4-gate firewall before it runs, with an OCSF-native audit trail mapped to EU AI Act / SOC 2 / NIST / HIPAA.
Змішані★ 0↓ 2 332/міс22 лип. 2026 р.
Власна ліцензія22 лип. 2026 р. · метрики 1.13.0