All tags
Catalogue tag

#ai-safety

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

20 records
Tagged “ai-safety”Ranked by health index
npm · Go · crates.io
86Excellenthealth index
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 4,869↓ 12.3K/moJul 20, 2026
MITJul 20, 2026 · metrics 1.13.0
npm
70Goodhealth index
jkheadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 72↓ 87.7K/moJul 14, 2026
MITJul 14, 2026 · metrics 1.13.0
npm
66Moderatehealth index
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9,089/moJul 22, 2026
Custom licenseJul 22, 2026 · metrics 1.13.0
npm
65Moderatehealth index
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4,611/moJul 19, 2026
MITJul 19, 2026 · metrics 1.13.0
npm
65Moderatehealth index
emiliaprotocol/emilia-protocol
Receipt Required for AI agents — no receipt, no irreversible action. An open, offline-verifiable authorization-receipt protocol: a named human signs the exact high-risk action (payment, deploy, delete, permissions) before it runs, and anyone can verify it later, trusting no one. Apache-2.0 · formally verified · IETF Internet-Drafts.
JavaScript★ 3↓ 934/moJul 15, 2026
Apache-2.0Jul 15, 2026 · metrics 1.13.0
npm
64Moderatehealth index
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3,459/moJul 19, 2026
Apache-2.0Jul 19, 2026 · metrics 1.13.0
npm
64Moderatehealth index
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2,900/moJul 17, 2026
Custom licenseJul 17, 2026 · metrics 1.13.0
PyPI
62Moderatehealth index
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9,632/moJul 16, 2026
Apache-2.0Jul 16, 2026 · metrics 1.13.0
PyPI
62Moderatehealth index
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1,793/moJul 17, 2026
Apache-2.0Jul 17, 2026 · metrics 1.13.0
PyPI
61Moderatehealth index
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 8,503/moJul 14, 2026
AGPL-3.0Jul 14, 2026 · metrics 1.13.0
Go
61Moderatehealth index
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 8Jul 19, 2026
MITJul 19, 2026 · metrics 1.13.0
PyPI
59Moderatehealth index
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5,683/moJul 16, 2026
Custom licenseJul 16, 2026 · metrics 1.13.0
PyPI
58Moderatehealth index
provael/provael
Provael — red-team open Vision-Language-Action (VLA) robot policies in simulation and report an Attack Success Rate (ASR). Prove it. Prevail.
Python★ 3↓ 2,044/moJul 19, 2026
Apache-2.0Jul 19, 2026 · metrics 1.13.0
Go
57Moderatehealth index
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 22Jul 21, 2026
Apache-2.0Jul 21, 2026 · metrics 1.13.0
crates.io
55Moderatehealth index
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 4↓ 96/moJul 15, 2026
Apache-2.0Jul 15, 2026 · metrics 1.13.0
PyPI
55Moderatehealth index
rozmiard/sclite
Schema-backed contract layer for governed AI/security workflows: intent, policy decisions, approvals, execution receipts, evidence, and replayable audit traces.
Python★ 2Jul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
npm
54Moderatehealth index
agledger-ai/sdk
AGLedger™ SDK — Accountability and audit infrastructure for agentic systems. Patent Pending.
TypeScript★ 0↓ 2,001/moJul 15, 2026
Custom licenseJul 15, 2026 · metrics 1.13.0
Go
45At riskhealth index
nlink-jp/mcp-guardian
MCP governance proxy - zero-dependency single binary (Go)
Go★ 0Jul 19, 2026
MITJul 19, 2026 · metrics 1.13.0
PyPI
42At riskhealth index
vdalal/agentx-security-sdk
Runtime action firewall for AI agents: the keyless Shield. Blocks catastrophic tool calls (DROP TABLE, SSRF, secret exfil) in-process, no key. MIT.
Python★ 0↓ 5,745/moJul 19, 2026
MITJul 19, 2026 · metrics 1.13.0
npm
32At riskhealth index
jyswee/a2a-trustgate
Compliance for AI systems in production — screen every AI-agent action through a 4-gate firewall before it runs, with an OCSF-native audit trail mapped to EU AI Act / SOC 2 / NIST / HIPAA.
Mixed★ 0↓ 2,332/moJul 22, 2026
Custom licenseJul 22, 2026 · metrics 1.13.0