Todas las etiquetas
Etiqueta del catálogo

#ai-safety

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

20 registros
Con la etiqueta «ai-safety»Ordenado por índice de salud
npm · Go · crates.io
86Excelenteíndice de salud
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 4869↓ 12.3K/mes20 jul 2026
MIT20 jul 2026 · métricas 1.13.0
npm
70Buenoíndice de salud
jkheadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 72↓ 87.7K/mes14 jul 2026
MIT14 jul 2026 · métricas 1.13.0
npm
66Moderadoíndice de salud
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9089/mes22 jul 2026
Licencia propia22 jul 2026 · métricas 1.13.0
npm
65Moderadoíndice de salud
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4611/mes19 jul 2026
MIT19 jul 2026 · métricas 1.13.0
npm
65Moderadoíndice de salud
emiliaprotocol/emilia-protocol
Receipt Required for AI agents — no receipt, no irreversible action. An open, offline-verifiable authorization-receipt protocol: a named human signs the exact high-risk action (payment, deploy, delete, permissions) before it runs, and anyone can verify it later, trusting no one. Apache-2.0 · formally verified · IETF Internet-Drafts.
JavaScript★ 3↓ 934/mes15 jul 2026
Apache-2.015 jul 2026 · métricas 1.13.0
npm
64Moderadoíndice de salud
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3459/mes19 jul 2026
Apache-2.019 jul 2026 · métricas 1.13.0
npm
64Moderadoíndice de salud
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2900/mes17 jul 2026
Licencia propia17 jul 2026 · métricas 1.13.0
PyPI
62Moderadoíndice de salud
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9632/mes16 jul 2026
Apache-2.016 jul 2026 · métricas 1.13.0
PyPI
62Moderadoíndice de salud
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1793/mes17 jul 2026
Apache-2.017 jul 2026 · métricas 1.13.0
PyPI
61Moderadoíndice de salud
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 8503/mes14 jul 2026
AGPL-3.014 jul 2026 · métricas 1.13.0
Go
61Moderadoíndice de salud
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 819 jul 2026
MIT19 jul 2026 · métricas 1.13.0
PyPI
59Moderadoíndice de salud
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5683/mes16 jul 2026
Licencia propia16 jul 2026 · métricas 1.13.0
PyPI
58Moderadoíndice de salud
provael/provael
Provael — red-team open Vision-Language-Action (VLA) robot policies in simulation and report an Attack Success Rate (ASR). Prove it. Prevail.
Python★ 3↓ 2044/mes19 jul 2026
Apache-2.019 jul 2026 · métricas 1.13.0
Go
57Moderadoíndice de salud
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 2221 jul 2026
Apache-2.021 jul 2026 · métricas 1.13.0
crates.io
55Moderadoíndice de salud
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 4↓ 96/mes15 jul 2026
Apache-2.015 jul 2026 · métricas 1.13.0
PyPI
55Moderadoíndice de salud
rozmiard/sclite
Schema-backed contract layer for governed AI/security workflows: intent, policy decisions, approvals, execution receipts, evidence, and replayable audit traces.
Python★ 218 jul 2026
MIT18 jul 2026 · métricas 1.13.0
npm
54Moderadoíndice de salud
agledger-ai/sdk
AGLedger™ SDK — Accountability and audit infrastructure for agentic systems. Patent Pending.
TypeScript★ 0↓ 2001/mes15 jul 2026
Licencia propia15 jul 2026 · métricas 1.13.0
Go
45En riesgoíndice de salud
nlink-jp/mcp-guardian
MCP governance proxy - zero-dependency single binary (Go)
Go★ 019 jul 2026
MIT19 jul 2026 · métricas 1.13.0
PyPI
42En riesgoíndice de salud
vdalal/agentx-security-sdk
Runtime action firewall for AI agents: the keyless Shield. Blocks catastrophic tool calls (DROP TABLE, SSRF, secret exfil) in-process, no key. MIT.
Python★ 0↓ 5745/mes19 jul 2026
MIT19 jul 2026 · métricas 1.13.0
npm
32En riesgoíndice de salud
jyswee/a2a-trustgate
Compliance for AI systems in production — screen every AI-agent action through a 4-gate firewall before it runs, with an OCSF-native audit trail mapped to EU AI Act / SOC 2 / NIST / HIPAA.
Mixto★ 0↓ 2332/mes22 jul 2026
Licencia propia22 jul 2026 · métricas 1.13.0