Todas las etiquetas
Etiqueta del catálogo

#ai-safety

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

42 registros
Con la etiqueta «ai-safety»Ordenado por índice de salud
npm · Go · crates.io
96Excepcionalíndice de salud
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 6138↓ 14.7K/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
PyPI · npm
94Excepcionalíndice de salud
sattyamjjain/agent-audit-kit
Static scanner for MCP-connected AI agent pipelines — 271 rules across 12 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10/10, GitHub Action, SARIF, public CVE-to-rule ledger.
Python★ 13↓ 2808/mes2 ago 2026
MIT2 ago 2026 · métricas 2.10.0
PyPI · RubyGems
90Excelenteíndice de salud
OWASP/www-project-agent-memory-guard
OWASP Foundation web repository
Python★ 155↓ 2902/mes25 ago 2026
Apache-2.025 ago 2026 · métricas 2.10.0
PyPI · Go · npm
90Excelenteíndice de salud
cordum-io/cordum
The action firewall for AI agents. Enforce policy and human approval before risky tool calls, shell commands, workflows, and production changes, with auditable evidence.
Go · TypeScript · Python★ 494↓ 17/mes9 ago 2026
Licencia propia9 ago 2026 · métricas 2.10.0
npm · PyPI
86Excelenteíndice de salud
issdandavis/SCBE-AETHERMOORE
Geometric AI governance and evaluation framework with a 14-layer security pipeline, semantic projection, and reproducible benchmark lanes.
Python · TypeScript★ 6↓ 4607/mes26 ago 2026
MIT26 ago 2026 · métricas 2.10.0
npm
84Excelenteíndice de salud
JKHeadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 77↓ 157.9K/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
npm
80Excelenteíndice de salud
lua-ai-global/governance
Zero-dependency TypeScript SDK for AI agent governance: policy enforcement, injection detection, tamper-evident audit, and standards mapping (EU AI Act, OWASP, NIST, ISO 42001)
TypeScript★ 25↓ 3545/mes26 jul 2026
MIT26 jul 2026 · métricas 2.10.0
npm
78Buenoíndice de salud
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9089/mes22 jul 2026
Licencia propia22 jul 2026 · métricas 2.10.0
npm
77Buenoíndice de salud
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4611/mes19 jul 2026
MIT19 jul 2026 · métricas 2.10.0
PyPI
77Buenoíndice de salud
alizahidraja/isnad
Grade every agent, scraper and model in a claim's chain — provenance, trust scoring and audit evidence for LLM pipelines
Python★ 37↓ 4204/mes29 ago 2026
Apache-2.029 ago 2026 · métricas 2.10.0
npm
75Buenoíndice de salud
bookedsolidtech/rea
Zero-trust governance layer for Claude Code. Policy-enforced MCP gateway with autonomy controls, middleware chain, audit log, HALT kill-switch, and prompt-injection defense.
TypeScript · Shell★ 0↓ 1950/mes27 jul 2026
MIT27 jul 2026 · métricas 2.10.0
PyPI · npm
75Buenoíndice de salud
gautamvarmadatla/mcpsafetywarden
MCP servers expose tools with no information about what they actually do at runtime. mcpsafetywarden sits between your agent and any MCP server, profiling tool behavior, blocking destructive calls, and running active security audits before you trust them in a workflow.
Python★ 9↓ 2923/mes1 ago 2026
Licencia propia1 ago 2026 · métricas 2.10.0
npm
73Buenoíndice de salud
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3459/mes19 jul 2026
Apache-2.019 jul 2026 · métricas 2.10.0
PyPI · npm
73Buenoíndice de salud
aiexponenthq/license-compliance-checker
License Compliance Checker — Multi-ecosystem license + AI model scanner for EU AI Act Article 53 GPAI compliance. SBOM, SARIF, training-data risk. Apache 2.0.
Python · TypeScript★ 1↓ 258/mes27 jul 2026
Apache-2.027 jul 2026 · métricas 2.10.0
PyPI
73Buenoíndice de salud
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 628/mes22 ago 2026
AGPL-3.022 ago 2026 · métricas 2.10.0
PyPI
73Buenoíndice de salud
b7n0de/proofbundle
Offline cryptographic receipts for AI evaluation results — Ed25519 + RFC 6962 Merkle + optional SD-JWT. Integrity, not truth
Python★ 2↓ 6574/mes23 jul 2026
MIT23 jul 2026 · métricas 2.10.0
Go
73Buenoíndice de salud
gumieri/nenya
A lightweight, highly secure AI API Gateway/Proxy written in Go. Acts as transparent middleware between local AI coding clients (OpenCode/Pi/Cursor) and upstream LLM providers (Gemini, DeepSeek, Zhipu z.ai).
Go★ 2524 jul 2026
Apache-2.024 jul 2026 · métricas 2.10.0
PyPI
73Buenoíndice de salud
philpaz/recusal
Deterministic governance for Claude and MCP tool calls. Pin approved capabilities, detect drift, and refuse unsafe or unapproved actions before execution. No model in the decision path.
Python★ 3↓ 2854/mes27 jul 2026
Apache-2.027 jul 2026 · métricas 2.10.0
npm
71Buenoíndice de salud
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2900/mes17 jul 2026
Licencia propia17 jul 2026 · métricas 2.10.0
PyPI
69Buenoíndice de salud
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9632/mes16 jul 2026
Apache-2.016 jul 2026 · métricas 2.10.0
PyPI
69Buenoíndice de salud
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1793/mes17 jul 2026
Apache-2.017 jul 2026 · métricas 2.10.0
PyPI
69Buenoíndice de salud
fathom-lab/styxx
Verification for the agent era. Your coding agent's PR summary cannot lie about its diff - one CI line. Plus the research: the first map of which AI minds can read each other. Every claim machine-verified against committed receipts, negatives included. pip install styxx
Python★ 14↓ 2231/mes19 ago 2026
MIT19 ago 2026 · métricas 2.10.0
PyPI
67Buenoíndice de salud
CognitiveThoughtEngine/constitutional-agent-governance
The WHY layer for AI agents: six constitutional gates + 12 hard constraints, plus cross-session risk composition — catches agents that pass every individual gate but accumulate risk across a sequence (stateless engines can't). pip install constitutional-agent
Python★ 0↓ 331/mes27 jul 2026
MIT27 jul 2026 · métricas 2.10.0
Go
67Buenoíndice de salud
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 819 jul 2026
MIT19 jul 2026 · métricas 2.10.0
crates.io
63Moderadoíndice de salud
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 5↓ 159/mes6 sept 2026
Apache-2.06 sept 2026 · métricas 2.10.0
PyPI
62Moderadoíndice de salud
sunglasses-dev/sunglasses
Sunglasses for AI agents. Protection layer + neighborhood watch.
Python★ 4↓ 2911/mes18 ago 2026
MIT18 ago 2026 · métricas 2.10.0
Go
62Moderadoíndice de salud
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 2221 jul 2026
Apache-2.021 jul 2026 · métricas 2.10.0
PyPI
62Moderadoíndice de salud
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5683/mes16 jul 2026
Licencia propia16 jul 2026 · métricas 2.10.0
PyPI
60Moderadoíndice de salud
YugantM/hvtracker
AI Agent Trust Registry — independent, evidence-based trust scores for 300+ open-source AI agents. Runtime-trust calibrated (MCP, dependencies, provenance drift), with side-by-side comparison and embeddable live badges. Ranked by verifiable signals, not hype.
HTML★ 526 jul 2026
MIT26 jul 2026 · métricas 2.10.0
PyPI
60Moderadoíndice de salud
haqaliz/belay
The agent harness: sandbox any agent, verify each step by replaying it against real state, and keep a deterministic trace.
Python★ 0↓ 2022/mes20 ago 2026
Apache-2.020 ago 2026 · métricas 2.10.0