Alle Tags
Katalog-Tag

#ai-safety

Alle Repositories im öffentlichen Register, die dieses Tag tragen — aus ihren GitHub-Topics oder den von ihren Paket-Registries veröffentlichten Schlagwörtern. Die Gesundheit wird nach derselben versionierten Methodik gemessen wie im übrigen Register.

42 Einträge
Getaggt als „ai-safety“Geordnet nach Gesundheitsindex
npm · Go · crates.io
96AußergewöhnlichGesundheitsindex
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 6.138↓ 14.7K/Monat28. Aug. 2026
MIT28. Aug. 2026 · Metriken 2.10.0
PyPI · npm
94AußergewöhnlichGesundheitsindex
sattyamjjain/agent-audit-kit
Static scanner for MCP-connected AI agent pipelines — 271 rules across 12 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10/10, GitHub Action, SARIF, public CVE-to-rule ledger.
Python★ 13↓ 2.808/Monat2. Aug. 2026
MIT2. Aug. 2026 · Metriken 2.10.0
PyPI · RubyGems
90ExzellentGesundheitsindex
OWASP/www-project-agent-memory-guard
OWASP Foundation web repository
Python★ 155↓ 2.902/Monat25. Aug. 2026
Apache-2.025. Aug. 2026 · Metriken 2.10.0
PyPI · Go · npm
90ExzellentGesundheitsindex
cordum-io/cordum
The action firewall for AI agents. Enforce policy and human approval before risky tool calls, shell commands, workflows, and production changes, with auditable evidence.
Go · TypeScript · Python★ 494↓ 17/Monat9. Aug. 2026
Eigene Lizenz9. Aug. 2026 · Metriken 2.10.0
npm · PyPI
86ExzellentGesundheitsindex
issdandavis/SCBE-AETHERMOORE
Geometric AI governance and evaluation framework with a 14-layer security pipeline, semantic projection, and reproducible benchmark lanes.
Python · TypeScript★ 6↓ 4.607/Monat26. Aug. 2026
MIT26. Aug. 2026 · Metriken 2.10.0
npm
84ExzellentGesundheitsindex
JKHeadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 77↓ 157.9K/Monat5. Sept. 2026
MIT5. Sept. 2026 · Metriken 2.10.0
npm
80ExzellentGesundheitsindex
lua-ai-global/governance
Zero-dependency TypeScript SDK for AI agent governance: policy enforcement, injection detection, tamper-evident audit, and standards mapping (EU AI Act, OWASP, NIST, ISO 42001)
TypeScript★ 25↓ 3.545/Monat26. Juli 2026
MIT26. Juli 2026 · Metriken 2.10.0
npm
78GutGesundheitsindex
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9.089/Monat22. Juli 2026
Eigene Lizenz22. Juli 2026 · Metriken 2.10.0
npm
77GutGesundheitsindex
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4.611/Monat19. Juli 2026
MIT19. Juli 2026 · Metriken 2.10.0
PyPI
77GutGesundheitsindex
alizahidraja/isnad
Grade every agent, scraper and model in a claim's chain — provenance, trust scoring and audit evidence for LLM pipelines
Python★ 37↓ 4.204/Monat29. Aug. 2026
Apache-2.029. Aug. 2026 · Metriken 2.10.0
npm
75GutGesundheitsindex
bookedsolidtech/rea
Zero-trust governance layer for Claude Code. Policy-enforced MCP gateway with autonomy controls, middleware chain, audit log, HALT kill-switch, and prompt-injection defense.
TypeScript · Shell★ 0↓ 1.950/Monat27. Juli 2026
MIT27. Juli 2026 · Metriken 2.10.0
PyPI · npm
75GutGesundheitsindex
gautamvarmadatla/mcpsafetywarden
MCP servers expose tools with no information about what they actually do at runtime. mcpsafetywarden sits between your agent and any MCP server, profiling tool behavior, blocking destructive calls, and running active security audits before you trust them in a workflow.
Python★ 9↓ 2.923/Monat1. Aug. 2026
Eigene Lizenz1. Aug. 2026 · Metriken 2.10.0
npm
73GutGesundheitsindex
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3.459/Monat19. Juli 2026
Apache-2.019. Juli 2026 · Metriken 2.10.0
PyPI · npm
73GutGesundheitsindex
aiexponenthq/license-compliance-checker
License Compliance Checker — Multi-ecosystem license + AI model scanner for EU AI Act Article 53 GPAI compliance. SBOM, SARIF, training-data risk. Apache 2.0.
Python · TypeScript★ 1↓ 258/Monat27. Juli 2026
Apache-2.027. Juli 2026 · Metriken 2.10.0
PyPI
73GutGesundheitsindex
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 628/Monat22. Aug. 2026
AGPL-3.022. Aug. 2026 · Metriken 2.10.0
PyPI
73GutGesundheitsindex
b7n0de/proofbundle
Offline cryptographic receipts for AI evaluation results — Ed25519 + RFC 6962 Merkle + optional SD-JWT. Integrity, not truth
Python★ 2↓ 6.574/Monat23. Juli 2026
MIT23. Juli 2026 · Metriken 2.10.0
Go
73GutGesundheitsindex
gumieri/nenya
A lightweight, highly secure AI API Gateway/Proxy written in Go. Acts as transparent middleware between local AI coding clients (OpenCode/Pi/Cursor) and upstream LLM providers (Gemini, DeepSeek, Zhipu z.ai).
Go★ 2524. Juli 2026
Apache-2.024. Juli 2026 · Metriken 2.10.0
PyPI
73GutGesundheitsindex
philpaz/recusal
Deterministic governance for Claude and MCP tool calls. Pin approved capabilities, detect drift, and refuse unsafe or unapproved actions before execution. No model in the decision path.
Python★ 3↓ 2.854/Monat27. Juli 2026
Apache-2.027. Juli 2026 · Metriken 2.10.0
npm
71GutGesundheitsindex
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2.900/Monat17. Juli 2026
Eigene Lizenz17. Juli 2026 · Metriken 2.10.0
PyPI
69GutGesundheitsindex
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9.632/Monat16. Juli 2026
Apache-2.016. Juli 2026 · Metriken 2.10.0
PyPI
69GutGesundheitsindex
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1.793/Monat17. Juli 2026
Apache-2.017. Juli 2026 · Metriken 2.10.0
PyPI
69GutGesundheitsindex
fathom-lab/styxx
Verification for the agent era. Your coding agent's PR summary cannot lie about its diff - one CI line. Plus the research: the first map of which AI minds can read each other. Every claim machine-verified against committed receipts, negatives included. pip install styxx
Python★ 14↓ 2.231/Monat19. Aug. 2026
MIT19. Aug. 2026 · Metriken 2.10.0
PyPI
67GutGesundheitsindex
CognitiveThoughtEngine/constitutional-agent-governance
The WHY layer for AI agents: six constitutional gates + 12 hard constraints, plus cross-session risk composition — catches agents that pass every individual gate but accumulate risk across a sequence (stateless engines can't). pip install constitutional-agent
Python★ 0↓ 331/Monat27. Juli 2026
MIT27. Juli 2026 · Metriken 2.10.0
Go
67GutGesundheitsindex
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 819. Juli 2026
MIT19. Juli 2026 · Metriken 2.10.0
crates.io
63MittelGesundheitsindex
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 5↓ 159/Monat6. Sept. 2026
Apache-2.06. Sept. 2026 · Metriken 2.10.0
PyPI
62MittelGesundheitsindex
sunglasses-dev/sunglasses
Sunglasses for AI agents. Protection layer + neighborhood watch.
Python★ 4↓ 2.911/Monat18. Aug. 2026
MIT18. Aug. 2026 · Metriken 2.10.0
Go
62MittelGesundheitsindex
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 2221. Juli 2026
Apache-2.021. Juli 2026 · Metriken 2.10.0
PyPI
62MittelGesundheitsindex
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5.683/Monat16. Juli 2026
Eigene Lizenz16. Juli 2026 · Metriken 2.10.0
PyPI
60MittelGesundheitsindex
YugantM/hvtracker
AI Agent Trust Registry — independent, evidence-based trust scores for 300+ open-source AI agents. Runtime-trust calibrated (MCP, dependencies, provenance drift), with side-by-side comparison and embeddable live badges. Ranked by verifiable signals, not hype.
HTML★ 526. Juli 2026
MIT26. Juli 2026 · Metriken 2.10.0
PyPI
60MittelGesundheitsindex
haqaliz/belay
The agent harness: sandbox any agent, verify each step by replaying it against real state, and keep a deterministic trace.
Python★ 0↓ 2.022/Monat20. Aug. 2026
Apache-2.020. Aug. 2026 · Metriken 2.10.0