全部标签
目录标签

#ai-safety

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

22 条记录
标签为“ai-safety”按健康指数排序
npm · Go · crates.io
86优秀健康指数
microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
Python★ 4,869↓ 12.3K/月2026年7月20日
MIT2026年7月20日 · 指标 1.13.0
npm
70良好健康指数
jkheadley/instar
Persistent Claude Code agents with scheduling, sessions, memory, and Telegram.
TypeScript★ 72↓ 87.7K/月2026年7月14日
MIT2026年7月14日 · 指标 1.13.0
npm
66中等健康指数
node9-ai/node9-proxy
The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.
TypeScript★ 209↓ 9,089/月2026年7月22日
自定义许可证2026年7月22日 · 指标 1.13.0
npm
65中等健康指数
IgorGanapolsky/ThumbGate
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches in strict mode.
JavaScript · HTML★ 24↓ 4,611/月2026年7月19日
MIT2026年7月19日 · 指标 1.13.0
npm
65中等健康指数
emiliaprotocol/emilia-protocol
Receipt Required for AI agents — no receipt, no irreversible action. An open, offline-verifiable authorization-receipt protocol: a named human signs the exact high-risk action (payment, deploy, delete, permissions) before it runs, and anyone can verify it later, trusting no one. Apache-2.0 · formally verified · IETF Internet-Drafts.
JavaScript★ 3↓ 934/月2026年7月15日
Apache-2.02026年7月15日 · 指标 1.13.0
npm
64中等健康指数
Keesan12/martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll back failures, learn from loops, and prove ROI across every repo.
TypeScript · JavaScript★ 38↓ 3,459/月2026年7月19日
Apache-2.02026年7月19日 · 指标 1.13.0
npm
64中等健康指数
alexandriashai/cbrowser
Cognitive Browser: The browser automation that thinks. Constitutional safety • Persona UX testing • Natural language interface • Self-healing selectors • Built for AI agents
TypeScript · JavaScript★ 18↓ 2,900/月2026年7月17日
自定义许可证2026年7月17日 · 指标 1.13.0
PyPI
64中等健康指数
b7n0de/proofbundle
Offline cryptographic receipts for AI evaluation results — Ed25519 + RFC 6962 Merkle + optional SD-JWT. Integrity, not truth
Python★ 2↓ 6,574/月2026年7月23日
MIT2026年7月23日 · 指标 1.13.0
PyPI
62中等健康指数
auspexai/tenant-sdk
SDK for authoring research tenants on AuspexAI
Python★ 0↓ 9,632/月2026年7月16日
Apache-2.02026年7月16日 · 指标 1.13.0
PyPI
62中等健康指数
crucible-security/crucible
pytest for AI agents - Autonomous red-teaming, behavioral monitoring & security testing for LLM agents
Python★ 45↓ 1,793/月2026年7月17日
Apache-2.02026年7月17日 · 指标 1.13.0
PyPI
61中等健康指数
auspexai/worker
AuspexAI volunteer worker for donating compute to AI research
Python★ 0↓ 8,503/月2026年7月14日
AGPL-3.02026年7月14日 · 指标 1.13.0
Go
61中等健康指数
nikicat/secrets-dispatcher
Per-operation approval and audit logging for secret access and git commit signing on Linux
Go · TypeScript★ 82026年7月19日
MIT2026年7月19日 · 指标 1.13.0
PyPI
59中等健康指数
veldtlabs/veldt-kya
KYA (Know Your Agents) — Open-source trust, governance, and evidentiary assurance infrastructure for autonomous systems. Built on KYP (Know Your Principal), a unified trust model for human users, AI agents, service accounts, and machine identities.
Python★ 2↓ 5,683/月2026年7月16日
自定义许可证2026年7月16日 · 指标 1.13.0
PyPI
58中等健康指数
provael/provael
Provael — red-team open Vision-Language-Action (VLA) robot policies in simulation and report an Attack Success Rate (ASR). Prove it. Prevail.
Python★ 3↓ 2,044/月2026年7月19日
Apache-2.02026年7月19日 · 指标 1.13.0
Go
57中等健康指数
trustabl/trustabl
Static analyzer for agent reliability.
Go★ 222026年7月21日
Apache-2.02026年7月21日 · 指标 1.13.0
crates.io
55中等健康指数
kahramanemir/Vallum
Security boundary CLI proxy between AI coding agents and your shell — redacts secrets, neutralizes prompt injection, wraps untrusted terminal output, and audits every command. Single Rust binary.
Rust★ 4↓ 96/月2026年7月15日
Apache-2.02026年7月15日 · 指标 1.13.0
PyPI
55中等健康指数
rozmiard/sclite
Schema-backed contract layer for governed AI/security workflows: intent, policy decisions, approvals, execution receipts, evidence, and replayable audit traces.
Python★ 22026年7月18日
MIT2026年7月18日 · 指标 1.13.0
npm
54中等健康指数
agledger-ai/sdk
AGLedger™ SDK — Accountability and audit infrastructure for agentic systems. Patent Pending.
TypeScript★ 0↓ 2,001/月2026年7月15日
自定义许可证2026年7月15日 · 指标 1.13.0
Go
45存在风险健康指数
nlink-jp/mcp-guardian
MCP governance proxy - zero-dependency single binary (Go)
Go★ 02026年7月19日
MIT2026年7月19日 · 指标 1.13.0
PyPI
42存在风险健康指数
vdalal/agentx-security-sdk
Runtime action firewall for AI agents: the keyless Shield. Blocks catastrophic tool calls (DROP TABLE, SSRF, secret exfil) in-process, no key. MIT.
Python★ 0↓ 5,745/月2026年7月19日
MIT2026年7月19日 · 指标 1.13.0
npm · crates.io
33存在风险健康指数
wiber/thetacog-mcp
npx thetacog-mcp — a decidable, on-chip, LLM-free placement receipt a stranger recomputes byte-for-byte. Free for builders; patent-licensed for production & financial use (US Patent App. 19/637,714).
JavaScript · HTML★ 0↓ 4,326/月2026年7月23日
自定义许可证2026年7月23日 · 指标 1.13.0
npm
32存在风险健康指数
jyswee/a2a-trustgate
Compliance for AI systems in production — screen every AI-agent action through a 4-gate firewall before it runs, with an OCSF-native audit trail mapped to EU AI Act / SOC 2 / NIST / HIPAA.
混合★ 0↓ 2,332/月2026年7月22日
自定义许可证2026年7月22日 · 指标 1.13.0