全部标签
目录标签

#ai-safety

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

42 条记录
标签为“ai-safety”按健康指数排序
PyPI
60中等健康指数
provael/provael
Provael — red-team open Vision-Language-Action (VLA) robot policies in simulation and report an Attack Success Rate (ASR). Prove it. Prevail.
Python★ 3↓ 2,044/月2026年7月19日
Apache-2.02026年7月19日 · 指标 2.10.0
PyPI
59中等健康指数
Python★ 0↓ 3,656/月2026年8月1日
Apache-2.02026年8月1日 · 指标 2.10.0
npm
57中等健康指数
neuradex/llm-rail
Deterministic workflow control CLI for LLM (Coding) agents
TypeScript★ 0↓ 5,751/月2026年9月5日
MIT2026年9月5日 · 指标 2.10.0
npm
56中等健康指数
agledger-ai/sdk
AGLedger™ SDK — Accountability and audit infrastructure for agentic systems. Patent Pending.
TypeScript★ 0↓ 2,001/月2026年7月15日
自定义许可证2026年7月15日 · 指标 2.10.0
PyPI
56中等健康指数
rozmiard/sclite
Schema-backed contract layer for governed AI/security workflows: intent, policy decisions, approvals, execution receipts, evidence, and replayable audit traces.
Python★ 22026年7月18日
MIT2026年7月18日 · 指标 2.10.0
Go
53中等健康指数
AccursedGalaxy/driver-os
Verifiable coding-agent harness in Go: external test gates, typed outcomes, sandboxed execution, multi-model providers, and signed proof bundles.
Go★ 02026年8月30日
MIT2026年8月30日 · 指标 2.10.0
Go
50中等健康指数
loop-eng/loopguard
Circuit breaker daemon for AI agent loops
Go★ 02026年8月25日
MIT2026年8月25日 · 指标 2.10.0
Go
45薄弱健康指数
nlink-jp/mcp-guardian
MCP governance proxy - zero-dependency single binary (Go)
Go★ 02026年7月19日
MIT2026年7月19日 · 指标 2.10.0
PyPI
36薄弱健康指数
vdalal/agentx-security-sdk
Runtime action firewall for AI agents: the keyless Shield. Blocks catastrophic tool calls (DROP TABLE, SSRF, secret exfil) in-process, no key. MIT.
Python★ 0↓ 5,745/月2026年7月19日
MIT2026年7月19日 · 指标 2.10.0
npm · crates.io
27存在风险健康指数
wiber/thetacog-mcp
npx thetacog-mcp — a decidable, on-chip, LLM-free placement receipt a stranger recomputes byte-for-byte. Free for builders; patent-licensed for production & financial use (US Patent App. 19/637,714).
JavaScript · HTML★ 0↓ 4,326/月2026年7月23日
自定义许可证2026年7月23日 · 指标 2.10.0
npm
25存在风险健康指数
jyswee/a2a-trustgate
Compliance for AI systems in production — screen every AI-agent action through a 4-gate firewall before it runs, with an OCSF-native audit trail mapped to EU AI Act / SOC 2 / NIST / HIPAA.
混合★ 0↓ 2,332/月2026年7月22日
自定义许可证2026年7月22日 · 指标 2.10.0
npm
20存在风险健康指数
golproductions/check
The anti-hallucination layer for AI agents.
混合★ 0↓ 2,728/月2026年9月5日
自定义许可证2026年9月5日 · 指标 2.10.0