Back to Topics Directory
Topic Hub

#ai-safety (20 Repositories)

Ranked open-source repositories tagged with #ai-safety, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

52.5%

Avg Review Latency

12.7h

Filter by language

20 repositories tagged #ai-safety

S TierPython 198 30 GFIs

DobermanCore/Doberman-Core

Your AI's guard dog. Doberman sits at runtime, gating every input, output and tool call to stop unsafe or unintended actions before they execute.

94.2%
Merge Rate
1d
First Review
82%
1st-Timers
13
Maintainers
S TierRust 541 5 GFIs

h5i-dev/h5i

Sandboxed collaboration for multi-agent teams: a Git-backed message forum with each agent isolated in its own disposable sandbox. Turn a Git repository into a secure message forum for AI agents.

92.8%
Merge Rate
3d
First Review
64%
1st-Timers
7
Maintainers
S TierPython 434

Floe-Labs/floe-guard

The spend meter and budget gate for AI voice agents. Meters STT + TTS + LLM + telephony per call, out of the box (Pipecat, LiveKit — Python & TypeScript). Hard-stops the next turn before it crosses your ceiling. Local, no account, no telemetry. Built by Floe — cost controls for Voice AI.

91.7%
Merge Rate
<1h
First Review
100%
1st-Timers
4
Maintainers
S TierPython 207

fu351/Doberman-Core

Your AI's guard dog. Doberman sits at runtime, gating every input, output and tool call to stop unsafe or unintended actions before they execute.

93.8%
Merge Rate
1d
First Review
83%
1st-Timers
11
Maintainers
A TierPython 118

WhitzardAgent/AgentGuard

AgentGuard: Zero-Trust Security Foundation for AI Agents

92.3%
Merge Rate
<1h
First Review
100%
1st-Timers
3
Maintainers
A TierRust 112

Firma-AI/openfirma

Runtime enforcement boundary for AI agents: a local sidecar that gates every outbound call against Cedar policies you own. Deterministic, call-level, no model on the hot path

83.9%
Merge Rate
17h
First Review
86%
1st-Timers
11
Maintainers
A TierPython 6.0k 12 GFIs

microsoft/agent-governance-toolkit

AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.

78.8%
Merge Rate
2d
First Review
61%
1st-Timers
31
Maintainers
A TierPython 155 2 GFIs

OWASP/www-project-agent-memory-guard

OWASP Foundation web repository

69.6%
Merge Rate
1d
First Review
50%
1st-Timers
4
Maintainers
B TierTypeScript 1.5k

kenryu42/cc-safety-net

An AI coding agent guardrail — a CLI hook that blocks destructive git and filesystem commands and secret file access before they execute. Supports Amp Code, Antigravity CLI, Claude Code, Codex, Copilot CLI, Cursor, Gemini CLI, Hermes Agent, Kimi Code, OpenClaw, OpenCode, and Pi.

71.4%
Merge Rate
13h
First Review
100%
1st-Timers
0
Maintainers
B TierPython 327

ttguy0707/CyberClaw

👾 下一代透明智能体架构 | Next-Gen Transparent Agent Architecture 🔍 全行为审计 | 🛡️ 两段式安全调用 | 🧠 双水位记忆 | ⏰ 心跳任务 📊 P0 级事故率降低 80% | 兼容 OpenClaw + Claude Code 技能生态

100.0%
Merge Rate
4h
First Review
100%
1st-Timers
1
Maintainers
B TierGo 498 1 GFIs

cordum-io/cordum

The action firewall for AI agents. Enforce policy and human approval before risky tool calls, shell commands, workflows, and production changes, with auditable evidence.

82.1%
Merge Rate
13h
First Review
50%
1st-Timers
1
Maintainers
B TierPython 127

x-zheng16/Awesome-Embodied-AI-Safety

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses | 500+ Papers | Perception, Cognition, Planning, Interaction, Agentic System

100.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 119

Prysai/Prysai-LLM-Playbook

An evidence-led, eight-locale LLM playbook: a transferable core, the Codex flagship track, and adapters for ChatGPT, Claude Code, Gemini, DeepSeek, and Grok.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 106

yzhao062/awesome-auditable-ai

Auditing AI agents: a curated list of papers, tools, datasets, benchmarks, and standards covering reliability, monitoring, failure attribution, and decision records.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 109

DmitrL-dev/AISecurity

AI Security Platform: Defense (61 Rust engines + Micro-Model Swarm) + Offense (39K+ payloads)

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierRust 294

Govcraft/rust-docs-mcp-server

🦀 Prevents outdated Rust code suggestions from AI assistants. This MCP server fetches current crate docs, uses embeddings/LLMs, and provides accurate context via a tool call.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 237

yzhao062/anywhere-agents

One config to rule all your AI agents: portable (every project, every session), effective (curated writing, routing, skills), and safer (destructive-command guard).

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierHTML 116

minetechnic2012-lang/claude-ops-inspector

Subagent Verification for Claude AI Code Networks 2026

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 302

agentcontrol/agent-control

Centralized agent control plane for governing runtime agent behavior at scale. Configurable, extensible, and production-ready.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierHTML 115

ChristoAnsek/audited-change-gate

Automated Proof-of-Carrying Change Management for AIOps 2026

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
Best Ai-safety Open Source Repositories & C-Rank™ | GetMerged