Back to Topics Directory
Topic Hub

#llm-security (21 Repositories)

Ranked open-source repositories tagged with #llm-security, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

36.8%

Avg Review Latency

60.5h

Filter by language

21 repositories tagged #llm-security

S TierPython 198 30 GFIs

DobermanCore/Doberman-Core

Your AI's guard dog. Doberman sits at runtime, gating every input, output and tool call to stop unsafe or unintended actions before they execute.

94.2%
Merge Rate
1d
First Review
82%
1st-Timers
13
Maintainers
S TierTypeScript 676

arcjet/arcjet-js

Runtime security for AI apps and agents: prompt injection detection, tool-call authorization, sensitive-data redaction, bot protection, and rate limiting. Drop it into your JS/TS code.

86.6%
Merge Rate
4h
First Review
80%
1st-Timers
8
Maintainers
S TierPython 207

fu351/Doberman-Core

Your AI's guard dog. Doberman sits at runtime, gating every input, output and tool call to stop unsafe or unintended actions before they execute.

93.8%
Merge Rate
1d
First Review
83%
1st-Timers
11
Maintainers
A TierGo 822

luckyPipewrench/pipelock

Open-source AI agent firewall for MCP security and agent egress. Scans mediated HTTP, MCP, A2A, and WebSocket traffic for exfiltration, SSRF, and prompt injection, and emits mediator-signed action receipts: verifiable audit evidence from outside the agent.

95.3%
Merge Rate
5d
First Review
100%
1st-Timers
4
Maintainers
A TierGo 339

adithyan-ak/AgentHound

Offensive security framework for AI agent infrastructure - recon, credential looting, model exfiltration, poisoning, and attack-path analysis across MCP, A2A, gateways, and AI services. BloodHound for the agentic stack.

92.6%
Merge Rate
5d
First Review
100%
1st-Timers
2
Maintainers
A TierPython 155 2 GFIs

OWASP/www-project-agent-memory-guard

OWASP Foundation web repository

69.6%
Merge Rate
1d
First Review
50%
1st-Timers
4
Maintainers
B TierPython 5.7k

Tencent/AI-Infra-Guard

A full-stack AI Red Teaming platform securing AI ecosystems via Agent Scan, Skills Scan, MCP scan, AI Infra scan and LLM jailbreak evaluation.

70.3%
Merge Rate
7d
First Review
80%
1st-Timers
10
Maintainers
B TierPython 7.0k 1 GFIs

NVIDIA-NeMo/Guardrails

NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems.

64.6%
Merge Rate
2d
First Review
25%
1st-Timers
30
Maintainers
B TierTypeScript 371 2 GFIs

Agent-Threat-Rule/agent-threat-rules

Open detection-rule standard for AI agent security threats — like Sigma, but for AI agents. Executable rules across 10 categories; merged into Microsoft AGT, Cisco AI Defense, MISP, OWASP, FINOS & SigmaHQ. MIT-licensed.

71.5%
Merge Rate
6d
First Review
100%
1st-Timers
1
Maintainers
B TierJUJupyter Notebook 59.0k

pathwaycom/llm-app

Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.

33.3%
Merge Rate
<1h
First Review
0%
1st-Timers
2
Maintainers
D TierJavaScript 165

edward-playground/aidefense-framework

An open-source knowledge base of defensive countermeasures to protect AI/ML systems. Features interactive views and maps defenses to known threats from frameworks like MITRE ATLAS, MAESTRO, and OWASP.

0.0%
Merge Rate
25d
First Review
0%
1st-Timers
0
Maintainers
D TierGo 228

praetorian-inc/julius

Simple LLM service identification - translate IP:Port to Ollama, vLLM, LiteLLM, or 60+ other AI services in seconds

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 560

Zyrexnn/Cybermes

Autonomous Offensive Security, Bug Bounty & Red Teaming Agent Framework powered by Hermes Agent, specialized reasoning skills, and multi-model LLM orchestration.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierGo 304

packyme/privacy-filter

LLM privacy gateway in Go — millisecond-latency PII and secret redaction. Used in production by PackyCode.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 486

liu00222/Open-Prompt-Injection

This repository provides a benchmark for prompt injection attacks and defenses in LLMs

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 109

DmitrL-dev/AISecurity

AI Security Platform: Defense (61 Rust engines + Micro-Model Swarm) + Offense (39K+ payloads)

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierTypeScript 349

adversa-ai/secureclaw

SecureClaw - Security Plugin and Skill for OpenClaw OWASP-Aligned

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 2.0k

msoedov/agentic_security

Agentic LLM Vulnerability Scanner / AI red teaming kit 🧪

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierGo 424

dataiku/kiji-proxy

Privacy proxy for your OpenAI requests

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 496

deadbits/vigil-llm

⚡ Vigil ⚡ Detect prompt injections, jailbreaks, and other potentially risky Large Language Model (LLM) inputs

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 345

getagentseal/agentseal

Security toolkit for AI agents. Scan your machine for dangerous skills and MCP configs, monitor for supply chain attacks, test prompt injection resistance, and audit live MCP servers for tool poisoning.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
Best Llm-security Open Source Repositories & C-Rank™ | GetMerged