← All tags

#ai-safety

80 posts

Claude Skill 23

200 AI agent skills, hardened with targeted behavioral guardrails. Free drop-in replacements.

JavaScript MIT Updated 4mos ago
MCP Server 9

Static scanner for MCP-connected AI agent pipelines — 225 rules across 11 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10...

Python MIT Updated 1mo ago
MCP Server 2

OpenCode memory governance plugin — Knowledge Graph + Hybrid Search + Smart Compaction + Tool Firewall + Stop Gate. 6 MCP tools with slash-c...

TypeScript Apache-2.0 Updated 1mo ago
MCP Server 4

Audit all locally configured MCP servers for permission risks, prompt injection threats, and schema drift

Python MIT Updated 1mo ago
MCP Server 9

Prompt-injection firewall for LLM applications — 33 input detectors, 9 output scanners, federated ed25519-signed threat-intel feed. Apache 2...

Python Apache-2.0 Updated 1mo ago
MCP Server 6

Still trusting results that AI generated, tested, and declared successful by itself? This Claude Code methodology adds evidence markers, aud...

Python MIT Updated 1mo ago
MCP Server 4

The verification layer for autonomous agents — an independent, capital-aware verdict before an irreversible action (/review), a cryptographi...

Python Apache-2.0 Updated 1mo ago
MCP Server 2

Runtime approval gates for AI agent tool calls. Intercept payments and emails before execution

Python Apache-2.0 Updated 2mos ago
Claude Skill 20

Security scanner for AI agent skills. Detects prompt injection, data exfiltration, and malicious payloads before you install.

Python NOASSERTION Updated 1mo ago
MCP Server 29

Lightweight AI safety middleware that protects humans by intercepting self-harm and criminal intent in LLM prompts. Features a 3-stage safet...

Python Apache-2.0 Updated 1mo ago
MCP Server 11

MCP middleware that blocks dangerous AI agent actions using a simple YAML config

TypeScript Updated 2mos ago
Claude Skill 7

🔍 Discover + safety-vet Claude Code extensions before you install them - a 0-100 trust score for discovery + a 1-5 static-scan risk verdict...

Python MIT Updated 2mos ago