200 AI agent skills, hardened with targeted behavioral guardrails. Free drop-in replacements.
#ai-safety
80 posts
Static scanner for MCP-connected AI agent pipelines — 225 rules across 11 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10...
OpenCode memory governance plugin — Knowledge Graph + Hybrid Search + Smart Compaction + Tool Firewall + Stop Gate. 6 MCP tools with slash-c...
Audit all locally configured MCP servers for permission risks, prompt injection threats, and schema drift
Prompt-injection firewall for LLM applications — 33 input detectors, 9 output scanners, federated ed25519-signed threat-intel feed. Apache 2...
Still trusting results that AI generated, tested, and declared successful by itself? This Claude Code methodology adds evidence markers, aud...
The verification layer for autonomous agents — an independent, capital-aware verdict before an irreversible action (/review), a cryptographi...
Runtime approval gates for AI agent tool calls. Intercept payments and emails before execution
Security scanner for AI agent skills. Detects prompt injection, data exfiltration, and malicious payloads before you install.
Lightweight AI safety middleware that protects humans by intercepting self-harm and criminal intent in LLM prompts. Features a 3-stage safet...
MCP middleware that blocks dangerous AI agent actions using a simple YAML config
🔍 Discover + safety-vet Claude Code extensions before you install them - a 0-100 trust score for discovery + a 1-5 static-scan risk verdict...