🛡️ A curated list of resources on agent skills security: attacks, defenses, frameworks, and benchmarks for securing AI agent tool use and sk...
#ai-safety
74 posts
Deterministic policy language for AI agents. Z3 + TLA+ dual-engine formal verification. Runtime enforcement <1ms.
Open-source prompt injection detector — 5 layers, 91.7% F1, ~27ms, offline, Apache 2.0
Pre-execution policy engine for AI agents. Every tool call checked before execution.
All AIs are sycophants.
MCP EU AI Act Compliance Scanner - Open source tool to detect EU AI Act violations in codebases
200 AI agent skills, hardened with targeted behavioral guardrails. Free drop-in replacements.
Static scanner for MCP-connected AI agent pipelines — 225 rules across 11 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10...
OpenCode memory governance plugin — Knowledge Graph + Hybrid Search + Smart Compaction + Tool Firewall + Stop Gate. 6 MCP tools with slash-c...
Audit all locally configured MCP servers for permission risks, prompt injection threats, and schema drift
Prompt-injection firewall for LLM applications — 33 input detectors, 9 output scanners, federated ed25519-signed threat-intel feed. Apache 2...
Still trusting results that AI generated, tested, and declared successful by itself? This Claude Code methodology adds evidence markers, aud...