← All tags

#ai-safety

74 posts

Claude Skill 41

🛡️ A curated list of resources on agent skills security: attacks, defenses, frameworks, and benchmarks for securing AI agent tool use and sk...

Updated 1w ago
MCP Server 15

Deterministic policy language for AI agents. Z3 + TLA+ dual-engine formal verification. Runtime enforcement <1ms.

Python Apache-2.0 Updated 1mo ago
MCP Server 9

Open-source prompt injection detector — 5 layers, 91.7% F1, ~27ms, offline, Apache 2.0

Python Apache-2.0 Updated 1mo ago
MCP Server 7

Pre-execution policy engine for AI agents. Every tool call checked before execution.

TypeScript NOASSERTION Updated 2mos ago
MCP Server 11

MCP EU AI Act Compliance Scanner - Open source tool to detect EU AI Act violations in codebases

Python MIT Updated 1mo ago
Claude Skill 23

200 AI agent skills, hardened with targeted behavioral guardrails. Free drop-in replacements.

JavaScript MIT Updated 3mos ago
MCP Server 9

Static scanner for MCP-connected AI agent pipelines — 225 rules across 11 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10...

Python MIT Updated 1w ago
MCP Server 2

OpenCode memory governance plugin — Knowledge Graph + Hybrid Search + Smart Compaction + Tool Firewall + Stop Gate. 6 MCP tools with slash-c...

TypeScript Apache-2.0 Updated 1w ago
MCP Server 4

Audit all locally configured MCP servers for permission risks, prompt injection threats, and schema drift

Python MIT Updated 1w ago
MCP Server 9

Prompt-injection firewall for LLM applications — 33 input detectors, 9 output scanners, federated ed25519-signed threat-intel feed. Apache 2...

Python Apache-2.0 Updated 1w ago
MCP Server 6

Still trusting results that AI generated, tested, and declared successful by itself? This Claude Code methodology adds evidence markers, aud...

Python MIT Updated 1w ago