← All tags

#ai-safety

74 posts

MCP Server 4

The verification layer for autonomous agents — an independent, capital-aware verdict before an irreversible action (/review), a cryptographi...

Python Apache-2.0 Updated 1w ago
MCP Server 2

Runtime approval gates for AI agent tool calls. Intercept payments and emails before execution

Python Apache-2.0 Updated 2w ago
Claude Skill 18

Security scanner for AI agent skills. Detects prompt injection, data exfiltration, and malicious payloads before you install.

Python NOASSERTION Updated 1w ago
MCP Server 29

Lightweight AI safety middleware that protects humans by intercepting self-harm and criminal intent in LLM prompts. Features a 3-stage safet...

Python Apache-2.0 Updated 2w ago
MCP Server 11

MCP middleware that blocks dangerous AI agent actions using a simple YAML config

TypeScript Updated 1mo ago
Claude Skill 5

🔍 Discover + safety-vet Claude Code extensions before you install them - a 0-100 trust score for discovery + a 1-5 static-scan risk verdict...

Python MIT Updated 1mo ago
MCP Server 25

Cryptographically signed policies that constrain AI agent behavior. Same attestation primitives as cilock, applied to agent execution: ident...

Go Apache-2.0 Updated 1w ago
MCP Server 11

Stop AI coding agents from leaking your API keys. Local proxy + MCP that swaps real secrets for phm_ tokens — works with Claude Code, Cursor...

Rust MIT Updated 3w ago
MCP Server 25

AI coding safety CLI for vibe coding workflows. Checkpoints, undo, anchors, MCP, and secret protection for Claude Code, Cursor, Codex, and O...

Python MIT Updated 3w ago
MCP Server 12

Policy-gated, durable, audited execution for AI agent tool calls. Every action gets a five-verdict policy check, a crash-safe checkpoint, an...

Python NOASSERTION Updated 2w ago
MCP Server 9

Bilingual hands-on roadmap for production-aware AI agents: MCP, memory, RAG, workflows, evaluation, safety, and agent colonies.

Python MIT Updated 3w ago
MCP Server 24

ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches...

JavaScript MIT Updated 1w ago