The verification layer for autonomous agents — an independent, capital-aware verdict before an irreversible action (/review), a cryptographi...
#ai-safety
74 posts
Runtime approval gates for AI agent tool calls. Intercept payments and emails before execution
Security scanner for AI agent skills. Detects prompt injection, data exfiltration, and malicious payloads before you install.
Lightweight AI safety middleware that protects humans by intercepting self-harm and criminal intent in LLM prompts. Features a 3-stage safet...
MCP middleware that blocks dangerous AI agent actions using a simple YAML config
🔍 Discover + safety-vet Claude Code extensions before you install them - a 0-100 trust score for discovery + a 1-5 static-scan risk verdict...
Cryptographically signed policies that constrain AI agent behavior. Same attestation primitives as cilock, applied to agent execution: ident...
Stop AI coding agents from leaking your API keys. Local proxy + MCP that swaps real secrets for phm_ tokens — works with Claude Code, Cursor...
AI coding safety CLI for vibe coding workflows. Checkpoints, undo, anchors, MCP, and secret protection for Claude Code, Cursor, Codex, and O...
Policy-gated, durable, audited execution for AI agent tool calls. Every action gets a five-verdict policy check, a crash-safe checkpoint, an...
Bilingual hands-on roadmap for production-aware AI agents: MCP, memory, RAG, workflows, evaluation, safety, and agent colonies.
ThumbGate Pre-Action Checks derive rules from repeated failures, flag risky tool calls, hard-block detected secret leaks, and block matches...