Agent skill for deepfake detection & media safety — detect AI-generated audio, images, and video with Resemble AI
#ai-safety
80 posts
Merge gates and safety checks for AI coding agents. Works with Claude Code, Cursor, Windsurf, Codex via MCP. Detect scope violations, missin...
AI Constraint Engine — enforces CLAUDE.md, .cursorrules, AGENTS.md rules as laws. 51 MCP tools, 991 tests. Official MCP Registry. npx speclo...
Enforce zero-trust rules for AI agents to prevent hallucinations, unsafe actions, and policy bypasses
Hands-on study companion for the Claude Certified Architect - Foundations (CCA-F) exam.
Real-time trustworthiness evaluation and safety interception for AI agents. Semantic analysis, safe alternative suggestions, multi-step atta...
Agentic security control plane for MCP and AI agent tool calls. MCP-native policy gateway with topology discovery and audit.
Dialectical reasoning architecture for LLMs (Thesis → Antithesis → Synthesis)
Security enforcement plugin for Claude Code. Blocks dangerous commands, audits every tool call, detects prompt injection.
LikenessGuard is an open-source reference project for pre-generation consent enforcement in AI image generation. V2 demonstrates the concept...
Your AI agent just burned $200. AgentGuard stops it at $5. Runtime cost guardrails for AI agents — budget enforcement, loop detection, kill...
Proof-backed AI agent for checking suspicious job posts, recruiter messages, and apply links, now live with case study and pilot intake.