← All tags

#ai-safety

74 posts

MCP Server 6

Agentic security control plane for MCP and AI agent tool calls. MCP-native policy gateway with topology discovery and audit.

Rust NOASSERTION Updated 1w ago
MCP Server 168

Dialectical reasoning architecture for LLMs (Thesis → Antithesis → Synthesis)

Python MIT Updated 4mos ago
MCP Server 6

Security enforcement plugin for Claude Code. Blocks dangerous commands, audits every tool call, detects prompt injection.

JavaScript Updated 2mos ago
MCP Server 2

LikenessGuard is an open-source reference project for pre-generation consent enforcement in AI image generation. V2 demonstrates the concept...

Python NOASSERTION Updated 3w ago
MCP Server 4

Your AI agent just burned $200. AgentGuard stops it at $5. Runtime cost guardrails for AI agents — budget enforcement, loop detection, kill...

Python MIT Updated 1w ago
MCP Server 6

Proof-backed AI agent for checking suspicious job posts, recruiter messages, and apply links, now live with case study and pilot intake.

TypeScript Updated 1w ago
MCP Server 8

MCP servers expose tools with no information about what they actually do at runtime. mcpsafetywarden sits between your agent and any MCP ser...

Python NOASSERTION Updated 2w ago
MCP Server 4

MCP & Claude Code security scanner — threat-models plugins, MCP servers, hooks, skills & connectors with an LLM before you trust them. Catch...

Go Apache-2.0 Updated 1w ago
MCP Server 11

Glass Box Framework — runtime constitutional verification for AI answers. Trust Cards with claim-level reasoning chains, formal ECS scoring,...

TypeScript Apache-2.0 Updated 1mo ago
MCP Server 4

Runtime artifact existence & freshness verification for AI agent completion claims — a lightweight, zero-LLM MCP gate that source-binds 'don...

TypeScript MIT Updated 1w ago
MCP Server 6

A safer MySQL CLI for AI coding agents: connection profiles, SSH tunnels, and automatic sensitive-data masking before query output reaches C...

Go MIT Updated 1w ago