← All tags

#agent-safety

21 posts

Claude Skill 1

Score every action an AI agent can take, cap its autonomy at the hottest dimension, and ship the mandate, exceptions and owner's manual behi...

Python NOASSERTION Updated 2d ago
Claude Skill 3

Safety-first workflow routing for Codex and Claude Code: explicit scope, authorization, evidence, and rollback for high-impact agent actions...

Python MIT Updated 2w ago
Claude Skill 1

Task Compass Skill: deterministic, auditable task routing for AI agents and OpenClaw

Python MIT Updated 1mo ago
Tool 112

Runtime safety for AI coding agents with real-time enforcement, system-event monitoring, and long-horizon provenance. Supports Claude Code,...

Rust Apache-2.0 Updated 1mo ago
MCP Server 2

Deterministic security engine for AI agents. See, block, rewind, prove.

Python AGPL-3.0 Updated 2mos ago
MCP Server 2

Runtime guardrails for TypeScript AI agents. Prevents duplicate tool calls, enforces per-user cost budgets, and gates irreversible actions....

TypeScript MIT Updated 4mos ago
MCP Server 4

Documentation for Cycles — AI agent governance, runtime budget, action authority, MCP integration

Vue Apache-2.0 Updated 1mo ago
MCP Server 4

Your AI agent just burned $200. AgentGuard stops it at $5. Runtime cost guardrails for AI agents — budget enforcement, loop detection, kill...

Python MIT Updated 1mo ago
MCP Server 17

Guardrails service for AI agents. Default-deny tool call evaluation with LLM safety analysis, priority-ordered decision matrix, and human-in...

Python NOASSERTION Updated 2mos ago
MCP Server 7

Execution control layer for AI agents — prevents duplicate or incorrect real-world actions under retries, uncertainty, and stale context.

Python Apache-2.0 Updated 1mo ago
Claude Skill 35

Trust nothing. Ship safely. — Skeptical-reading and prompt-injection defense skill for AI agents. Provenance tagging, red-flag patterns, ref...

Shell MIT Updated 4mos ago