Score every action an AI agent can take, cap its autonomy at the hottest dimension, and ship the mandate, exceptions and owner's manual behi...
#agent-safety
21 posts
Safety-first workflow routing for Codex and Claude Code: explicit scope, authorization, evidence, and rollback for high-impact agent actions...
Task Compass Skill: deterministic, auditable task routing for AI agents and OpenClaw
Runtime safety for AI coding agents with real-time enforcement, system-event monitoring, and long-horizon provenance. Supports Claude Code,...
Deterministic security engine for AI agents. See, block, rewind, prove.
Runtime guardrails for TypeScript AI agents. Prevents duplicate tool calls, enforces per-user cost budgets, and gates irreversible actions....
Documentation for Cycles — AI agent governance, runtime budget, action authority, MCP integration
Your AI agent just burned $200. AgentGuard stops it at $5. Runtime cost guardrails for AI agents — budget enforcement, loop detection, kill...
Guardrails service for AI agents. Default-deny tool call evaluation with LLM safety analysis, priority-ordered decision matrix, and human-in...
Execution control layer for AI agents — prevents duplicate or incorrect real-world actions under retries, uncertainty, and stale context.
Trust nothing. Ship safely. — Skeptical-reading and prompt-injection defense skill for AI agents. Provenance tagging, red-flag patterns, ref...