← All tags

#ai-safety

80 posts

Claude Skill 1

AI Skill 执行可靠性审查工具。评估 Skill 被 AI 执行时的可复现性、可信度与业务适配性。Halu = Hallucination, Catch = 捕获。

Python MIT Updated 1mo ago
Claude Skill 33

AI agents now operate with authority. Authority without discipline is how complex systems fail. Nuclear’s control loop, ported to AI-assiste...

Python MIT Updated 21h ago
Claude Skill 89

Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effecti...

Python NOASSERTION Updated 3mos ago
Claude Skill 51

🔬 Verifiable AI-Augmented Engineering Framework - Stop AI hallucinations with formal traceability (REQ→ART→TC). Agent Skills for Claude Cod...

Python CC-BY-SA-4.0 Updated 1mo ago
MCP Server 8

C# SQL Agent MCP server featuring raw SQL input, strict AST validation, and an embedded Admin UI. Eliminates LLM hallucinations and security...

C# Apache-2.0 Updated 1mo ago
MCP Server 17

Trust infrastructure for AI agents. Know who produced a value, when, and that it hasn't been tampered with. Zero dependencies. Pure C.

C MIT Updated 1mo ago
MCP Server 7

A carnivorous honeypot for AI agents. Every deployment generates a unique persona so no two instances look alike. Detects, fingerprints, and...

Python Apache-2.0 Updated 6mos ago
MCP Server 4

MCP server for AI security intelligence. Check any MCP server for supply-chain threats before installing -- from Claude, Cursor, or Windsurf...

TypeScript NOASSERTION Updated 5mos ago
MCP Server 2

A local firewall for AI agent commands

Go MIT Updated 2mos ago
MCP Server 8

A safety layer for AI coding agents. CLAUDE.md/AGENTS.md generator, MCP runtime guardrail, pre-commit hook, GitHub Action.

TypeScript MIT Updated 3mos ago