← All tags

#evals

8 posts

Claude Skill 8

Minimal, self-checking, simulation-driven improvement for AI agents

Python MIT Updated 2w ago
Claude Skill 1

Califa Cards: A Comprehensive Skill Card Standard and Testing Framework

1 skills Python Apache-2.0 Updated 1mo ago
MCP Server 8

A portable, model-agnostic test suite of declarative YAML scenarios for evaluating AI assistant security against prompt-injection and data-e...

Python Apache-2.0 Updated 2mos ago
MCP Server 83

Composable Pi coding agent with MCP, LSP, agent chains, prompt presets, and local eval telemetry

TypeScript MIT Updated 1mo ago
MCP Server 26.1k

Mastra is the modern TypeScript framework for AI-powered applications and agents.

TypeScript NOASSERTION Updated 1mo ago