Multi-engine LLM benchmark & monitoring CLI for Apple Silicon
Models
Model releases, capabilities, pricing, and benchmarks.
375 posts
An adversarially benchmarked reference implementation for pre-action agent authorization. Framework-agnostic tool permissions, identity veri...
🛡️ A curated list of resources on agent skills security: attacks, defenses, frameworks, and benchmarks for securing AI agent tool use and sk...
🚄 Query Renfe train schedules with ease using an MCP server that integrates GTFS data and supports real-time pricing and flexible date form...
Agentic Firewall v2 is a runtime security middleware and red-team benchmark harness designed to secure Model Context Protocol (MCP) applicat...
Race a baseline vs a skill or MCP on real tasks. Hard checks show if it got better, faster, or cheaper. Numbers, not vibes.
Quantitative finance and derivative pricing
A Model Context Protocol (MCP) server for Magic: The Gathering Commander format, providing comprehensive card information, rulings, pricing,...
A Model Context Protocol (MCP) server for Elsevier Scopus. Allows AI assistants like Claude to search academic papers, retrieve abstracts, a...
Benchmark, evaluate, and optimize skills to ensure reliable performance across all LLMs
Evidence-backed reference for AI capabilities, pricing tiers, availability, and platform support across ChatGPT, Claude, Gemini, Copilot, an...
MCP server for real-time news with bias scoring, live stock/ETF/crypto data, AI options pricing, balanced news synthesis, and meme search. 1...