Engineering-first OpenAI Codex guide: 14 modules, 9 installable skills, reproducible benchmarks, testing, review, orchestration, and a Livin...
Models
Model releases, capabilities, pricing, and benchmarks.
375 posts
An entire marketing department as one Claude skill. 14 modules: audits scored 0-100, an 18-tactic hook engine, copy graded before you see it...
An agent skill that plays ARC-AGI-3. One rule: say what an action will do before you spend it. Claude Code on Opus 5 finished all 25 public...
DeepSeek has released an experimental AI model called V4-Flash-Vision-Exp that the company claims performs comparably to Anthropic's Opus 4....
Tech-lead orchestration for Claude Code — the top-tier model (Fable) keeps architecture & decisions, cheap subagents (Sonnet/Opus) do the ro...
Cross-platform Agent Skill for IT infrastructure architecture, server/network/UPS sizing, SCADA/IT-OT, hardware selection, live pricing, BOM...
A cross-harness, pure-skill implementation of the Fable 5 orchestrator pattern proposed by Anthropic — your agent as the boss; Codex, Cursor...
Agent Skill that researches, compares, and shortlists hotels with evidence-backed, like-for-like pricing — prepares the booking decision, ne...
Distilled operating skills for daily-driver Claude models — few dense rules, executable gates over long prose.
Fable-model operating discipline for any AI coding agent - Claude Code, Cursor, Copilot, Codex, Gemini, Windsurf, Cline, Amp & more. One npx...
A permission-aware Agent Skill for reproducible, auditable batch episode collection across VLA, policy, world-model, and embodied benchmark...
Agent Skill for ASD-STE100 Simplified Technical English. Makes LLMs write text that survives one read. Benchmarked on 12 models: 95.5% fewer...