Agent Skill for cross-border product launch readiness: admission checks, target-market benchmarks, localization, documents, labels, pricing,...
Models
Model releases, capabilities, pricing, and benchmarks.
320 posts
An Agent Skill helping you to optimize Xcode incremental and clean builds by running benchmarks and optimizing build settings.
AI reasoning skills distilled from 4,665 real Claude Fable 5 chain-of-thought traces. Mathematically tuned. Grade A (100%) emulation accurac...
Claude Code skills that make Claude (Fable) the conductor of AI worker fleets: many parallel OpenAI Codex workers plus Opus design agents, o...
The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / pr...
Eco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowered effort. Measured...
Architecture-first skill lifecycle for AI agents. 5 modes: CREATE → EVAL → EDIT → REVIEW → PACKAGE. Integrates Anthropic's eval engine (grad...
Project-local agent harness skill for traceable AI workflows
Six Claude Code skills that harden Opus 4.8 toward frontier behavior — written by Fable 5, pressure-tested on the target model with transcri...
Use cultivar to test your Agent Skills, run them in sandboxes, and across different agents.
Fable 5-grade work discipline for any Claude model — a Claude Code skill + guard hooks (plan gate, model ceiling, per-task enforcement) that...
[COLM'26] SkillLearnBench is the first benchmark for evaluating continual learning methods that automatically generate agent skills.