← All topics

Models

Model releases, capabilities, pricing, and benchmarks.

320 posts

Claude Skill 12

Agent Skill for cross-border product launch readiness: admission checks, target-market benchmarks, localization, documents, labels, pricing,...

Python NOASSERTION Updated 1mo ago
Claude Skill 1.2k

An Agent Skill helping you to optimize Xcode incremental and clean builds by running benchmarks and optimizing build settings.

6 skills Python MIT Updated 3mos ago
Claude Skill 11

AI reasoning skills distilled from 4,665 real Claude Fable 5 chain-of-thought traces. Mathematically tuned. Grade A (100%) emulation accurac...

JavaScript NOASSERTION Updated 1mo ago
Claude Skill 43

Claude Code skills that make Claude (Fable) the conductor of AI worker fleets: many parallel OpenAI Codex workers plus Opus design agents, o...

Python MIT Updated 2w ago
Claude Skill 1.9k

The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / pr...

JavaScript MIT Updated 2w ago
Claude Skill 26

Eco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowered effort. Measured...

JavaScript MIT Updated 3w ago
Claude Skill 115

Architecture-first skill lifecycle for AI agents. 5 modes: CREATE → EVAL → EDIT → REVIEW → PACKAGE. Integrates Anthropic's eval engine (grad...

Python MIT Updated 1mo ago
Claude Skill 11

Project-local agent harness skill for traceable AI workflows

Python MIT Updated 4w ago
Claude Skill 50

Six Claude Code skills that harden Opus 4.8 toward frontier behavior — written by Fable 5, pressure-tested on the target model with transcri...

TypeScript MIT Updated 2w ago
Claude Skill 35

Use cultivar to test your Agent Skills, run them in sandboxes, and across different agents.

Python MIT Updated 2w ago
Claude Skill 111

Fable 5-grade work discipline for any Claude model — a Claude Code skill + guard hooks (plan gate, model ceiling, per-task enforcement) that...

Python MIT Updated 1w ago
Claude Skill 75

[COLM'26] SkillLearnBench is the first benchmark for evaluating continual learning methods that automatically generate agent skills.

Python MIT Updated 2w ago