faster
Claude Code Benchmark
Latest internal run: secret-scanner CLI. Use the Distill.codes CLI to repeat a paired local comparison.
Fable 5 xhigh
Same task, same prompt, same model, separate fresh directories. Both runs produced implementations that passed the task smoke verifier.
Main reductions
fewer output tokens
less source LOC
lower estimated provider cost
Supporting clean result
Fable 5 and Opus 5 are recommended for evaluation. Sonnet 5 showed a smaller and unstable effect in our runs.
faster
fewer output tokens
less source LOC
Methodology
- same task and prompt
- same model and effort inside each pair
- separate fresh directories or worktrees
- smoke verifier for task acceptance
- input, cache, and output tokens tracked separately
- source LOC tracked separately from support and test files
Distill.codes does not guarantee smaller diffs, lower token use, or faster runs on every task. Short or simple tasks may show little difference; longer and more complex tasks tend to reveal workflow waste more clearly.
Run it locally
Use your proxy URL to run direct and Distill.codes Claude Code tasks in separate local folders. The CLI saves the report locally.
npx distill-codes bench "https://proxy.distill.codes/<proxy-key>/essential/anthropic"Try it on your workload
Use the trial to compare review time, output tokens, files touched, LOC, and whether the final patch still satisfies your task.
Choose plan and start 3-day trial