Claude Haiku 4.5 in Claude Code: tokens per question

Haiku is the lightest Claude tier in this comparison, suited to focused coding questions where a short answer is enough.

The figures

Bare Claude Code
55,500

cache-weighted input tokens per question

With capsul
19,300

cache-weighted input tokens per question

less input sent
x2.88

90% interval: x1.62 to x4.49

answer checks passed out of five, bare then with capsul
2 / 4
tool steps per question, bare then with capsul
2.4 / 4.4
output tokens per question, bare then with capsul
2,710 / 444
API-equivalent cost per question, bare then with capsul
$0.0812 / $0.0265

Its API list price is the lowest in the Claude group. That makes it a natural candidate for locating a function, explaining a test failure or making a small, well-scoped edit. A cheap model can still spend heavily if the agent reads many files before answering, so compare the input column rather than the prompt alone.

The measured interval matters particularly here. This row should not be read as proof that capsul saves input on every Haiku question. Look at the question-level results and answer checks before choosing it for a workflow that repeats often.

Question by question

Five one-line questions about a real TypeScript codebase, each asked 3 times per arm. Weighted input per question, mean of the repetitions.

QuestionBarecapsulLess inputCorrect runs, bare / capsul
Q157,70018,000x3.201 / 1
Q253,60024,800x2.161 / 3
Q318,20010,700x1.700 / 3
Q4118,00030,600x3.873 / 3
Q529,70012,200x2.433 / 3

List prices

ModelInputOutputCache readCache write 5 minCache write 1 hContext
Claude Haiku 4.51.005.000.101.252.00200,000
Claude Sonnet 52.0010.000.202.504.001,000,000
Claude Sonnet 4.63.0015.000.303.756.001,000,000
Claude Opus 4.65.0025.000.506.2510.001,000,000
Claude Opus 4.75.0025.000.506.2510.001,000,000
Claude Opus 4.85.0025.000.506.2510.001,000,000
Claude Opus 55.0025.000.506.2510.001,000,000
Claude Opus 5.54.0020.000.205.008.001,000,000
Claude Fable 510.0050.001.0012.5020.001,000,000
Claude Fable 5.110.0050.000.2512.5020.001,000,000
Anthropic Claude API list prices in US dollars per million tokens, checked September 23, 2026.

Anthropic list prices, checked September 23, 2026.

How this was measured

Two arms on the same model: Claude Code as anyone runs it, and the same question through capsul. The unit is cache-weighted input: fresh tokens at full weight, cache reads at a tenth. Campaign of September 22, 2026, 3 repetitions per question and arm.

30 / 30 cells · /benchmarks/2026-09-22-claude.json

Questions

How many tokens does Claude Haiku 4.5 use per question in Claude Code?

In the September 22, 2026 benchmark, a bare Claude Code session on Claude Haiku 4.5 sent 55,500 cache-weighted input tokens per one-line coding question on average. With capsul on the same model, 19,300.

How much does a coding question cost on Claude Haiku 4.5?

At Anthropic list prices, the bare agent's traffic came to $0.0812 per question and capsul's to $0.0265. A Claude subscription does not bill per token: read it as how much of a plan each question uses, compared with other models.

Is the saving on Claude Haiku 4.5 established?

Yes. The 90% interval runs from x1.62 to x4.49, entirely above one, on 30 measured cells.

Does capsul change Claude Haiku 4.5's answers?

It passed more of them: 2 of five answer checks bare, 4 of five with capsul. Five checks is a small sample: read it as no loss rather than a gain.