/ capsul models

Every model, measured

What each model sends per question in Claude Code and Codex CLI, with and without capsul. Every figure comes from a dataset you can download.

Cache-weighted input tokens per one-line coding question, mean of the repetitions. Computed from the published datasets: Claude campaign of September 22, 2026, Codex campaign of September 4, 2026.

Claude models, in Claude Code

Cache-weighted input tokens per one-line coding question, mean of the repetitions.
ModelBare agentWith capsulLess inputAnswer checks
Claude Haiku 4.555,50019,300x2.882 / 4
Claude Sonnet 554,20018,900x2.864 / 4
Claude Sonnet 4.631,40016,600x1.894 / 4
Claude Opus 4.645,70016,400x2.784 / 4
Claude Opus 4.738,70020,100x1.925 / 4
Claude Opus 4.865,80013,800x4.755 / 4
Claude Opus 545,30016,200x2.804 / 4
Claude Opus 5.542,30015,900x2.663 / 4
Claude Fable 541,40017,600x2.354 / 4
Claude Fable 5.149,70023,400x2.123 / 4

GPT models, in Codex CLI

Cache-weighted input tokens per one-line coding question, mean of the repetitions.
ModelBare agentWith capsulLess inputAnswer checks
gpt-5.6-sol33,10014,700x2.263 / 3
gpt-5.6-terra24,00011,100x2.163 / 4
gpt-5.6-luna32,10013,600x2.364 / 3
gpt-5.547,2008,130x5.814 / 4
gpt-5.431,8008,220x3.864 / 4
gpt-5.4-mini42,30012,100x3.493 / 4