Same-task comparison · Cost and waiting

Coding with Haiku: weigh cost and waiting together

Bito gives the same coding work to five Claude configurations. Session cost, elapsed time and implementation checks shape the choice, alongside Haiku's long-prompt price boundary.

This continues our Haiku and everyday model-cost readings, using Bito's October 9 report and official prices checked October 11. Figures are BitShovel diagrams of the author's data; we did not rerun the coding tasks.

Session cost and time are plotted separately for five Claude Code configurations; the table retains the original blended scores. Each unit has its own zero baseline, with no winner marked.
BitShovel's editorial diagram of Bito's October 9 table. USD costs use the author's session token counts and October 8 API rates; time is elapsed per session. This is neither a vendor image nor a BitShovel test. · Open full-size image
In this articleContents5

中文

Similar cost, different waiting

Across Bito's six main coding tasks, Haiku medium costs about $0.31 per session and takes 4.8 minutes, versus Sonnet medium's $0.64 and 3.7 minutes. Haiku high reaches $0.62 and 7.3 minutes. Its 79% versus 77% is within the author's roughly three-point run-to-run noise. Lower spending and earlier delivery are separate considerations.

Bito author-run tests · Blended session values across six main tasks · Reported October 9, checked October 11. Scores are percentages of the grading maximum, not pass rates. Costs use October 8 rates; effort settings and compute usage differ. Roughly three-point score noise does not support ranking a two-point gap.
Model / Claude Code effortScore / % of maximumUSD / sessionMinutes / session
Opus 5.5 · medium (test default)85%$2.718.5
Sonnet 5.5 · high82%$1.197.6
Haiku 5.5 · high79%$0.627.3
Sonnet 5.5 · medium (test default)77%$0.643.7
Haiku 5.5 · medium (test default)76%$0.314.8
Sources and further reading

中文

Looking up a setting differs from implementing and checking

The authors separate one-shot requests, short multi-turn tasks and long sessions. Haiku high approaches Opus medium on quick code lookups, while implementation and final checks widen the gap. Our reading is to identify the required endpoint: finding a clue is useful, but delivering a working change also requires checks. A blended percentage cannot decide which outcome is sufficient.

Sources and further reading

中文

A growing session can cross into a higher price tier

Official rates apply per request. At up to 100,000 prompt tokens, Haiku costs $0.10/$0.50 per million input/output tokens; above that, $0.50/$2.50. Cache reads and writes count toward prompt length. A cache hit does not remove the threshold. Assess growing sessions by material per request, call count and total cost. The table's usage-based API costs do not price a Claude subscription.

Sources and further reading

中文

Read the task comparison with its method

Bito uses fixed prompts from its services, a pinned Claude Code version and code-checked answer keys. Opus 5.5 grades five dimensions while also competing. A companion article describes a usual four sessions per task; the main table lacks per-cell counts, the version number, full prompts and failure logs. Medium and high differ in compute budget. This informs selection without making every result independently reproducible or comparable to our AA index and Pi pass rates.

Sources and further reading

中文

Connect savings to the outcome you need

This continues the everyday-cost question: start with bounded lookups and explanations, then observe retries and completeness during multi-turn changes. For sustained implementation and review, judge a more expensive configuration alongside waiting, rework and human checking. Compare candidates on your own shared task and delivery standard, rather than inheriting a precise rank from one study.

Sources and further reading