{"i":24,"slug":"haiku-retry-or-escalate","chart":{"id":"retry-escalate-cost-per-correct","title":"Expected list-price cost per correct answer, by retry policy (calculation)","subtitle":"Eight hard tasks, each equally likely. A failed try is paid for. Calls recorded in Claude Code","kind":"bar","unit":"usd","yLabel":"USD per correct answer","series":[{"name":"Cost per correct answer (calculation)","points":{"$k":["label","value","n","highlight"],"$r":[["Sonnet 5.5 low effort once (calculation)",0.01219,16,false],["Sonnet 5.5 every time (calculation)",0.01322,24,true],["Haiku, one retry, then Sonnet (calculation)",0.05664,48,false],["Haiku once (calculation)",0.06772,24,false],["Haiku up to 3 tries, then Sonnet (calculation)",0.07193,48,false]]}}],"note":"Calculation, not a run. These are median-input scenarios, not measured mean costs. Total modelled list-price cost of all tries over the 8 tasks, divided by the expected number of correct answers. Per-call cost is each task’s median call at list price (reported tokens × list price); the calls ran on a flat subscription. A try runs only when the earlier tries failed the validator, and tries are independent at each task’s observed pass rate. n is the number of recorded calls behind each bar. A calculation has no confidence interval; the sensitivity chart shows a range. Highlighted: the baseline, Sonnet 5.5 every time.","sourceIds":["calc-repricing","agent-provider-h2h-hard","agent-effort-ladder","price-anthropic"]}}