{"i":4,"slug":"hard-model-head-to-head","chart":{"id":"hard-h2h-frontier","title":"Quality vs cost frontier on hard tasks","subtitle":"Strict pass rate against list-price cost per strict pass","kind":"scatter","unit":"rate","xLabel":"USD per strict pass (list-price calculation)","yLabel":"Strict pass rate","series":[{"name":"Claude Code","points":{"$k":["label","x","value","n","highlight"],"$r":[["Claude Sonnet 5.5 · Claude Code",0.01435,1,24,true],["Claude Opus 5.5 · Claude Code",0.02824,1,24,false],["Claude Opus 5.5 (high) · Claude Code",0.03337,1,24,false],["Claude Fable 5.1 · Claude Code",0.09331,1,24,false],["Claude Haiku 4.5 · Claude Code",0.0672,0.4583,24,false]]}},{"name":"Codex CLI","points":[{"label":"GPT-6.1 Sol (medium) · Codex CLI","x":0.02564,"value":1,"n":16,"highlight":false},{"label":"GPT-6.1 Sol (high) · Codex CLI","x":0.01514,"value":1,"n":16,"highlight":false}]}],"note":"Upper-left is better. Highlighted points are on the frontier: no other configuration passes at least as often for at most the same cost per pass. Frontier: Claude Sonnet 5.5 · Claude Code. Costs are calculations from tokens. Pass rates with their 95% intervals are in the pass-rate chart.","sourceIds":["agent-provider-h2h-hard","calc-repricing","price-anthropic","price-openai"]}}