{"i":16,"slug":"single-call-vs-agent-loop","chart":{"id":"agent-loop-total-time","title":"Total time per attempt: single call vs agent loop","subtitle":"Median per configuration; whiskers = fastest and slowest attempt","kind":"dot-range","unit":"seconds","yLabel":"Seconds","series":[{"name":"Total time per attempt","points":{"$k":["label","value","lo","hi","n"],"$r":[["Claude Haiku 4.5 (single call) · Claude Code",39.01,15.27,75.13,24],["Claude Haiku 4.5 (agent loop) · Claude Code",56.77,24.53,223.7,24],["Claude Sonnet 5.5 (single call) · Claude Code",7.75,2.26,34.79,24],["Claude Sonnet 5.5 (agent loop) · Claude Code",7.41,2.75,24.19,16],["GPT-6 Luna (single call) · Codex CLI",5.16,3.59,11.32,16],["GPT-6 Luna (agent loop) · Codex CLI",9.32,3.78,15.89,14]]}}],"note":"Whiskers are a range (fastest and slowest attempt), not a confidence interval. Agent loop: wall time of the whole session, failures and time-outs included. One host, one network. The Claude single-call cells are reference cells from the hard head-to-head (8 tasks × 3 repetitions), reused, not rerun. The reference cells ran in a different hour.","whisker":"minmax","sourceIds":["agent-agent-loop","agent-provider-h2h-hard"]}}