{"i":14,"slug":"cli-model-latency-tokens","chart":{"id":"cli-vs-api-prompt-overhead","title":"Hidden prompt: input tokens for the same one-line request","subtitle":"Reported input tokens, matched cohort","kind":"bar","unit":"tokens","yLabel":"Input tokens per call","series":[{"name":"Input tokens","points":{"$k":["label","value","n"],"$r":[["OpenAI API · GPT-6 Luna · none",17,5],["OpenAI API · GPT-6.1 Sol · low",17,5],["OpenAI API · GPT-6.1 Sol · high",17,5],["Codex CLI · GPT-6 Luna · none",18859,5],["Codex CLI · GPT-6.1 Sol · low",19551,5],["Codex CLI · GPT-6.1 Sol · high",19555,5]]}}],"note":"The CLI wraps every request in its own system prompt and tool context; the bare API sends only the request. Part of the CLI input is served from cache.","sourceIds":["agent-provider-explorer"]}}