The Fine Print Behind the '20% Off' Headline
OpenAI's own framing of the cut as "over 20%" undersells what's actually happening to the more expensive side of the bill. Input tokens for GPT-5.6 Sol fell exactly 20%, from $5 to $4 per million tokens, but output tokens - the ones that cost six times more per token to begin with - dropped 33%, from $30 to $20 per million [1]. Cached input tokens, the discounted rate OpenAI charges for context a model has already processed once, fell from $0.50 to $0.40 per million [1]. Because output tokens are the expensive half of almost every real workload, the products that benefit most are the ones that generate a lot of text rather than just read it: coding agents drafting whole files, customer-support bots writing full responses, and agentic tools like Codex that loop through many turns of model output [2]. A chatbot that mostly reads a prompt and returns a short answer barely notices the difference; an agent that writes thousands of tokens of code or tool calls sees close to a third off its heaviest cost line.



