The Real Story Isn't Performance - It's Margin Compression
Grok 4.6 ships at $2 per million input tokens and $6 per million output tokens below the 200K-token threshold, against $5/$30 for GPT-5.6 Sol and $5/$25 for Claude Opus 5 [1][2]. That is not a rounding difference - investor Gavin Baker quantified it as roughly Claude Fable 5 Max's performance at an 85 percent discount, 80 percent cheaper on input tokens and 88 percent cheaper on output [3]. The pressure lands squarely on OpenAI and Anthropic's inference margins in the run-up to their planned IPOs, and market commentary has already flagged the knock-on risk to Microsoft's equity stakes in both companies as cost-to-outcome becomes the real competitive battleground rather than raw intelligence scores [3].



