Same Model, Steeper Price
OpenAI rolled out Ultrafast for GPT-6.1 Sol on October 8, 2026 [1], but it doesn't change what the model knows or can do - it's the same gpt-6.1-sol model id, selected simply by setting service_tier to ultrafast [2]. What OpenAI is selling is the gap between generated tokens, packaged as a standalone premium: $12 per million input tokens and $60 per million output tokens, six times Standard's $2/$10 [3], with long-context requests above 272K tokens priced even higher at $24/$90 and cached input available at $0.60 per million tokens against $15 for cache writes [8]. The same 6x ratio shows up one tier up: GPT-6 Astra Ultrafast costs $60 input and $300 output, exactly six times Astra Standard [13]. Ultrafast for Sol also ships with both US and EU data residency, a step beyond Astra's Ultrafast tier, which remains US-only [9], and Amazon made it available on Bedrock around the same time [7]- evidence OpenAI is building inference speed into a durable, cross-platform pricing ladder rather than a one-off launch perk.



