Sarah Friar has outlined OpenAI's approach to AI infrastructure and pricing in a new company post, framing lower costs and greater model efficiency as central to what she calls "abundant intelligence."
OpenAI said it recently cut GPT-5.6 Luna pricing by 80 percent, now $0.20 per million input tokens and $1.20 per million output tokens, while GPT-5.6 Terra fell 20 percent to $2 and $12 respectively. The company also introduced a Fast mode for GPT-5.6 Sol, offering up to 2.5 times the processing speed at double the price, with intelligence unchanged.
On efficiency, OpenAI said GPT-5.6 Sol helped its engineering teams cut end-to-end model-serving costs by 20 percent and improved speculative decoding to lift token-generation efficiency by more than 15 percent. OpenAI also said changes to reasoning retention and context management raised the model's score on the ARC-AGI-3 benchmark from 13.3 percent to 38.3 percent while using six times fewer output tokens.
OpenAI said it has more than one billion active users and over two million businesses, with users sending roughly 50 percent more daily messages and using ChatGPT for twice as many task types six months after signup. OpenAI said agentic work through Codex now accounts for 99.8 percent of weekly output tokens, with Finance among the teams making agentic tools central to their work.
