OpenAI Slashes API Prices for Luna, Terra Models Amid Intense Competition

OpenAI is introducing a Fast mode for GPT-5.6 Sol that replaces Priority Processing; existing API requests marked as priority will automatically use Fast mode, delivering up to 2.5x faster responses at twice the standard price.
OpenAI reports a 20% reduction in the end-to-end cost of serving GPT-5.6 Sol, driven by improvements across models, inference systems and agent workflows, with token-generation efficiency improving by more than 15%.
GPT-5.6 Luna is marketed as the fastest and most affordable in the lineup, with near-frontier-class performance and execution that is described as nearly nine times faster for high-volume tasks.
OpenAI highlights that Luna targets high-volume workloads such as routine coding, document processing, classification and background agent tasks, underscoring its use-case focus beyond general AI tasks.
OpenAI slashed the price of its GPT-5.6 Luna model by 80% on July 30, dropping the cost from $1 to just $0.20 per million input tokens, according to eWeek. Output tokens fell from $6 to $1.20 per million. The company also cut prices for its mid-tier GPT-5.6 Terra model by 20%, to $2 per million input tokens and $12 per million output tokens.
The cuts come as cheap Chinese AI models put growing pressure on OpenAI to compete on price, Cryptopolitan noted. OpenAI says better software and infrastructure — not just market pressure — made the reductions possible. The flagship GPT-5.6 Sol model keeps its current price, but gets a new speed upgrade.
OpenAI markets Luna as the fastest and cheapest model in its GPT-5.6 lineup. The company says Luna runs nearly nine times faster than other models for high-volume tasks, according to Fone Arena. It is built for routine jobs like coding, document sorting, and background agent tasks — not complex reasoning.
The 80% price cut makes Luna one of the most affordable frontier-class AI models available via API. Developers building apps that need to process millions of documents or run constant background checks stand to save the most. InfoWorld noted the cuts also reduce how many usage credits Luna consumes inside ChatGPT Work and Codex subscriptions, boosting value for enterprise users.
OpenAI is replacing its old Priority Processing option for GPT-5.6 Sol with a new "Fast mode." Fast mode delivers responses up to 2.5 times quicker than standard Sol, according to Fone Arena. The catch: it costs twice the standard price. Any existing API requests already marked as priority will switch to Fast mode automatically.
Sol's base price stays the same for now. The Fast mode addition gives developers a clear choice — pay more for speed, or stick with standard pricing for slower but cheaper results. This tiered approach lets OpenAI keep Sol profitable while still offering competitive options lower in the lineup.
OpenAI says it cut the end-to-end cost of serving GPT-5.6 Sol by 20% through improvements to its models, inference systems, and agent workflows. Token-generation efficiency improved by more than 15%, the company reported. These gains ripple down to Luna and Terra, making lower prices possible without hurting margins.
InfoWorld described the cuts as part of OpenAI's push to make AI tools more affordable for high-volume workloads. Better infrastructure means each token costs less to generate and serve. OpenAI says these savings are being passed on to developers rather than kept as profit — at least for now.
The timing of these cuts is not random. Chinese open-weight AI models have grown fast in popularity, largely because they are cheap or free to run, according to Yahoo Finance. OpenAI faces pressure from both Chinese rivals and open-source alternatives that undercut its pricing on routine tasks.
Cryptopolitan framed the Luna price cut as a direct response to that competitive squeeze. Luna dropped from $1/$6 to $0.20/$1.20 per million tokens — an 80% slash that signals OpenAI is serious about defending its share of the high-volume, cost-sensitive developer market. The broader AI pricing war shows no signs of slowing down.
Publishers
19
Articles
6
Reach
25