GPT-5.6 Luna Is Now 80% Cheaper
OpenAI has cut API prices for two of its three GPT-5.6 models. Luna, its fastest and most affordable model, now costs 80% less, and Terra, the balanced model for everyday work, costs 20% less.
Transcript
Starting today, OpenAI cut the price of GPT 5.6 Luna, its fastest model, by eighty percent.
Luna now costs twenty cents per million input tokens and one dollar twenty cents per million output. Terra costs twenty percent less, at two dollars input and twelve dollars output.
OpenAI says the gains come from better models and serving systems. Within a human led process, Sol rewrote production kernels, cutting the cost of serving the model by twenty percent.
Sol pricing is unchanged, so the frontier model did not get cheaper. The new Fast mode runs up to 2.5 times faster at twice the price, with no added intelligence.
Inboxsmith helps small businesses handle calls and messages so nothing gets missed. Please like and subscribe for more news.
Sources
Every claim in this video comes from the top ranking coverage of this topic. The claims and where each one came from:
- Starting today, GPT-5.6 Luna, our fastest and most affordable model, will cost 80% less, while GPT-5.6 Terra, our balanced model for everyday work, will cost 20% less.(OpenAI's official announcement)
- Starting July 30, API pricing is $2 per million input tokens and $12 per million output tokens for Terra, and $0.20 per million input tokens and $1.20 per million output tokens for Luna.(OpenAI's official announcement)
- Our efficiency edge comes from improving the models, the inference systems that run them, and the agentic harness that connects them to tools and context.(OpenAI's official announcement)
- Within a human-led process, Sol autonomously rewrote and optimized production kernels, designed and ran hundreds of experiments to improve token generation, and monitored training, intervening when problems arose.(OpenAI's official announcement)
- The kernel work helped reduce the end-to-end cost of serving the model by 20%, while its experiments increased token-generation efficiency by more than 15%.(OpenAI's official announcement)
- Sol pricing remains unchanged.(OpenAI's official announcement)
- We're also introducing Fast mode in the API, which replaces our Priority Processing offering.(OpenAI's official announcement)
- For GPT-5.6 Sol, Fast mode now delivers up to 2.5x faster speeds than Standard processing at twice the price, with no change in intelligence.(OpenAI's official announcement)
