OpenAI, one of the world’s leading artificial intelligence companies, has announced significant cuts in the API prices of two of its GPT-5.6 models. The price of the cheapest model has been reduced by 80 percent, while the mid-range model has received a 20 percent price cut.
According to the company, improved infrastructure, more efficient systems and hardware optimization made these reductions possible.
Significant price cuts announced
According to media reports, leading artificial intelligence company OpenAI has significantly reduced the API prices of its latest GPT-5.6 models.
The changes came into effect on July 30, while the GPT-5.6 models were introduced for general availability on July 9. This means users and developers began benefiting from major price cuts just three weeks after the launch.
The company clarified that the reductions were made possible by improvements in model performance, inference systems, efficient hardware utilization, production software and context management.
Which models received price cuts?
Of the three major GPT-5.6 models, prices have been reduced for two, while the price of the most powerful model remains unchanged.
New pricing details
The biggest reduction has been given to GPT-5.6 Luna, the fastest and cheapest model in the family.
GPT-5.6 Luna
The price of this model has been reduced by 80 percent. The input price has fallen from $1 to $0.20 per million tokens, while the output price has been reduced from $6 to $1.20 per million tokens.
GPT-5.6 Terra
This model has received a 20 percent price reduction. The new price is $2 for input and $12 for output per million tokens, compared with $2.50 and $15 previously.
GPT-5.6 Sol
There has been no change in the price of this model. Its input price remains at $5 and its output price at $30 per million tokens.
Fast mode also introduced
The company has also introduced a new “Fast Mode” for GPT-5.6 Sol. Under the new feature, processing speed can increase by up to 2.5 times compared with the standard mode.
It will replace the previous “Priority Processing” option. Fast Mode will cost twice the standard API rate. While processing will be faster, there will be no change in the model’s intelligence or quality.
What will be the impact on users?
According to the company, there will be no change in the subscription fees or usage limits for ChatGPT and Codex.
However, institutional users and developers using the Terra and Luna models will consume fewer credits, allowing them to accomplish more work within the same budget.
The artificial intelligence industry has seen continuous improvements in model speed, efficiency and cost over the past several years.
Companies are not only introducing more intelligent models but are also working to reduce the cost of running them so that more businesses, software developers and organizations can benefit from these services.
According to experts, improvements in model training and inference systems are continuously reducing the cost per token, directly benefiting users through lower prices.
OpenAI’s decision to reduce prices so significantly just three weeks after launch appears to indicate that competition in the artificial intelligence market is rapidly intensifying. Lower prices are likely to encourage developers, startups and businesses to adopt advanced AI models.
The 80 percent reduction in particular indicates that the company wants to encourage large-scale use of its cheaper model. This could make the development of AI applications, chatbots, customer service systems, programming assistants and automated systems considerably more affordable than before.
On the other hand, keeping the price of the most powerful model unchanged appears to be part of a strategy under which organizations seeking the highest performance will continue to pay a premium, while lower-cost alternatives will be available for general use.
Also Read: 150-year-overdue book returned to Australian library