[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"article-openai-cuts-gpt-56-prices-ai-bills-en":3,"article-related-openai-cuts-gpt-56-prices-ai-bills-en":29,"series-model-release-b3fd7185-d626-4e48-ae4d-d38170255e54":76},{"id":4,"slug":5,"title":6,"content":7,"summary":8,"source":9,"source_url":10,"author":11,"image_url":12,"cover_image":12,"category":13,"language":14,"translated_content":11,"related_article_id":15,"keywords":16,"key_takeaways":22,"views":26,"created_at":27,"published_at":28,"topic_cluster_id":11},"b3fd7185-d626-4e48-ae4d-d38170255e54","openai-cuts-gpt-56-prices-ai-bills-en","OpenAI Cuts GPT-5.6 Prices as AI Bills Climb","\u003Cp data-speakable=\"summary\">\u003Ca href=\"\u002Ftag\u002Fopenai\">OpenAI\u003C\u002Fa> cut prices for GPT-5.6 Luna and Terra as AI model pricing gets more competitive.\u003C\u002Fp>\u003Cp>OpenAI has cut the cost of using two of its newest models by as much as 80%, and the timing says a lot about where the AI business is headed. On July 30, CEO \u003Ca href=\"https:\u002F\u002Fopenai.com\" target=\"_blank\" rel=\"noopener\">Sam Altman\u003C\u002Fa> announced lower prices for \u003Ca href=\"https:\u002F\u002Fopenai.com\u002Findex\u002Fgpt-5-6\" target=\"_blank\" rel=\"noopener\">GPT-5.6 Luna\u003C\u002Fa> and \u003Ca href=\"https:\u002F\u002Fopenai.com\u002Findex\u002Fgpt-5-6\" target=\"_blank\" rel=\"noopener\">GPT-5.6 Terra\u003C\u002Fa>, while also adding a faster option for \u003Ca href=\"https:\u002F\u002Fopenai.com\u002Findex\u002Fgpt-5-6\" target=\"_blank\" rel=\"noopener\">GPT-5.6 Sol\u003C\u002Fa> in the API.\u003C\u002Fp>\u003Cp>The move is about more than a cheaper bill. It shows OpenAI wants buyers to think in terms of cost per useful output, not just raw model quality. For customers already watching token spend in \u003Ca href=\"https:\u002F\u002Fopenai.com\u002Fcodex\" target=\"_blank\" rel=\"noopener\">Codex\u003C\u002Fa> and \u003Ca href=\"https:\u002F\u002Fopenai.com\u002Fchatgpt\" target=\"_blank\" rel=\"noopener\">ChatGPT Work\u003C\u002Fa>, the new pricing changes the math immediately.\u003C\u002Fp>\u003Ctable>\u003Cthead>\u003Ctr>\u003Cth>Model\u003C\u002Fth>\u003Cth>Old pricing\u003C\u002Fth>\u003Cth>New pricing\u003C\u002Fth>\u003Cth>Change\u003C\u002Fth>\u003C\u002Ftr>\u003C\u002Fthead>\u003Ctbody>\u003Ctr>\u003Ctd>GPT-5.6 Luna\u003C\u002Ftd>\u003Ctd>Not stated\u003C\u002Ftd>\u003Ctd>$0.20 input \u002F $1.20 output per million tokens\u003C\u002Ftd>\u003Ctd>80% drop\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>GPT-5.6 Terra\u003C\u002Ftd>\u003Ctd>Not stated\u003C\u002Ftd>\u003Ctd>$2 input \u002F $12 output per million tokens\u003C\u002Ftd>\u003Ctd>20% drop\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>GPT-5.6 Sol Fast mode\u003C\u002Ftd>\u003Ctd>Standard mode only\u003C\u002Ftd>\u003Ctd>Up to 2.5x speed for 2x price\u003C\u002Ftd>\u003Ctd>New API option\u003C\u002Ftd>\u003C\u002Ftr>\u003C\u002Ftbody>\u003C\u002Ftable>\u003Ch2>OpenAI is pricing for usage, not hype\u003C\u002Fh2>\u003Cp>Altman wrote on X that OpenAI wants to offer the “best price\u002Fintelligence tradeoff at every level.” That phrasing matters because it shifts the conversation away from \u003Ca href=\"\u002Ftag\u002Fbenchmark\">benchmark\u003C\u002Fa> bragging and toward practical buying decisions.\u003C\u002Fp>\n\u003Cfigure class=\"my-6\">\u003Cimg src=\"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785542568883-mf9l.png\" alt=\"OpenAI Cuts GPT-5.6 Prices as AI Bills Climb\" class=\"rounded-xl w-full\" loading=\"lazy\" \u002F>\u003C\u002Ffigure>\n\u003Cp>OpenAI said Luna now costs 80% less to use, while Terra is 20% cheaper. It also said those lower prices flow through to paid subscriptions when customers use Codex and \u003Ca href=\"\u002Ftag\u002Fchatgpt\">ChatGPT\u003C\u002Fa> Work, which means the discount is not confined to the API.\u003C\u002Fp>\u003Cp>That matters for teams that build products on top of OpenAI models. Lower token costs can turn a feature that was barely economical into one that can be shipped at scale, especially when usage spikes and the bill grows faster than the user base.\u003C\u002Fp>\u003Cul>\u003Cli>Luna: $0.20 per million input tokens and $1.20 per million output tokens\u003C\u002Fli>\u003Cli>Terra: $2 per million input tokens and $12 per million output tokens\u003C\u002Fli>\u003Cli>Sol Fast mode: up to 2.5x speed for 2x the price\u003C\u002Fli>\u003Cli>Pricing changes apply to paid subscriptions in Codex and ChatGPT Work\u003C\u002Fli>\u003C\u002Ful>\u003Ch2>The pressure on token spending is real\u003C\u002Fh2>\u003Cp>EMARKETER senior analyst Jacob Bourne told \u003Ca href=\"https:\u002F\u002Fwww.businessinsider.com\" target=\"_blank\" rel=\"noopener\">Business Insider\u003C\u002Fa> that “the era of tokenmaxxing is over.” His point is simple: companies have learned how quickly AI costs can balloon without clear business value.\u003C\u002Fp>\u003Cblockquote>“Enterprises have figured out how easy it is to burn tokens without getting value back, and they're pushing back on those increasing AI bills.” — Jacob Bourne, EMARKETER senior analyst\u003C\u002Fblockquote>\u003Cp>That pushback is showing up across the market. The companies buying AI are no longer impressed by raw capability alone. They want predictable spend, clearer unit economics, and fewer surprises when usage ramps up.\u003C\u002Fp>\u003Cp>This is also why pricing has become such a hot topic for frontier model vendors. If the best models remain expensive to run, customers will keep splitting workloads across multiple providers, or they will reserve premium models for only the hardest tasks.\u003C\u002Fp>\u003Ch2>Efficiency has become the new selling point\u003C\u002Fh2>\u003Cp>OpenAI said the price cuts came from improvements across the model itself, the \u003Ca href=\"\u002Ftag\u002Finference\">inference\u003C\u002Fa> systems that run it, and the agentic harness that connects the model to tools and context. In plain English, the company is saying it got better at turning compute into useful tokens.\u003C\u002Fp>\n\u003Cfigure class=\"my-6\">\u003Cimg src=\"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785542570041-y0wa.png\" alt=\"OpenAI Cuts GPT-5.6 Prices as AI Bills Climb\" class=\"rounded-xl w-full\" loading=\"lazy\" \u002F>\u003C\u002Ffigure>\n\u003Cp>The company added that better routing keeps hardware productive, optimized production software generates tokens more efficiently, and smarter context management helps agents avoid repeating work they already completed. That is the real story here: cost reductions are being treated as an engineering problem, not a marketing one.\u003C\u002Fp>\u003Cp>\u003Ca href=\"\u002Ftag\u002Fmicrosoft\">Microsoft\u003C\u002Fa> CEO \u003Ca href=\"https:\u002F\u002Fwww.microsoft.com\" target=\"_blank\" rel=\"noopener\">Satya Nadella\u003C\u002Fa> made a similar point during an earnings call on Wednesday, saying cost efficiency was central to \u003Ca href=\"https:\u002F\u002Fwww.microsoft.com\u002Fen-us\u002Fai\" target=\"_blank\" rel=\"noopener\">MAI-Thinking-1\u003C\u002Fa>.\u003C\u002Fp>\u003Cp>He said, “We are building a new model system where the harness, context, memory, and action space are separate from any one model family, thereby moving the frontier on the cost to outcome curve.” That is a very Microsoft way of saying the same thing OpenAI is now emphasizing: the wrapper around the model matters almost as much as the model itself.\u003C\u002Fp>\u003Cul>\u003Cli>OpenAI said the 5.6 series rolled out about three weeks earlier\u003C\u002Fli>\u003Cli>GPT-5.6 Sol was not included in the price-cut announcement\u003C\u002Fli>\u003Cli>OpenAI paused a broader rollout at the U.S. government's request before the 5.6 launch\u003C\u002Fli>\u003Cli>Moonshot AI's open-weight \u003Ca href=\"https:\u002F\u002Fwww.moonshot.ai\" target=\"_blank\" rel=\"noopener\">Kimi K3\u003C\u002Fa> has increased price pressure on closed-model vendors\u003C\u002Fli>\u003C\u002Ful>\u003Ch2>What this means for buyers and rivals\u003C\u002Fh2>\u003Cp>Arun Chandrasekaran, a distinguished vice president analyst at \u003Ca href=\"https:\u002F\u002Fwww.gartner.com\" target=\"_blank\" rel=\"noopener\">Gartner\u003C\u002Fa>, told Business Insider that the price cuts give buyers more room to negotiate with frontier AI labs. He said pricing and commercial terms have been notoriously hard to change until now.\u003C\u002Fp>\u003Cp>That is a meaningful shift for procurement teams. If OpenAI lowers prices this visibly, other vendors will have to answer with either cheaper access, stronger enterprise terms, or more explicit proof that their models save money elsewhere in the workflow.\u003C\u002Fp>\u003Cp>The competitive pressure is coming from multiple directions. \u003Ca href=\"https:\u002F\u002Fwww.anthropic.com\" target=\"_blank\" rel=\"noopener\">Anthropic\u003C\u002Fa> is trying to balance subscription and usage pricing against limited compute, while \u003Ca href=\"https:\u002F\u002Fcloud.google.com\u002Fai\" target=\"_blank\" rel=\"noopener\">Google\u003C\u002Fa> has been talking up model efficiency. OpenAI is now telling the market that better economics can be part of the product story, not a concession made after the fact.\u003C\u002Fp>\u003Cp>OpenAI has also filed confidentially for an IPO, which makes this price war more interesting. Public-market investors tend to ask harder questions about margins, infrastructure spend, and whether growth is buying enough revenue to justify the compute bill.\u003C\u002Fp>\u003Cp>The next test is whether these lower prices stay in place long enough to reset buyer expectations. If they do, AI procurement teams will start treating token costs the way cloud teams treat storage and bandwidth: a line item to optimize, not a mystery to accept.\u003C\u002Fp>\u003Ch2>OpenAI’s new pricing sets a marker\u003C\u002Fh2>\u003Cp>OpenAI is telling the market that model quality alone is no longer enough to win enterprise budgets. The companies that can make strong models cheaper to run will have an easier time keeping customers, especially as teams compare every token against actual business output.\u003C\u002Fp>\u003Cp>The obvious question now is whether rivals will match these cuts or try to outdo OpenAI on speed, context handling, and enterprise controls. Either way, the next pricing update from a frontier lab will matter less as a headline and more as a signal of who can still make the economics work.\u003C\u002Fp>","OpenAI slashed prices for GPT-5.6 Luna and Terra, signaling a sharper fight on AI cost efficiency.","www.businessinsider.com","https:\u002F\u002Fwww.businessinsider.com\u002Fopenai-price-cuts-gpt-terra-luna-2026-7",null,"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785542568883-mf9l.png","model-release","en","39170c12-7e99-4fb8-aebc-d4f155953b6f",[17,18,19,20,21],"OpenAI","Sam Altman","GPT-5.6","AI pricing","token costs",[23,24,25],"OpenAI cut GPT-5.6 Luna prices by 80% and Terra by 20%.","The discounts also apply to paid usage in Codex and ChatGPT Work.","The move intensifies pressure on AI vendors to prove better cost efficiency.",1,"2026-08-01T00:02:28.341618+00:00","2026-08-01T00:02:28.332+00:00",{"tags":30,"relatedLang":35,"relatedPosts":39},[31,33],{"name":17,"slug":32},"openai",{"name":18,"slug":34},"sam-altman",{"id":15,"slug":36,"title":37,"language":38},"openai-cuts-gpt-56-prices-ai-bills-zh","OpenAI 降價 GPT-5.6，AI 成本戰升溫","zh",[40,46,52,58,64,70],{"id":41,"slug":42,"title":43,"cover_image":44,"image_url":44,"created_at":45,"category":13},"2fee41e9-10d2-4755-9777-081159b6e609","opus-5-premium-ai-becoming-commodity-en","Opus 5 proves premium AI is becoming a commodity","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785499372952-zhqg.png","2026-07-31T12:02:29.544844+00:00",{"id":47,"slug":48,"title":49,"cover_image":50,"image_url":50,"created_at":51,"category":13},"5ff2c0fd-3296-48d3-8c74-acdadf2d5605","openai-free-gpt56-access-scientists-en","OpenAI Gives Scientists Free GPT-5.6 Access","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785434595001-84cv.png","2026-07-30T18:02:38.762824+00:00",{"id":53,"slug":54,"title":55,"cover_image":56,"image_url":56,"created_at":57,"category":13},"b5c87cdc-dec6-41ad-b5af-ef649d069859","google-gemini-36-flash-35-lite-pro-missing-en","Google ships Gemini 3.6 Flash and 3.5 Lite","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785144797179-t7qi.png","2026-07-27T09:32:53.415721+00:00",{"id":59,"slug":60,"title":61,"cover_image":62,"image_url":62,"created_at":63,"category":13},"63d81bb8-cd12-4273-864f-584d9d7db6d1","kimi-k3-forces-silicon-valley-to-pick-sides-en","Kimi K3 Is Forcing Silicon Valley to Pick Sides","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785139418320-5fll.png","2026-07-27T08:03:09.991025+00:00",{"id":65,"slug":66,"title":67,"cover_image":68,"image_url":68,"created_at":69,"category":13},"43cd3860-e32d-4585-94c3-c9de55a8dad9","opus-5-fewer-refusals-ship-faster-en","Opus 5 lets you ship with fewer refusals","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1785045795303-vwbt.png","2026-07-26T06:02:43.598776+00:00",{"id":71,"slug":72,"title":73,"cover_image":74,"image_url":74,"created_at":75,"category":13},"7c168597-37ea-4dcd-b25b-8cb0e73532d2","claude-opus-5-undercuts-fable-5-price-en","Claude Opus 5 undercuts Fable 5 on price","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1784982778159-kwk7.png","2026-07-25T12:32:33.412584+00:00",[77,82,87,92,97,102,107,112,117,122],{"id":78,"slug":79,"title":80,"created_at":81},"d4cffde7-9b50-4cc7-bb68-8bc9e3b15477","nvidia-rubin-ai-supercomputer-en","NVIDIA Unveils Rubin: A Leap in AI Supercomputing","2026-03-25T16:24:35.155565+00:00",{"id":83,"slug":84,"title":85,"created_at":86},"eab919b9-fbac-4048-89fc-afad6749ccef","google-gemini-ai-innovations-2026-en","Google's AI Leap with Gemini Innovations in 2026","2026-03-25T16:27:18.841838+00:00",{"id":88,"slug":89,"title":90,"created_at":91},"5f5cfc67-3384-4816-a8f6-19e44d90113d","gap-google-gemini-ai-checkout-en","Gap Teams Up with Google Gemini for AI-Driven Checkout","2026-03-25T16:27:46.483272+00:00",{"id":93,"slug":94,"title":95,"created_at":96},"f6d04567-47f6-49ec-804c-52e61ab91225","ai-model-release-wave-march-2026-en","Navigating the AI Model Release Wave of March 2026","2026-03-25T16:28:45.409716+00:00",{"id":98,"slug":99,"title":100,"created_at":101},"895c150c-569e-4fdf-939d-dade785c990e","small-language-models-transform-ai-en","Small Language Models: Llama 3.2 and Phi-3 Transform AI","2026-03-25T16:30:26.688313+00:00",{"id":103,"slug":104,"title":105,"created_at":106},"38eb1d26-d961-4fd3-ae12-9c4089680f5f","midjourney-v8-alpha-features-pricing-en","Midjourney V8 Alpha: A Deep Dive into Its Features and Pricing","2026-03-26T01:25:36.387587+00:00",{"id":108,"slug":109,"title":110,"created_at":111},"bf36bb9e-3444-4fb8-ab19-0df6bc9d8271","rag-2026-indispensable-ai-bridge-en","RAG in 2026: The Indispensable AI Bridge","2026-03-26T01:28:34.472046+00:00",{"id":113,"slug":114,"title":115,"created_at":116},"60881d6d-2310-44ef-b1fb-7f98e9dd2f0e","xiaomi-mimo-trio-agents-robots-voice-en","Xiaomi’s MiMo trio targets agents, robots, and voice","2026-03-28T03:05:08.899895+00:00",{"id":118,"slug":119,"title":120,"created_at":121},"f063d8d1-41d1-4de4-8ebc-6c40511b9369","xiaomi-mimo-v2-pro-1t-moe-agents-en","Xiaomi MiMo-V2-Pro: 1T MoE Model for Agents","2026-03-28T03:06:19.238032+00:00",{"id":123,"slug":124,"title":125,"created_at":126},"a1379e9a-6785-4ff5-9b0a-8cff55f8264f","cursor-composer-2-started-from-kimi-en","Cursor’s Composer 2 started from Kimi","2026-03-28T03:11:59.132398+00:00"]