[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"article-openai-api-pricing-august-2026-token-costs-en":3,"article-related-openai-api-pricing-august-2026-token-costs-en":30,"series-tools-0eba42be-cb97-4711-9189-e7f4d0d9ffbe":75},{"id":4,"slug":5,"title":6,"content":7,"summary":8,"source":9,"source_url":10,"author":11,"image_url":12,"cover_image":12,"category":13,"language":14,"translated_content":11,"related_article_id":15,"keywords":16,"key_takeaways":23,"views":27,"created_at":28,"published_at":29,"topic_cluster_id":11},"0eba42be-cb97-4711-9189-e7f4d0d9ffbe","openai-api-pricing-august-2026-token-costs-en","OpenAI API Pricing Hits $0.05 to $180\u002FM Tokens","\u003Cp>\u003Ca href=\"\u002Ftag\u002Fopenai\">OpenAI\u003C\u002Fa> \u003Ca href=\"\u002Ftag\u002Fapi\">API\u003C\u002Fa> users now face a spread from $0.05 to $180 per million tokens, depending on model and meter. BenchLM’s August 7, 2026 pricing tracker says the newest GPT-5.6 tiers and older Pro models can differ by hundreds of times on the same workload.\u003C\u002Fp>\u003Cp data-speakable=\"summary\">BenchLM updated OpenAI API pricing for August 2026, covering \u003Ca href=\"\u002Ftag\u002Ftoken\">token\u003C\u002Fa> rates, cache discounts, batch pricing, and long-context charges.\u003C\u002Fp>\u003Ctable>\u003Cthead>\u003Ctr>\u003Cth>項目\u003C\u002Fth>\u003Cth>數值\u003C\u002Fth>\u003C\u002Ftr>\u003C\u002Fthead>\u003Ctbody>\u003Ctr>\u003Ctd>Last synced\u003C\u002Ftd>\u003Ctd>August 7, 2026\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Lowest listed input rate\u003C\u002Ftd>\u003Ctd>$0.05 per million tokens\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Highest listed output rate\u003C\u002Ftd>\u003Ctd>$180 per million tokens\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>GPT-5.6 Sol\u003C\u002Ftd>\u003Ctd>$5 input \u002F $30 output per million tokens\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>GPT-5.6 Terra\u003C\u002Ftd>\u003Ctd>$2 input \u002F $12 output per million tokens\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>GPT-5.6 Luna\u003C\u002Ftd>\u003Ctd>$0.20 input \u002F $1.20 output per million tokens\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Long-context GPT-5.6 Sol\u003C\u002Ftd>\u003Ctd>$10 input \u002F $45 output per million tokens\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Batch API discount\u003C\u002Ftd>\u003Ctd>50% off\u003C\u002Ftd>\u003C\u002Ftr>\u003C\u002Ftbody>\u003C\u002Ftable>\u003Ch2>What changed\u003C\u002Fh2>\u003Cp>BenchLM’s registry now reflects the GPT-5.6 family, which went GA on July 9, 2026. Sol, Terra, and Luna all ship with a 1.05M-token context window, while Terra and Luna saw price cuts on July 30: Terra fell 20% and Luna fell 80%.\u003C\u002Fp>\n\u003Cfigure class=\"my-6\">\u003Cimg src=\"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786300371726-ccpx.png\" alt=\"OpenAI API Pricing Hits $0.05 to $180\u002FM Tokens\" class=\"rounded-xl w-full\" loading=\"lazy\" \u002F>\u003C\u002Ffigure>\n\u003Cp>The pricing page also keeps older OpenAI families live, including GPT-5.5, GPT-5.4, GPT-5.1, GPT-4o, and the o-series. That matters because the cheapest listed options still start at GPT-5 nano’s $0.05 per million input tokens, while GPT-5.5 Pro and o1-pro sit at the top end of the chart.\u003C\u002Fp>\u003Cul>\u003Cli>GPT-5.6 Sol: $5 input \u002F $30 output per million tokens\u003C\u002Fli>\u003Cli>GPT-5.6 Terra: $2 input \u002F $12 output per million tokens\u003C\u002Fli>\u003Cli>GPT-5.6 Luna: $0.20 input \u002F $1.20 output per million tokens\u003C\u002Fli>\u003Cli>GPT-5.5 Pro: $30 input \u002F $180 output per million tokens\u003C\u002Fli>\u003Cli>Long-context meters raise flagship rates above short-context pricing\u003C\u002Fli>\u003C\u002Ful>\u003Ch2>Why it matters\u003C\u002Fh2>\u003Cp>For developers, the bill is now shaped by more than prompt length. Cached input bills at 10% of standard input rates, Batch API cuts both directions in half, and GPT-5.6 cache writes cost 1.25x standard input with a 30-minute minimum cache life.\u003C\u002Fp>\n\u003Cfigure class=\"my-6\">\u003Cimg src=\"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786300372225-7tb6.png\" alt=\"OpenAI API Pricing Hits $0.05 to $180\u002FM Tokens\" class=\"rounded-xl w-full\" loading=\"lazy\" \u002F>\u003C\u002Ffigure>\n\u003Cp>That means the same app can swing from cheap to expensive based on routing, prefix reuse, and whether it runs asynchronously. BenchLM says a cached and batched Sol workload can land near a quarter of list price, which makes model selection and request design part of cost control, not just ops cleanup.\u003C\u002Fp>\u003Cp>OpenAI’s API is still billed separately from \u003Ca href=\"\u002Ftag\u002Fchatgpt\">ChatGPT\u003C\u002Fa> plans, so a paid consumer subscription does not buy API credits. The practical question for teams is simple: do you need flagship quality, or can Terra, Luna, or an older model hit the target at a much lower rate?\u003C\u002Fp>\u003Cp>The takeaway is not just that OpenAI prices vary widely, but that the cheapest viable model and the request pattern around it now decide the real bill.\u003C\u002Fp>","BenchLM’s August 2026 tracker shows OpenAI API pricing from $0.05 to $180 per million tokens, plus cache, batch, and long-context meters.","benchlm.ai","https:\u002F\u002Fbenchlm.ai\u002Fopenai\u002Fapi-pricing",null,"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786300371726-ccpx.png","tools","en","0afe05ab-6675-498f-b179-531a8460290b",[17,18,19,20,21,22],"OpenAI","API pricing","tokens","Batch API","GPT-5.6","BenchLM",[24,25,26],"OpenAI API pricing now spans $0.05 to $180 per million tokens across models and meters.","GPT-5.6 Sol, Terra, and Luna share a 1.05M-token context window, but Terra and Luna were cut on July 30.","Cached input, Batch API, and long-context rates can change the final bill more than the headline model price.",0,"2026-08-09T18:32:26.984366+00:00","2026-08-09T18:32:26.978+00:00",{"tags":31,"relatedLang":34,"relatedPosts":38},[32],{"name":17,"slug":33},"openai",{"id":15,"slug":35,"title":36,"language":37},"openai-api-pricing-august-2026-token-costs-zh","OpenAI API 單價落差到 180 美元","zh",[39,45,51,57,63,69],{"id":40,"slug":41,"title":42,"cover_image":43,"image_url":43,"created_at":44,"category":13},"0aab53f7-2569-4c05-93f1-9bed53def12b","deepseek-codex-ai-coding-costs-en","DeepSeek in Codex Will Cut AI Coding Costs Hard","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786321971893-69cj.png","2026-08-10T00:32:33.297704+00:00",{"id":46,"slug":47,"title":48,"cover_image":49,"image_url":49,"created_at":50,"category":13},"0d0d262b-65b7-4388-82e5-bab51244f9c0","token-vs-word-chinese-tokenization-matters-en","Token vs. word: why Chinese tokenization still matters","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786320173873-ptb3.png","2026-08-10T00:02:32.1747+00:00",{"id":52,"slug":53,"title":54,"cover_image":55,"image_url":55,"created_at":56,"category":13},"d11db4f6-f2c8-4125-8d46-db91fd9c9121","usage-limits-chatgpt-enterprise-edu-controls-en","Usage limits belong in ChatGPT Enterprise and Edu controls","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786298573885-sm6r.png","2026-08-09T18:02:23.482195+00:00",{"id":58,"slug":59,"title":60,"cover_image":61,"image_url":61,"created_at":62,"category":13},"0ad2ef7c-aa71-4608-bd16-e0f27e9dd3de","prepare-for-gemini-3-5-pro-on-launch-day-en","Prepare for Gemini 3.5 Pro on launch day","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786276969432-cdi1.png","2026-08-09T12:02:26.327419+00:00",{"id":64,"slug":65,"title":66,"cover_image":67,"image_url":67,"created_at":68,"category":13},"34ace48c-c860-49c3-85ba-16ae03cf58b1","kitesurf-turns-workers-into-agent-browser-en","Kitesurf turns Workers into an agent browser","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786127613785-xbat.png","2026-08-07T18:32:57.093672+00:00",{"id":70,"slug":71,"title":72,"cover_image":73,"image_url":73,"created_at":74,"category":13},"8bcb444e-69ae-4fcc-a60d-592d1159af49","cuda-warps-memory-divergence-explained-en","CUDA warps turn GPU threads into one machine","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786064603311-wyof.png","2026-08-07T01:02:59.616194+00:00",[76,81,86,91,96,101,106,111,116,121],{"id":77,"slug":78,"title":79,"created_at":80},"8008f1a9-7a00-4bad-88c9-3eedc9c6b4b1","surepath-ai-mcp-policy-controls-en","SurePath AI's New MCP Policy Controls Enhance AI Security","2026-03-26T01:26:52.222015+00:00",{"id":82,"slug":83,"title":84,"created_at":85},"27e39a8f-b65d-4f7b-a875-859e2b210156","mcp-standard-ai-tools-2026-en","MCP Standard in 2026: Integrating AI Tools","2026-03-26T01:27:43.127519+00:00",{"id":87,"slug":88,"title":89,"created_at":90},"165f9a19-c92d-46ba-b3f0-7125f662921d","rag-2026-transforming-enterprise-ai-en","How RAG in 2026 is Transforming Enterprise AI","2026-03-26T01:28:11.485236+00:00",{"id":92,"slug":93,"title":94,"created_at":95},"6a2a8e6e-b956-49d8-be12-cc47bdc132b2","mastering-ai-prompts-2026-guide-en","Mastering AI Prompts: A 2026 Guide for Developers","2026-03-26T01:29:07.835148+00:00",{"id":97,"slug":98,"title":99,"created_at":100},"3ab2c67e-4664-4c67-a013-687a2f605814","garry-tan-open-sources-claude-code-toolkit-en","Garry Tan Open-Sources a Claude Code Toolkit","2026-03-26T08:26:20.245934+00:00",{"id":102,"slug":103,"title":104,"created_at":105},"66a7cbf8-7e76-41d4-9bbf-eaca9761bf69","github-ai-projects-to-watch-in-2026-en","20 GitHub AI Projects to Watch in 2026","2026-03-26T08:28:09.752027+00:00",{"id":107,"slug":108,"title":109,"created_at":110},"9f332fda-eace-448a-a292-2283951eee71","practical-github-guide-learning-ml-2026-en","A Practical GitHub Guide to Learning ML in 2026","2026-03-27T01:16:50.125678+00:00",{"id":112,"slug":113,"title":114,"created_at":115},"1b1f637d-0f4d-42bd-974b-07b53829144d","aiml-2026-student-ai-ml-lab-repo-review-en","AIML-2026 Is a Bare-Bones Student Lab Repo","2026-03-27T01:21:51.661231+00:00",{"id":117,"slug":118,"title":119,"created_at":120},"6d1bf3f6-e191-4d30-b55b-8a0722fa6afe","ai-trending-github-repos-and-research-feeds-en","AI Trending Tracks Repos and Research Feeds","2026-03-27T01:31:35.709532+00:00",{"id":122,"slug":123,"title":124,"created_at":125},"010539a1-4c3a-4bd3-937a-26616422ee0d","awesome-ai-for-science-research-tools-map-en","Awesome AI for Science Is Becoming a Real Research Map","2026-03-27T01:46:50.89513+00:00"]