Daily Tech Briefing
AI 科技速览

每天 5 分钟内学习 AI。获取最新的人工智能新闻,理解其重要性,并学习如何将其应用于您的工作。

AI 快讯
VentureBeat ML · 2026/7/30 21:48:00
AI price wars: OpenAI cuts GPT-5.6 Luna prices by 80% as model competition shifts toward cost

AI price wars: OpenAI cuts GPT-5.6 Luna prices by 80% as model competition shifts toward cost

AI 中文解读
OpenAI 突然大幅降价,将 GPT-5.6 系列中最小的 Luna 模型价格砍掉了 80%,一场 AI 价格战正式打响。与此同时,Anthropic 和 Google 也刚发布了性能更强或成本更低的新模型,整个市场开始从拼“模型多聪明”转向拼“谁用起来更划算”。简单说,以前用一次 AI 模型可能要花不少钱,现在 Luna 的输入和输出加起来每百万 token 只要 1.40 美元,比很多竞争对手都便宜。OpenAI 还把中档模型 Terra 降价了 20%,同时给旗舰 Sol 模型加了个“极速模式”,速度能快 2.5 倍,但算力费也因此翻倍。创始人 Sam Altman 直接称这是“今天的大幅降价”。这场价格战对普通人来说是好消息:开发者用 AI 做产品的成本变低,意味着更多免费或便宜的应用会出现;普通用户用高级模型写文章、做翻译、处理数据的开销也会减少。未来我们可能看到更多 AI 工具像水电一样便宜又顺手,整个行业正在变得更亲民。
To quote an ancient Jedi Master "Begun, the AI price wars have!"OpenAI is sharply reducing the prices of two models in its GPT-5.6 frontier series, cutting GPT-5.6 Luna, the smallest and fastest model in the series, by 80% and GPT-5.6 Terra, the mid-tier model, by 20%, while adding a premium Fast mode for its flagship GPT-5.6 Sol model.The cuts place Luna much closer to the lowest-cost commercial models in the market and arrive just a few days after Anthropic released its highly performant Claude Opus 5 at the same price as Opus 4.8, and Google introduced Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, two rival models built around lower inference costs, faster execution and more efficient agent workloads. OpenAI is successfully undercutting Google's price per intelligence and attempting to sway Anthropic users, who may not mind paying more, with a speed boost. OpenAI says Luna will now cost $0.20 per million input tokens and $1.20 per million output tokens, for a combined input-plus-output price of $1.40 per million tokens. Terra will cost $2 per million input tokens and $12 per million output tokens, for a combined price of $14.Pricing for Sol Standard remains unchanged at $5 per million input tokens and $30 per million output tokens. OpenAI is also adding Sol Fast mode at twice the Standard price: $10 per million input tokens and $60 per million output tokens. The company says Fast mode delivers up to 2.5 times the throughput without changing the model’s underlying intelligence.OpenAI co-founder and CEO Sam Altman took to X to announce the changes as "major price cuts today."VentureBeat Frontier AI model API pricing comparisonModelInput ($/1M)Output ($/1M)Total ($/1M)SourceMiMo-V2.5 Flash$0.10$0.30$0.40Xiaomideepseek-v4-flash$0.14$0.28$0.42DeepSeekdeepseek-v4-pro$0.435$0.87$1.305DeepSeekGPT-5.6 Luna$0.20$1.20$1.40OpenAIMiniMax-M3$0.30$1.20$1.50MiniMaxLongCat-2.0 — limited-time promo$0.30$1.20$1.50LongCatGemini 3.1 Flash-Lite$0.25$1.50$1.75GoogleQwen3.7-Plus$0.40$1.60$2.00Alibaba CloudMiMo-V2.5$0.40$2.00$2.40XiaomiGemini 3.5 Flash-Lite$0.30$2.50$2.80GoogleLongCat-2.0 — standard$0.75$2.95$3.70LongCatMiMo-V2.5 Pro (≤256K)$1.00$3.00$4.00XiaomiGLM-5.2$1.40$4.40$5.80Z.aiGrok 4.5$2.00$6.00$8.00xAIMiMo-V2.5 Pro (>256K)$2.00$6.00$8.00XiaomiGemini 3.6 Flash$1.50$7.50$9.00GoogleQwen3.7-Max$2.50$7.50$10.00Alibaba CloudGemini 3.5 Flash$1.50$9.00$10.50GoogleGemini 3.1 Pro Preview (≤200K)$2.00$12.00$14.00GoogleGPT-5.6 Terra$2.00$12.00$14.00OpenAIGPT-5.4$2.50$15.00$17.50OpenAIKimi K3$3.00$15.00$18.00Moonshot AIGemini 3.1 Pro Preview (>200K)$4.00$18.00$22.00GoogleClaude Opus 5$5.00$25.00$30.00AnthropicGPT-5.5$5.00$30.00$35.00OpenAIGPT-5.5 Instant (chat-latest)$5.00$30.00$35.00OpenAISakana Fugu Ultra (≤272K)$5.00$30.00$35.00Sakana AIGPT-5.6 Sol — Standard mode$5.00$30.00$35.00OpenAIClaude Fable 5 / Claude Mythos 5$10.00$50.00$60.00AnthropicGPT-5.6 Sol — Fast mode$10.00$60.00$70.00OpenAIPricing is shown per one million tokens. Total cost is calculated as input price plus output price. Cached-input pricing is excluded to keep the comparison consistent across providers.OpenAI moves Luna into the low-cost tierThe most consequential change is the Luna price cut.When OpenAI introduced the GPT-5.6 series, Luna was priced at $1 per million input tokens and $6 per million output tokens, for a combined total of $7. The new pricing reduces that combined figure to $1.40.That places Luna below Google’s Gemini 3.5 Flash-Lite, which costs a combined $2.80 per million input and output tokens, and far below Gemini 3.6 Flash at $9. Luna also now costs less than OpenAI’s own GPT-5.4 and Terra models by a wide margin.It is not the cheapest model in the broader market. Xiaomi’s MiMo-V2.5 Flash, DeepSeek’s flash model and several other APIs remain less expensive on a pure token basis. But the reduction brings an OpenAI frontier-series model into direct competition with the market’s low-cost inference tier.OpenAI says the GPT-5.6 series represents its frontier mo
分享
阅读原文