OpenAI 發布 GPT-6 Sol 與 Luna:API 定價降 50%,效能基準公開

2026年9月23日 05:18
站內 AI 整理稿

OpenAI has released GPT-6 Sol and GPT-6 Luna, 2 new models in its GPT-6 family.They sit below GPT-6 Astra, which launched earlier this month.OpenAI trained both with methods similar to Astra’s.The aim is to bring Astra’s advances to faster, more affordable models.Deployable today?Yes.

Both models are live in the OpenAI API as gpt-6-sol and gpt-6-luna.They are API-only models, so there are no weights to self-host.Three tiers, one recipe The GPT-6 family now has 3 tiers.Astra is the top model for the hardest work.Sol targets complex coding and professional tasks at lower cost.

Luna targets fast, high-volume everyday work.OpenAI team states better caching and inference let it serve these models more cheaply.It is cutting Sol and Luna API prices by 50% against their GPT-5.6 promotional pricing.ModelInput (per 1M tokens)Output (per 1M tokens)GPT-6 Astra$10.00$50.

00GPT-6 Sol$2.00 (was $4)$10.00 (was $20)GPT-6 Luna$0.10 (was $0.20)$0.50 (was $1.20) One detail is worth noting.Luna’s output price falls from $1.20 to $0.50, a cut of about 58%, not 50%.Benchmarks: what OpenAI reports Professional work: On AutomationBench 1.0.6, Sol at xhigh effort scores 33.

2% at $0.27 per task.Claude Opus 5 at max effort scores 26.9% at 11.1x that cost.Low-effort Astra scores 30.3% at 3.9x Sol’s cost.Luna at high effort gains 5.4 points over its predecessor at 58% lower cost per task.On Agents’ Last Exam, Sol at max effort scores 56.4%.

That beats Claude Opus 5’s best score at 60% lower cost per task.Coding: On DeepSWE v1.1, Sol at max effort scores 68.8%.That is 1.1 points behind Claude Fable 5 at xhigh, at about 80% lower cost per task.Luna at max effort scores 66.6%, comparable to Opus 5 and Fable 5 at medium effort.

In those comparisons, Luna costs 93% less per task than Opus 5 and 96% less than Fable 5.On FrontierCode 1.1 Main, which grades whether code is ready to merge, Sol matches Claude Fable 5.1 at xhigh at much lower cost.Computer use: On OSWorld 2.0 offline, Sol at xhigh scores 60.5% versus 60.

3% for Opus 5 at medium.Sol’s cost per task is about 80% lower.Luna at max beats GPT-5.6 Sol at medium for 1/10 of the cost.Factuality: OpenAI’s internal test uses de-identified ChatGPT conversations where users flagged model errors.Sol makes about half as many mistakes as its predecessor.

Luna at higher effort matches GPT-5.6 Sol at about 1/100 of its cost.OpenAI also carried Astra’s communication style over.Expect clearer, slightly shorter answers with less jargon, especially in coding conversations.

Prompt caching for long-running agents Agents resend the same instructions, tools and history on every turn.GPT-6 ships an improved prompt caching system with higher cache hit rates by default.Cached input reads get discounts of up to 90%.

Eligible shared prefixes reused within a 30-minute window now qualify.New controls for developers: A Prompt Caching Dashboard tracks hit rates over time.A diagnostics tool explains misses, for example "reason": "toolschanged".Explicit breakpoints let you choose where a cached prefix ends.

Reasoning effort can change mid-conversation via configurationupdate without breaking cache.allowed_tools restricts callable tools while keeping definitions stable.Prewarming prepares known context before the first user request.The full prompt caching guide covers each pattern.

GitHub reports these changes cut the share of prompt tokens needing fresh processing by more than 50%, helping Copilot respond faster.Availability API: gpt-6-sol and gpt-6-luna.ChatGPT Work and Codex: Plus, Pro, Business, Enterprise and Edu users.Free and Go: Luna in the ChatGPT desktop app.

Not yet in Chat.The ChatGPT rollout is gradual through launch day.Interactive Explainer window.addEventListener('message',function(e){var d=e.data;if(d&&d.mtpEmbed==='gpt6-sol-luna'&&d.height){var f=document.getElementById('mtp-g6-frame');if(f&&e.source===f.contentWindow){f.style.height=d.

height+'px';}}}); Key Takeaways Sol costs $2/$10 and Luna $0.10/$0.50 per 1M tokens.Sol at xhigh beats Opus 5 max on AutomationBench at 9% of the cost.Luna scores 66.6% on DeepSWE v1.1, costing 93% less per task than Opus 5.Cached input reads get up to 90% off, with new cache controls.

Both are live in the API, ChatGPT Work and Codex; not yet in Chat.Check out the Technical Blog.All credit goes to the researcher of this project.Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?

now you can join us on telegram as well.Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?Connect with us The post OpenAI Releases GPT-6 Sol and Luna: 50% Cheaper API Pricing and Benchmarks appeared first on MarkTechPost.

Related

相關文章

量子位生成式AI

Jev vs Decitron:同為決策AI,為什麼不是一回事?

它來自TypeSafe AI。這支有OpenAI背景的團隊沒有繼續卷生成和推理,而是換了個方向:讓AI直接做判斷。他們甚至把口號直接寫成:Decisions, not strings(要決策,不要文本)。Jev的走紅,也讓“Decision”(決策)重新成為AI圈的熱門詞。

剛剛

Win11 驚現 AI 惡意軟件 ClosedQuorum:入侵後自主決策,專挖最有價值信息

該惡意軟件主要針對 Windows 11 平臺,會在成功入侵用戶電腦後調用谷歌 Gemini、DeepSeek、Qwen 和 Mistral 等 AI 模型自主決策,專門挖掘攻擊目標最有價值的信息。調用 Gemini、DeepSeek 等模型,入侵後自主決策Talos 深入分析後發現,ClosedQuorum 在成功入侵目標設備後,會藉助多個 AI 模型來決策攻擊方案,包括竊取瀏覽器登錄憑證、提取加密貨幣信息、持久化執行惡意軟件,以及向其他設備橫向散播惡意軟件。

剛剛
量子位生成式AI

阿里千問AI平臺全面升級模型服務、Agent服務、AI應用

阿里巴巴宣布千問AI平台全面升級,以推動Agent進入生產為核心,新增Agent服務與行業AI解決方案,並推出API優速模式、Agent Studio、千問AI座艙等能力。平台強調從消耗Token轉向交付結果,並預計至2032年算力規模將超過20GW,以支撐完整AI體系。Agent Studio提供企業級全棧服務,支援模型路由、託管運行與生態工具接入,同時強化模型服務的效能與成本優化。

剛剛
量子位生成式AI

GPT-6 Astra搓3D刷屏後,3D生成的競爭規則變了

GPT-6 Astra的出現讓3D生成行業競爭標準從「Production-Ready」轉向「Agent-Ready」,強調生成能力需能被AI Agent理解和調用。影眸Hyper3D推出Agentic Mode,能自主分析雜亂輸入、規劃建模路徑,並支援後續參數化調整、材質替換與動畫預設,同時開放MCP接口讓外部Agent串接3D生成流程。

剛剛