SpaceXAI 發布 Grok 4.7:更大的基礎模型,價格與 Grok 4.6 同為 2/6 美元

2026年9月22日 04:10
站內 AI 整理稿

SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work.Grok 4.7 is built on a larger base model and a longer reinforcement learning run.It still ships at the same price and speed as Grok 4.6.Is it deployable?Yes, as a hosted model.You can call grok-4.

7 today through the xAI API, Cursor, Grok Build, OpenRouter, Vercel, and Cloudflare.What Changed Under the Hood SpaceXAI lists 4 changes over Grok 4.6: A new, larger base model: Grok 4.7 does not reuse the Grok 4.6 base.

A longer RL run on harder tasks: The task mix is weighted toward problems that take many hours to complete.Better self-verification and long-context handling: The company says the model checks its own work more carefully.

Native Grok Bot harness support: It was trained to understand the Grok Bot harness for conversational and knowledge work.The developer docs list the API specs: PropertyValueModel namegrok-4.

7Context window500,000 tokensKnowledge cutoffMay 2026ModalitiesText and image input, text outputReasoning effortlow, medium, high (default), xhighAPIsResponses API, Chat CompletionsToolsFunction calling, web search, X search, code execution Benchmarks The launch table compares Grok 4.

7 at xHigh effort with Grok 4.6 High, GPT-5.6 Sol Max, and Fable 5.1 Max.The Grok 4.7 DeepSWE score was run at high effort.All scores are vendor-reported.BenchmarkGrok 4.7 xHighGrok 4.6 HighGPT-5.6 Sol MaxFable 5.1 MaxInput price ($/M)$2$2$4$10Output price ($/M)$6$6$20$50CursorBench 4.046.3%40.4%41.

7%51.8%DeepSWE v1.171.0%65.2%72.7%70.0%EEBench64.0%53.0%39.4%56.4%AA Briefcase v1.11,6571,5461,4871,678Terminal-Bench 4.038.0%20.3%37.3%57.9%Harvey Legal Agent Benchmark19.6%15.8%2.5%6.7%HealthBench Professional56.7%48.5%60.5%62.1% High effort Grok 4.7 improves on Grok 4.6 in every row.

The largest jump is on Terminal-Bench 4.0, from 20.3% to 38.0%.EEBench rose 11 points to 64.0%, the top score in the table.On Harvey’s legal agent benchmark, Grok 4.7 scored 19.6% against 6.7% for Fable 5.1 Max.Grok 4.7 does not lead across the board.Fable 5.

1 Max tops 4 of 7 benchmarks, including a 57.9% Terminal-Bench score.GPT-5.6 Sol Max holds the top DeepSWE v1.1 result at 72.7%.Price is the other axis.Fable 5.1 Max costs 5x more on input and about 8.3x more on output.GPT-5.6 Sol Max costs 2x more on input and about 3.3x more on output.

On a CursorBench 4.0 cost-per-task chart, SpaceXAI places Grok 4.7 at the frontier in price-performance.On GDPval, which tests professional knowledge work, Grok 4.7 xhigh scored 1,695 Elo.That is up from 1,605 for Grok 4.6 high.Fable 5.1 max leads at 1,735, and GPT-6 Astra max scored 1,542.

SpaceXAI also says Grok 4.7 is better at creating documents and presentations.Safety and Cybersecurity Grok 4.7 ships with an entirely new safeguard stack.SpaceXAI calls it the strongest model it has tested on refusals and jailbreak resistance.It topped LatchBio’s biosafety benchmark at 62.4%.

On HackerBench v0.3, SpaceXAI’s own benchmark for risky and malicious cyber tasks, the model let 3.3% of risky dual-use prompts through.The company says it rarely blocks legitimate security work.

Select cybersecurity partners now get invite-only access to its red-team capabilities for defense research.Pricing and Availability Grok 4.7 costs $2 per million input tokens and $6 per million output tokens.It is available in Cursor on all plans and is the default model in Grok Build.

It is also served through the Grok API, OpenRouter, Vercel, and Cloudflare.Grok 4.7 Fast is the same model on faster infrastructure, with twice the output speed at twice the price.The docs say it runs only in Cursor and Grok Build, not on the public xAI API.

It is also excluded from Grok Build’s free tier.A US regional endpoint at https://us.api.x.ai/v1 keeps inference in the United States at a 10% premium.SpaceXAI recommends setting a promptcachekey for reliable cache hits.

Copy CodeCopiedUse a different Browserimport os from xaisdk import Client from xaisdk.chat import user client = Client(apikey=os.getenv("XAIAPI_KEY")) chat = client.chat.create(model="grok-4.7") chat.append(user("Explain this repo.")) print(chat.sample().

content) Interactive Explainer (function(){var f=document.getElementById("mtp-qwen-image-21");window.addEventListener("message",function(e){if(!f||e.source!==f.contentWindow)return;var d=e.data;if(d&&typeof d.mtpQwenImage21==="number"){f.style.height=d.

mtpQwenImage21+"px";}});})(); Key Takeaways Grok 4.7 keeps Grok 4.6 pricing at $2 input and $6 output per million tokens.It leads the launch table on EEBench (64.0%) and Harvey Legal Agent Benchmark (19.6%).Fable 5.1 Max still leads on 4 of 7 benchmarks, including Terminal-Bench 4.0.

The API offers a 500,000-token context window and 4 reasoning levels up to xhigh.Only 3.3% of risky dual-use prompts passed on SpaceXAI’s HackerBench v0.3.Check out the announcement, docs, and X post.All credit goes to the researcher of this project.

Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?now you can join us on telegram as well.Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?

Connect with us The post SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6 appeared first on MarkTechPost.

Related

相關文章

IT之家模型更新

曝阿里試水千問平板 QwenBook:定位原生智能體電腦,主打辦公、打通 WPS

作者:汪淼 責編:汪淼 評論: 感謝網友 HH_KK 的線索投遞!9 月 22 日消息,據《晚點 LatePost》今日報道,阿里雲旗下無影團隊正在研發一款名叫 QwenBook 的 AI 設備,定位是原生智能體電腦,但產品形態更接近一臺“千問平板”,產品定義和硬件部分由團隊自研,無影過去做雲電腦積累的系統和端雲能力,也被帶進了這個項目。

剛剛
IT之家模型更新

北京大學材料科學與工程學院特聘研究員竇錦虎評價 MiMo-V2.6-Pro:達到訓練有素博士研究人員水平

作者:歸瀧 責編:歸瀧 評論: 感謝網友 Autumn_Dream、野原小哀 的線索投遞!9 月 22 日消息,小米今日凌晨正式發佈並開源了全新的 Xiaomi MiMo-V2.6 系列,包含 Pro 與 Flash 兩個原生全模態模型。小米稱這是其探索 RSI(遞歸自我改進)路徑的關鍵一步,通過規模化擴展強化學習算力,讓模型在持續的探索與反饋中拓展智能邊界。

剛剛
鈦媒體模型更新

阿里巴巴吳泳銘:未來模型將擴展至5-10T參數規模,到2032年運營的數據中心規模超過20GW | 鈦快訊

TechPulse2026.09.22 11:14 · 來自浙江全文1656字00:00 / 05:27未來機器思考總量將是人類的1000倍,而今天這一比例不到3%,至少還有幾萬倍的增長空間。 9月22日,在2026杭州雲棲大會上,阿里巴巴集團CEO吳泳銘表示,機器正在成為思考的主力,「思考」將像工業革命中的「動力」一樣,成為規模化供給的商品。未來機器思考總量將是人類的1000倍,而今天這一比例不到3%,至少還有幾萬倍的增長空間。吳泳銘認為,今天機器動力已經承擔了99.

剛剛