Google AI Just Released Gemini 3.7 Flash: A Coding and Agent Model at $0.75/1M Input Tokens
Google has released Gemini 3.7 Flash, the newest model in its Flash tier, three weeks after Gemini 3.6 Flash.The model card describes it as a refinement of 3.6 Flash with algorithmic improvements to the core reasoning foundation — not a new pretraining run.
It accepts text, images, audio, and video across a 1M-token context window, returns up to 64K output tokens, and supports customizable thinking configurations that trade quality against cost and latency.The knowledge cutoff stays at March 2026.
The gains concentrate in three places: software engineering, document-heavy knowledge work, and web development.The sharper argument is price.Gemini 3.7 Flash ships at $0.75 per 1M input tokens and $3.75 per 1M output tokens — half the original 3.
6 Flash list rate, and roughly a third the blended cost of Claude Sonnet 5 or GPT-5.6 Terra.Is it Deployable?Yes, API and enterprise only.There are no open weights.
Access runs through hosted surfaces: the Gemini API and Google AI Studio, Google Antigravity, Android Studio, the Gemini Enterprise Agent Platform, and the Gemini Enterprise app.Consumers reach it through Gemini Spark on Google AI Pro and Ultra plans.
Company fit: Startups and mid-market teams gain the most, because the introductory price makes always-on agents affordable without a Pro-tier budget.Regulated enterprises get a governed path through Gemini Enterprise.
Teams with data-residency or air-gap requirements are excluded — there is nothing to self-host.Industries: Google’s own eval set points at legal, financial services, biosciences, and enterprise operations.The Harvey LAB-AA, GDP.pdf, and AutomationBench results are the tells.
Applications: Long-running coding agents, document-heavy back-office automation, UI generation from screenshots or design systems, and PDF-to-structured-data pipelines.The Benchmark Picture On FrontierCode 1.1 Main, which measures production code quality, Gemini 3.7 Flash scores 43.6% against 34.
4% for 3.6 Flash.On DeepSWE v1.1, a long-horizon software engineering eval, it reaches 65.3%.On WebDev Arena it posts an Elo of 1588 versus 1538, the top score in Google’s comparison table.Document and workflow results move further.GDP.pdf, an expert PDF comprehension eval, goes from 22.0% to 34.0%.
AutomationBench, a private enterprise workflow set, goes from 17.0% to 30.4% — ahead of both Claude Sonnet 5 at 10.7% and GPT-5.6 Terra at 23.6%.Long-context retrieval on GDM-MRCR v2 at 128k reaches 97.0%.GPT-5.6 Terra is ahead on DeepSWE (69.6%), Terminal-bench 2.1 (87.4%), Terminal-bench 3.0 (20.
8%), and OSWorld-2.0 (50.2%).On GDPval-AA v2 knowledge work, 3.7 Flash scores 1525 Elo against 1598 for Sonnet 5 and 1628 for Muse Spark 1.2.CharXiv Reasoning is a regression: 84.5% without tools, down from 85.2% for 3.6 Flash.On the Artificial Analysis Intelligence Index, 3.
7 Flash scores 56, against 57 for both GPT-5.6 Terra and Muse Spark 1.2.(function(){ var f=document.getElementById("mtp-g37-x7k2-f"); window.addEventListener("message",function(e){ if(!e.data||e.data.mtpFrame!=="g37flash")return; var h=parseInt(e.data.height,10); if(h&&h>200&&h<4000)f.style.
height=h+"px"; },false); })(); Pricing is the real argument Gemini 3.7 Flash lists at $0.75 per 1M input tokens and $3.75 per 1M output tokens.That rate is introductory and expires December 31, 2026; from January 1, 2027 it becomes $1.50 and $7.50.In Google's own table, Claude Sonnet 5 sits at $2.
00/$10.00 and GPT-5.6 Terra at $2.00/$12.00.At an 80/20 input-output mix, that is a blended $1.35 per 1M tokens today against $3.60 for Sonnet 5 and $4.00 for GPT-5.6 Terra.For teams running agents at volume, the intelligence-per-dollar gap is the reason to evaluate, not the individual eval wins.
Key Takeaways Gemini 3.7 Flash is a refinement of 3.6 Flash, not a new base model, shipped just three weeks later.Coding gains are real: FrontierCode 43.6% vs 34.4%, DeepSWE 65.3% vs 48.6%, WebDev Arena 1588 Elo.Price is the strongest claim — $0.75/$3.
75 per 1M until December 31, 2026, then it doubles.GPT-5.6 Terra still leads on terminal and computer-use agents; CharXiv is a small regression.API and enterprise only.No open weights, so no self-hosting or air-gapped deployment.Check out the Technical Details.
Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?now you can join us on telegram as well.Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?
Connect with us The post Google AI Just Released Gemini 3.7 Flash: A Coding and Agent Model at $0.75/1M Input Tokens appeared first on MarkTechPost.
Related
相關文章

阿里視頻大模型Wan3.0正式上線,行業評價“穩定、真實、有質感”
阿里巴巴影片生成大模型Wan3.0正式上線,單次可生成30秒影片,並首次支援doc、xls、ppt、pdf、md等文檔輸入。企業用戶普遍評價其「穩定、真實、有質感」,能穩定保持角色與場景一致性,並已進入短劇、影視、廣告等生產流程。即日起可於阿里雲百鍊、千問等平台體驗,標準版並推出限時7折優惠。

月之暗面第一代萬億參數多模態模型 Kimi K2.5 官宣月底結束服役
作者:歸瀧 責編:歸瀧 評論: 8 月 24 日消息,月之暗面 Kimi 官方微博今日宣佈,其第一代萬億參數多模態模型 —— Kimi K2.5 本月底即將結束服役。據此前報道,今年 1 月,月之暗面宣佈推出並開源了其最新的 Kimi K2.

消息稱知名 AI 研究員 Luke Metz 離開 OpenAI,加入 Meta 超級智能實驗室
作者:遠洋 責編:遠洋 評論: 感謝網友 華南吳彥祖 的線索投遞!8 月 24 日消息,據知情人士向 Axios 證實,知名 AI 研究員 Luke Metz 已加入 Meta 的超級智能實驗室(Superintelligence Labs)。

Anthropic 最強大模型 Fable 5 遇冷,企業用戶轉向更便宜 AI 產品
作者:遠洋 責編:遠洋 評論: 8 月 24 日消息,據英國《金融時報》報道,Anthropic 的美國客戶正在使用更便宜的替代品來替代其最強大的 AI 工具,這在其預計將實現有史以來規模最大的 IPO 之前,對其高支出的商業模式提出了質疑。

阿里雲視頻生成模型 Wan3.0 正式上線,支持單次生成 30 秒視頻、文檔輸入
作者:遠洋 責編:遠洋 評論: 8 月 24 日消息,阿里雲消息,今天,視頻生成模型 Wan3.0 正式上線。官方稱,Wan3.0 在生成時長、萬能創作、全能參考以及真實世界還原等維度全面升級,單次可生成 30 秒視頻,並首次支持 doc、xls、ppt、pdf、md 等文檔格式輸入,力求準確還原真實世界。
