何夕2077生成式AI

深度求索發佈閃電模型更新

2026年8月1日 00:00

重點摘要

Ad Skip to content All Topics AI and society AI in practice AI research Frontier Radar Short News Read full article about: Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids Google Deepmind has introduced Gemini Robotics 2, which it calls its most advanced vision-language-action (VLA) model yet. VLA models combine image recognition, language processing, and action control to help robots operate in physical environments. Deepmind says the model can control systems ranging from tabletop arms to full-body humanoid robots. The company describes Gemini Robotics 2 as an "intelligence layer" for a new generation of adaptive robots. It can manage full-body movement, perform fine motor tasks, and coordinate multiple robots, according to Deepmind.

站內 AI 整理稿

Ad Skip to content All Topics AI and society AI in practice AI research Frontier Radar Short News Read full article about: Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids Google Deepmind has introduced Gemini Robotics 2, which it calls its most advanced vision-language-action (VLA) model yet.

VLA models combine image recognition, language processing, and action control to help robots operate in physical environments.Deepmind says the model can control systems ranging from tabletop arms to full-body humanoid robots.

The company describes Gemini Robotics 2 as an "intelligence layer" for a new generation of adaptive robots.It can manage full-body movement, perform fine motor tasks, and coordinate multiple robots, according to Deepmind.Developers can apply for early access through the waitlist.

Google Deepmind also introduced Gemini Robotics ER 2, a model designed for "embodied reasoning." The term refers to understanding the physical world and deciding which actions to take based on that information.ER 2 acts as a higher-level control system for robots and replaces Gemini Robotics ER 1.

6, released in April.The new model is available in Google AI Studio.

Comment Source: Gemini Robotics Read full article about: Thinking Machines bets on efficiency over size with its second model, Inkling Small Thinking Machines, the AI lab from former OpenAI CTO Mira Murati, has released Inkling Small.

According to Artificial Analysis, the open-weights reasoning model scores 40 on the Intelligence Index, one point below Inkling (41), with less than a third of the parameters (276 billion total, 12 billion active).AA says no open model of equal or smaller size scores higher.

Inkling Small beats its bigger sibling on several coding and reasoning tests, including Humanity's Last Exam (32% vs.30%) and GPQA Diamond (89% vs.87%).

It falls behind on agent-based tasks and factual knowledge but is far more token-efficient, averaging 24K output tokens per task compared to 45K for Deepseek V4 Flash and 78K for GPT-5.4 mini.

Mira Murati's Thinking Machines ships a smaller, more efficient reasoning model that punches above its weight.| Image: Artificial Analysis The model handles text, image, and speech inputs, has a 256K-token context window, and ships under Apache 2.0.

Weights are on Hugging Face, and users can fine-tune it in the browser via Tinker Playground.Thinking Machines positions its models as a foundation for fine-tuning with users' own data.Some see this as the next frontier in AI.

Comment Source: Thinking Machines | Artificial Analysis Ad Read full article about: New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost Deepseek has released V4 Flash "0731," a major upgrade to its budget AI model.

According to the Artificial Analysis Intelligence Index, the new version scores 50 points, ten more than the previous V4 Flash that launched in April 2026.That puts it just one point behind OpenAI's budget model GPT-5.

6 Luna, but it costs about 60 percent less per task, even after OpenAI's 80 percent price cut.A big reason for the gap is Deepseek's 98 percent cache discount, well above the industry-standard 90 percent.The model also uses 12 percent fewer tokens than its predecessor.

The Artificial Analysis Intelligence Index shows Deepseek V4 Flash "0731" scoring 50 points after its update, nearly matching OpenAI's GPT-5.6 Luna while claiming the top spot for price-to-performance ratio.

| Image: Artificial Analysis The model improves across every tested category compared to the previous version, with the biggest gains in agentic tasks.On GDPval, a benchmark designed to test models on complex real-world office work, it climbs from 1,189 to 1,559 Elo points.

It also hallucinates less often.The architecture stays the same: 284 billion total parameters, 13 billion active, with a one-million-token context window.The model weights are available under an MIT license on Hugging Face.

Comment Source: Artificial Analysis Ad Read full article about: EU pools up to €30 billion for AI gigafactories while US tech giants casually spend 20 times more The European Commission has opened bidding to build up to seven so-called AI gigafactories across Europe.

The goal is to sharply expand Europe's AI computing capacity.Up to 10 billion euros in EU and national funding is expected to draw at least 20 billion euros in private investment.

The facilities would give startups, companies, research institutions, and government agencies access to the infrastructure needed to train and run large AI models.Eighteen member states, including Germany and France, are taking part.

The Commission has also signed letters of intent with AMD, Nvidia, and Qualcomm to secure access to hardware.Applications are due November 12, 2026, with construction of the first facilities set to begin in 2027.The project is part of the EU's "AI Continent" strategy.For comparison, major U.S.

tech companies alone plan to spend more than $600 billion on data centers this year, and that figure keeps rising.Europe's total package of around 30 billion euros is roughly 20 times smaller.If all that computing power is actually needed, Europe's investment would be a drop in the bucket.

Comment Source: EU Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems Ad Aschenbrenner's AI thesis could be correct, his timing and leverage were not Leopold Aschenbrenner’s AI hedge fund Situational Awareness had to unload nearly its entire publicly traded portfolio to Ken Griffin’s Citadel after racking up heavy losses on leveraged AI stock positions.

Just days earlier, Aschenbrenner had reported a six-month return of 439 percent and pulled in fresh capital.Then margin calls forced the fire sale.Read full article Comment Ad Read full article about: OpenAI goes full China pricing mode with an 80 percent cut to its most affordable GPT-5.

6 model OpenAI is cutting GPT-5.6 Luna prices by 80 percent and Terra by 20 percent, effective July 30.Luna drops to $0.20 per million input tokens and $1.20 per million output tokens, while Terra falls to $2 and $12.Sol pricing stays the same.

OpenAI says Luna matches the performance of leading models from a year ago, but a task that cost a dollar with those models now runs about 6 cents on Luna, nearly nine times faster.All models are available through ChatGPT Work, Codex, and the OpenAI API.

OpenAI's smallest AI model, Luna, aims to dominate competitors on price-to-performance.| Image: OpenAI OpenAI says the cuts are possible because GPT-5.6 Sol made the company's own infrastructure more efficient.

The model allegedly optimized GPU software on its own, cutting deployment costs by 20 percent.It also improved token generation by more than 15 percent through speculative decoding.Growing price pressure across the AI market likely played a role too, especially from low-cost Chinese providers.

Microsoft is now openly promoting its own MAI models as cheaper alternatives to OpenAI.The price war could hurt the broader market if it slows revenue growth at frontier labs whose balance sheets are tied to massive infrastructure investments.

Comment Source: OpenAI Ex-OpenAI researcher bets $100 billion will flow into training data because scaling alone won't cut it Former OpenAI employee Andrew Ho and Cambridge researcher Adam Hunt see a growing problem with large language models.

Instead of becoming more versatile, the models are becoming more specialized, excelling at coding and math while stagnating or even regressing in other areas.

Ho is leaving OpenAI to start a company focused on specialized training data and predicts that AI labs will need to spend more than $100 billion on targeted data collection.

Read full article Comment Ad Language models can't spark scientific revolutions, but world models might Ad Read full article about: Microsoft AI bets on cheap specialist models instead of chasing the frontier Microsoft AI is making token efficiency a competitive focus, favoring small specialist models over general-purpose frontier models.

AI CEO Mustafa Suleyman writes that the industry has to weigh top performance against cost.Rather than one all-purpose model, the company trains compact models for single fields.

Its latest cybersecurity model MAI-Cyber-1-Flash tops the CyberGym benchmark by 12 percentage points over Anthropic's Mythos at half the cost, Suleyman says.But that result requires the MDASH system, which orchestrates several models and still routes hard tasks to OpenAI's reasoning models.

Microsoft also says MAI-Image-2.5-Flash cuts GPU costs by up to 84 percent compared with GPT-Image-2.Suleyman also wants swappable models that keep Microsoft from relying on one model family.Whether the small MAI models partly replacing OpenAI can match its performance remains doubtful.

Competition is moving from individual models to harnesses, the software that routes tasks and supplies context.Orchestrators send most work to cheaper specialists and reserve frontier models for hard cases.Anthropic modeled this approach for Claude Fable 5, while Sakana built Fugu around it.

Comment Source: Microsoft Load more BETA-TEST × BETA-TEST ×

Related

相關文章

MarkTechPost AI生成式AI

使用 NVIDIA NeMo Retriever、託管 NIM、LanceDB、重新排序與基於事實生成建立多模態 RAG 管線

在本教學中,我們將使用 NVIDIA NeMo Retriever 建立一個先進的多模態檢索增強生成管線。首先設定 Python 3.12 環境、安裝必要套件,並在無需 GPU 或外部 API 金鑰的情況下進行離線 PDF 文字提取。接著,我們透過託管的 NVIDIA NIM 端點來偵測頁面元素、提取表格、圖表與資訊圖形、產生稠密向量嵌入,並將處理後的內容儲存至 LanceDB。最後,我們實作了稠密檢索、視覺語言重新排序、後設資料過濾搜尋、附行內引用的基於事實回應生成,以及輕量級的 recall-at-k 評估,以驗證跨多模態文件內容的檢索品質。

41 分鐘前
MarkTechPost AI生成式AI

NVIDIA AI 推出 NOOA:將 AI 代理轉化為單一 Python 類別的物件導向框架

NVIDIA 實驗室開源了 NOOA(NVIDIA 物件導向代理),這是一個與模型無關的 Python 框架,用於建構 AI 代理。傳統的代理開發分散在提示模板、工具架構、回呼程式碼和工作流程圖中,而 NOOA 將所有這些整合到一個 Python 類別中:方法代表模型可採取的動作,欄位代表代理狀態,文件字串作為提示,型別註解則是執行時期強制執行的合約。主體為「...」的方法由 LLM 驅動的迴圈在執行時期完成,而具有正常主體的方法則保持確定性的 Python 程式碼。開發者與模型因此共享同一介面,使代理行為能像一般軟體一樣進行測試、追蹤、重構和版本控制。NVIDIA 報告在 SWE-bench Verified 上達到 82.2%,在 CyberGym L1 上達到 86.8%,平均 RHAE 為 85.1%。

1 小時前

六巨頭定AI插件新標準,撞臉Claude,Anthropic沒上桌

六大科技巨頭(AWS、Anysphere、GitHub、微軟、OpenAI、Vercel)聯合發布AI智能體插件統一開放規範Agent Plugins 1.0.0,旨在統一插件打包格式,減少開發者重複勞動。該規範的結構與Anthropic的Claude Code插件系統高度相似,但Anthropic並未參與制定,而是繼續經營自己的封閉生態。

3 小時前
鈦媒體生成式AI

DeepSeek重啟融資,三年市值對齊騰訊?

DeepSeek重啟第二輪融資,以5000億元人民幣估值尋求籌集80億美元,但網傳一份由小型醫藥私募發起的專項基金募資材料引發網友質疑,後經DeepSeek員工證實部分數據屬實。該公司近期宣布API大幅漲價,可能打破其以低價換規模的估值邏輯,面臨客戶流失風險。市場關注其能否從「價格屠夫」轉型為價值提供商,以及三年內市值能否對齊騰訊等巨頭。

4 小時前