多智能體做數學

2026年8月31日 00:00
站內 AI 整理稿

Computer Science > Artificial Intelligence arXiv:2608.23691 (cs) [Submitted on 24 Aug 2026] Title:Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment Authors:Stephen Chung, Wenyu Du, William J.

Wesley View a PDF of the paper titled Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment, by Stephen Chung and 2 other authors View PDF HTML (experimental) Abstract:We study autonomous mathematical discovery in the Station, an open-world multi-agent environment in which AI agents from different model families pursue a shared research goal without a central coordinator or scripted pipeline.

Agents choose their own research directions, conduct experiments, collaborate, and build a shared scientific literature.

Across 12 construction problems from the AlphaEvolve catalogue and two additional case studies, the Station obtained results novel relative to the prior literature on five problems: a new infinite family of finite-field Kakeya sets, new exact 604-point kissing configurations in dimension 11, new records for the discretized Kakeya needle and sign uncertainty problems, and a substantially improved lower bound for Erdős's minimum-overlap problem.

Agents also discovered novel infinite families for Book Ramsey numbers.Importantly, the agents produced not only numerical constructions but also theorems and analyses explaining how those constructions work, making the results more interpretable and easier for mathematicians to build upon.

We release all raw agent dialogues, proofs, and verification code, providing a transparent record of how these discoveries emerged.Comments: 38 pages, 12 figures, 3 tables.

Source code at this https URL and raw agent dialogues, proofs, and verification artifacts at this https URL Subjects: Artificial Intelligence (cs.AI); Discrete Mathematics (cs.DM); Multiagent Systems (cs.MA) MSC classes: 68T42 (Primary) 68T20, 05C55, 52C17 (Secondary) ACM classes: I.2.11; I.2.

3 Cite as: arXiv:2608.23691 [cs.AI] (or arXiv:2608.23691v1 [cs.AI] for this version) https://doi.org/10.48550/arXiv.2608.

23691 Focus to learn more arXiv-issued DOI via DataCite Submission history From: Stephen Chung [view email] [v1] Mon, 24 Aug 2026 18:00:03 UTC (1,585 KB) Full-text links: Access Paper: View a PDF of the paper titled Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment, by Stephen Chung and 2 other authorsView PDFHTML (experimental)TeX Source view license Current browse context: cs.

AI < prev | next > new | recent | 2026-08 Change to browse by: cs cs.DM cs.MA References & Citations NASA ADSGoogle Scholar Semantic Scholar export BibTeX citation Loading...BibTeX formatted citation × loading...

Data provided by: Bookmark Bibliographic Tools Bibliographic and Citation Tools Bibliographic Explorer Toggle Bibliographic Explorer (What is the Explorer?) Connected Papers Toggle Connected Papers (What is Connected Papers?) Litmaps Toggle Litmaps (What is Litmaps?) scite.

ai Toggle scite Smart Citations (What are Smart Citations?) Code, Data, Media Code, Data and Media Associated with this Article alphaXiv Toggle alphaXiv (What is alphaXiv?) Links to Code Toggle CatalyzeX Code Finder for Papers (What is CatalyzeX?) DagsHub Toggle DagsHub (What is DagsHub?

) GotitPub Toggle Gotit.pub (What is GotitPub?) Huggingface Toggle Hugging Face (What is Huggingface?) ScienceCast Toggle ScienceCast (What is ScienceCast?) Demos Demos Replicate Toggle Replicate (What is Replicate?) Spaces Toggle Hugging Face Spaces (What is Spaces?) Spaces Toggle TXYZ.

AI (What is TXYZ.AI?) Related Papers Recommenders and Search Tools Link to Influence Flower Influence Flower (What are Influence Flowers?) Core recommender toggle CORE Recommender (What is CORE?

) Author Venue Institution Topic About arXivLabs arXivLabs: experimental projects with community collaborators arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy.arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community?Learn more about arXivLabs.Which authors of this paper are endorsers?| Disable MathJax (What is MathJax?)

Related

相關文章

OpenAI被曝按噸囤Mac mini,幾萬臺全拿去訓AI了

新智元·2026年08月31日 11:53AI實驗室按噸囤Mac mini訓Agent,帶火蘋果Mac營收。等等,AI實驗室囤Mac mini,竟然是按「噸」囤的?!The Information剛剛爆料,OpenAI已經採購了數萬臺Mac mini和Mac Studio,到現在還在追著蘋果要貨。

剛剛

Grok Bot王炸登場,深度接入X,全自動監測全球最新動態

新智元·2026年08月31日 11:52全自動全球熱點監測最關鍵一環。Grok Bot 現在可以直接接入 X 了。它不只是「能讀 X 了」這麼簡單。你在 Grok Bot 裡關聯你的 X 賬號,系統自動幫你創建開發者賬號,付費用戶還直接送你一筆 X API 的調用額度。

剛剛

Cisco 給 9 萬員工配上個人 Agent:記住你的一切,還能替你跨系統辦事

思科(Cisco)近日在公司內部展開一項大規模的人工智慧部署計畫,一口氣為旗下多達九萬名員工配備專屬的個人 Agent。這款數位助理並非簡單的問答機器人,而是被賦予「記憶」能力,能夠記住每位員工的工作習慣、職務內容與日常需求,更關鍵的是,它具備跨系統執行任務的能力,可以直接代辦許多繁瑣的流程性工作。這項決策在企業軟體與 AI 應用領域投下震撼彈,也讓外界得以一窺大型科技公司如何將生成式 AI 真正落地到內部營運。 根據了解,這套個人 Agent 的核心設計理念在於「理解你的一切」。

剛剛
量子位AI Agent

OpenAI買幾萬臺Mac搞強化訓練!英偉達的活被蘋果搶了

OpenAI和Anthropic大量採購Mac mini與Mac Studio進行強化學習訓練,因其統一內存架構在AI任務上具優勢,帶動Mac季度銷售額成長近29%。蘋果意外在企業AI市場取得成功,但面臨供應短缺與英偉達推出競爭產品DGX Spark的挑戰。

剛剛
量子位AI Agent

全國第三,公司第二,“初創黑馬”靈犀智湧用ROSS Harness把機器人送進工業具身智能第一梯隊

靈犀智湧用一臺由Demo級本體組裝而成的機器人,成為工業場景賽除行業頭部企業外唯一獲獎的機器人公司。 第二屆世界人形機器人運動會的51個賽項,劃出了兩條清晰的評價體系:競技賽檢驗運動極限與協同能力,場景賽則直接對標真實崗位的作業標準。工業場景裝配上料崗正是後者中最貼近產線的賽項,評判標準直指工業級交付的核心指標:長時程穩定性、毫米級精度,以及擾動下的自主恢復能力。

剛剛
何夕2077AI Agent

Agent設計開源

Agent設計近期宣布開源,這項動態在開發者與設計工具社群中引發關注。根據公開訊息,其核心方向是讓 Builder 直接在原生設計工具中給出解決方案,將設計與工程之間的傳統鴻溝進一步收斂。值得注意的是,整個流程的產物是可直接運作的頁面,而非傳統意義上的靜態設計稿,這意味著設計階段的成果能更無縫地對接開發實作。 在技術實作層面,Agent 會直接修改 HTML 原始碼,而非僅輸出視覺規範或標註文件。這項設計背後的關鍵在於,整個系統由設計系統驅動真實組件,讓每一次調整都直接反映在實際運行的程式碼結構中。

4 小時前