多智能體系統算出開放數學新結果

2026年10月6日 00:00
站內 AI 整理稿

Computer Science > Artificial Intelligence arXiv:2609.

40324 (cs) [Submitted on 30 Sep 2026 (v1), last revised 1 Oct 2026 (this version, v2)] Title:Cogentic: Multi-Agent Orchestration for Automated Proof Discovery Authors:Yang Cai, Vineet Gupta, Yanchen Jiang, Christopher Liaw, Aranyak Mehta, Grigoris Velegkas, Di Wang View a PDF of the paper titled Cogentic: Multi-Agent Orchestration for Automated Proof Discovery, by Yang Cai and 6 other authors View PDF HTML (experimental) Abstract:We present Cogentic, a multi-agent harness for automated proof discovery on open research problems.

While frontier language models can generate strong mathematical ideas in a single shot, single-shot generation is often insufficient for open problems that require exploring multiple competing conjectures, overcoming subtle technical obstructions, and retaining intermediate progress over a long horizon.

Cogentic addresses these challenges through an iterative prove-verify loop in which an orchestrator allocates a population of independent provers across distinct proof directions, subjects their output to adversarial verification by several specialized components, and promotes confirmed intermediate results into a persistent verified ledger that later rounds build on.

The harness is designed to be able to solve research-level math and theoretical computer science problems.Using either Gemini 3.

1 Pro or an early version of Gemini 4 Argon as the base model, Cogentic produced novel results on open problems across online learning, auction theory, and mechanism design.Each result was independently verified by domain experts and is developed in full in companion papers.

We list these results, and new ones as they are verified, at this https URL .Subjects: Artificial Intelligence (cs.AI); Computer Science and Game Theory (cs.GT) Cite as: arXiv:2609.40324 [cs.AI] (or arXiv:2609.40324v2 [cs.AI] for this version) https://doi.org/10.48550/arXiv.2609.

40324 Focus to learn more arXiv-issued DOI via DataCite Submission history From: Yanchen Jiang [view email] [v1] Wed, 30 Sep 2026 17:55:22 UTC (212 KB) [v2] Thu, 1 Oct 2026 21:13:54 UTC (213 KB) Full-text links: Access Paper: View a PDF of the paper titled Cogentic: Multi-Agent Orchestration for Automated Proof Discovery, by Yang Cai and 6 other authorsView PDFHTML (experimental)TeX Source view license Current browse context: cs.

AI < prev | next > new | recent | 2026-09 Change to browse by: cs cs.GT References & Citations NASA ADSGoogle Scholar Semantic Scholar export BibTeX citation Loading...BibTeX formatted citation × loading...

Data provided by: Bookmark Bibliographic Tools Bibliographic and Citation Tools Bibliographic Explorer Toggle Bibliographic Explorer (What is the Explorer?) Connected Papers Toggle Connected Papers (What is Connected Papers?) Litmaps Toggle Litmaps (What is Litmaps?) scite.

ai Toggle scite Smart Citations (What are Smart Citations?) Code, Data, Media Code, Data and Media Associated with this Article alphaXiv Toggle alphaXiv (What is alphaXiv?) Links to Code Toggle CatalyzeX Code Finder for Papers (What is CatalyzeX?) DagsHub Toggle DagsHub (What is DagsHub?

) GotitPub Toggle Gotit.pub (What is GotitPub?) Huggingface Toggle Hugging Face (What is Huggingface?) ScienceCast Toggle ScienceCast (What is ScienceCast?) Demos Demos Replicate Toggle Replicate (What is Replicate?) Spaces Toggle Hugging Face Spaces (What is Spaces?) Spaces Toggle TXYZ.

AI (What is TXYZ.AI?) Related Papers Recommenders and Search Tools Link to Influence Flower Influence Flower (What are Influence Flowers?) Core recommender toggle CORE Recommender (What is CORE?

) Author Venue Institution Topic About arXivLabs arXivLabs: experimental projects with community collaborators arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy.arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community?Learn more about arXivLabs.Which authors of this paper are endorsers?| Disable MathJax (What is MathJax?)

Related

相關文章

鈦媒體AI Agent

中國等不來Muse

窮奇觀察2026.10.04 11:19 · 來自浙江全文11574字三個小循環,整合不出一個大循環 中國用戶等不來Muse。不是技術不夠,不是動作不快,是這道題根本不存在。Muse解的是"打通",美國沒有小循環,反壟斷拆完牆,瀏覽器是萬能鑰匙,一個Agent能替用戶敲開所有門。中國要解的是"整合"。字節、騰訊、阿里,三家莊園自給自足,門都是自家的,不需要外部鑰匙。題不一樣,抄得了作業本,抄不了題目。Muse的地基先看Muse自己有多猛。9月2日,Meta發佈Muse模型Spark 1.

1 天前
IT之家AI Agent

曝美國陸軍著手組建自主系統司令部,推動機器人進入技術未來戰爭

作者:清源 責編:清源 評論: 感謝網友 咩咩洋 的線索投遞! 10 月 3 日消息,據美媒 Axios 獲得的一份備忘錄,美國陸軍著手組建新的自主系統司令部,同時任命專門負責採購的官員,加快智能裝備採購和列裝。這些部署由代理陸軍部長亞當 · 特爾下達。就在當地時間 1 日,美國國防部長皮特 · 赫格塞思公佈了 Meridian 和 Agincourt 兩個項目,重點都是推動機器人技術進入未來戰爭。這份題為《陸軍自主化舉措》的備忘錄提出多項要求:組建陸軍未來與自主系統司令部。

2 天前
量子位AI Agent

Jev估值100億美元!創始人Diogo Almeida回答一切

o Almeida,頂著一頭新染的紅髮閃亮登場了!(doge Diogo在最新一期Latent Space訪談中坦言: 公開benchmark極其容易被操縱,即便開發者沒有主動作弊,最終也可能被榜單牽著走。 所以相比於一張通用榜單,他更願意相信長期積累的產品體驗與信任: 直到你把模型放進自己的工作流,並針對那個工作流去評估、去測量,才能判斷它在真正重要的流程中表現如何。

2 天前