Fly Language Model (FLM) Wires the Full Fruit Fly Connectome Into a Frozen 1.2B LLM, and Its Own Controls Show the Wiring Does Not Help

2026年9月12日 18:51
站內 AI 整理稿

The Fly Language Model (FLM) is a public chatbot that couples the complete retained MaleCNS v1.0 fruit fly connectome to a frozen LiquidAI LFM2.5-1.2B-Instruct backbone.

The developer who created the FLM calls it the world’s first Fly Language Model, built on an architecture called GPF (Generative Pre-trained Fly).

It does not use the GPF label, explicitly disclaims being the first connectome language model, and reports that a parameter-matched control without the fly graph performs slightly better.Deployable: Yes, locally.The nftechie/flm repo is MIT-licensed and runs on Python 3.

12 (macOS or Linux, MPS, CUDA, or CPU) with no API key.What was actually built The system is a reservoir computer bolted onto a language model.All 166,700 retained nodes and 25,582,938 directed edges of the MaleCNS graph participate.

The graph, the backbone, and the random input and output projections are all fixed.Only a 278,528-parameter readout is trained, which is about 0.0238% of the 1,170,340,608 backbone parameters.At each token, a fixed Gaussian projection compresses the 2,048-dimensional token embedding to 128 channels.

Each reservoir node receives one channel with a random sign.The whole graph then updates with x = tanh(W(0.6x + 0.4Bc)), where W holds incoming-normalized anatomical contact counts.

States are pooled into 128 bins, passed through two trained bias-free matrices (U at 128 by 128, V at 2,048 by 128), and projected through the frozen vocabulary head as a bounded residual added to the backbone logits.The residual is capped at an RMS of 0.25 across vocabulary coordinates.

The results On a freshly frozen set of 32 SmolTalk everyday-conversation dialogues (1,236 target tokens), three fit seeds gave: ConditionNLL (nats/token)Frozen backbone1.381995Fly readout1.359816 ± 0.000110Direct-input readout1.359328 ± 0.000108Relabeled, no refit1.381265 ± 0.000802No edges1.

381995 The fly readout improved on the backbone by 0.0222 nats per token (perplexity 3.98 to 3.90).But a direct-input control, which feeds the same 128-channel token projection straight into an identical readout with no graph, did better in all 3 seeds by 0.000488 nats per token.

The paired bootstrap interval (+0.00000502 to +0.00104) does not support a fly-specific gain.Two other controls matter.Setting W to zero removes the residual exactly, reproducing the backbone’s per-token losses, so the graph verifiably participates.

Relabeling node identities without retraining returns NLL near baseline, which shows the readout depends on its learned interface alignment, not that fly topology beats random wiring.The research report also proves the recurrence contracts initial-state differences by at most 0.6 per token.

After 10 tokens that bound is 0.00605; after 20 it is 0.0000366.Piling in 166,700 cells does not buy long memory.Context still comes from the backbone.

Prior work and the ‘first’ claim The research report cites ngxson/fly-hf, an earlier prototype that used a 49,393-cell central-brain subset of MaleCNS as a reservoir trained on TinyStories without a pretrained backbone, and states plainly that it makes no claim to be the first connectome-based language model.

FLM’s distinction is scale (the full retained graph) and the frozen-backbone design that keeps the source of language competence identifiable.Interactive explainer (function(){var f=document.getElementById("mtp-flm-explainer-frame");window.addEventListener("message",function(e){if(e.data&&e.data.

mtpFlm==="resize"&&f)f.style.height=e.data.h+"px";});})(); Key Takeaways Full 166,700-node fly connectome drives a frozen LFM2.5-1.2B; only 278,528 parameters train.Fly readout cuts NLL by 0.0222 nats/token, but a no-graph control beats it in every seed.

Disconnection zeroes the residual exactly; relabeling breaks it.The graph participates, it does not win.State forgets at 0.6 per token, so the connectome adds no long-range memory.MIT code runs locally on Python 3.12; study artifacts stay private, so results are not independently reproducible yet.

Check out the Paper, GitHub repo, and live demo.Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?now you can join us on telegram as well.

Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?Connect with us The post Fly Language Model (FLM) Wires the Full Fruit Fly Connectome Into a Frozen 1.

2B LLM, and Its Own Controls Show the Wiring Does Not Help appeared first on MarkTechPost.

Related

相關文章

量子位生成式AI

無問芯穹與華環電子簽署戰略合作,共同探索國產異構算力AI基礎設施新方向

無問芯穹與華環電子簽署戰略合作協議,雙方將結合各自在AI軟體平台、網路通信與硬體研發的優勢,共同探索國產異構算力基礎設施的協同方案。此次合作聚焦於智算中心解決方案及「Token工廠」新模式,目標是推動計算、網路與AI原生基礎設施深度融合,為AI規模化應用提供高效穩定的支撐。

9 分鐘前
IT之家生成式AI

優步全球範圍裁員 10%,被裁員工稱 AI 已大舉滲透日常工作

作者:清源 責編:清源 評論: 9 月 18 日消息,據《商業內幕》今天(18 日)晚間報道,在優步(Uber),AI 已經滲透到員工工作的許多環節,從回答 Slack 裡的內部問題,到替乘客行程中聯繫客服時收到的消息撰寫回復。6 名近期遭裁員的員工透露,過去幾個月,AI 在工作中的使用範圍明顯擴大,其中一些人甚至會通過提示詞讓 AI 完成相當一部分任務。

3 小時前
鈦媒體生成式AI

月之暗面遞表之後,Kimi 的成色要被驗算三遍

舒澤品牌手記2026.09.18 18:16 · 來自浙江全文4982字00:00 / 14:05Anthropic 的 30 萬次指控,會成為招股書的第幾頁?文 | 舒澤品牌手記9月17日,月之暗面發佈了一套金融行業解決方案。按官方披露,中信建投、中金公司、易方達等數十家金融機構已經在用 Kimi 處理投研建模、風險排查和盡調材料——研究人員把管理層報表、審計報告和盡調文件交給 Kimi,拿回一份可以繼續調整假設的 Excel 模型。同一天,深圳商報記者就港股上市進展、股東架構調整等事項向月之暗面發去採訪函。

5 小時前

Calibre上手 AI 互動寫作:電子書管理器搖身變成"文字冒險遊戲引擎"

這個遊戲默認藏而不發,不會跟著 Calibre 啟動就冒出來。用戶得主動在"首選項 — 工具欄和菜單"裡把它請到主工具欄,才算真正激活。它的玩法很清晰:由 AI 在後臺搭起並掌管一個虛構世界,用戶通過不斷輸入文字來推著故事往前走,等於把"讀電子書"這件事,翻轉成了"和 AI 一起寫故事"。

7 小時前