超越特定領域的世界模型:JEPA-Anything 以單一方法涵蓋七大領域

2026年10月6日 05:12
站內 AI 整理稿

Researchers from PhAI Labs, CUHK, Fudan, Stanford, Oxford and Princeton have released JEPA-Anything, a domain-agnostic framework for building world models.Instead of designing a new predictive model for each field, it applies one shared learning recipe to very different systems.

It extends joint-embedding predictive architectures (JEPAs) with a method called Orthogonal Predictive Factorization (OPF).The research team tested it across 7 domains: vision, biology, clinical trajectories, control, molecular dynamics, physical fields and weather.

What problem does JEPA-Anything solve?A standard JEPA, such as I-JEPA or V-JEPA 2, uses a context encoder, an EMA target encoder and one predictor.The predictor outputs one monolithic target embedding.

The research team call this a capacity-allocation problem: high-variance structure dominates, and weaker modes get conflicting gradients.How does Orthogonal Predictive Factorization work?OPF splits the latent target of width d into K learned subspaces of width r, with d = K × r.

Most experiments use K = 4.Each factor gets a dedicated predictor.The factor predictions are then recombined through the Moore-Penrose pseudoinverse of the projector matrix.The result is 1 complete latent state for decoding, planning or rollout.

Three regularizers keep the factors useful: Orthogonality loss: keeps columns within each projector orthonormal and different projectors in non-overlapping subspaces.Factor-activity loss: a hinge on per-coordinate standard deviation so no factor goes dead.

Encoder-variance loss: sends a direct anti-collapse signal to the online encoder.The OPF loss is simply added to each domain’s original training loss.Domain adapters handle tokenization and encoders; the core library exposes the shared core as OrthogonalFactorProjection.

Orthogonality matters for stable synthesis.On CITRIS Interventional Pong, a capacity-matched unconstrained multi-head model had a condition number of 438.52.The orthogonal version reached 1.00005, with cross-factor overlap near zero.window.addEventListener("message",function(e){if(e.data&&e.data.

mtpJepaH){var f=document.getElementById("mtp-jepa-any-frame");if(f&&e.source===f.contentWindow){f.style.height=e.data.mtpJepaH+"px";}}}); What results does the research report?Group I, terminal readout: On single-cell data, zero-shot PBMC clustering (AvgBIO) rose to 0.7752 versus 0.

7194 for Cell-JEPA.Norman perturbation Pearson rose from 0.787 to 0.814.For forecasting over 1,000 clinical events on UK Biobank data, mean PRAUC was 0.718 versus 0.711 for the matched standard JEPA.Group II, latent world dynamics: On Interventional Pong, single-intervention MSE fell 34.83%.

Unseen combined interventions improved 12.90%, and 6-step free rollout improved 8.58%.JEPA-Anything improved reported metrics on all 10 matched dynamics tasks.Benchmarks include CausalWorld, DeepMind Control, PDEBench and WeatherBench2.On APEBench Burgers, 6-step rollout error dropped about 44.

7%, improving in every seed.For 100-step molecular rollouts with a TrajCast-style backbone, it posted the lowest MAE and RMSD on water, quartz, paracetamol and benzene.Planning results are mixed.With parameters matched within 0.3%, JEPA-Anything improved CEM return on Walker2d and HalfCheetah.

Hopper favored standard JEPA.Group III, scientific analysis: Factor analysis nominated IL-18 plus CD73 blockade as a cancer intervention.Wet-lab tests supported it in co-cultures, patient-derived organoids, tumor fragments and mice.

Latent orbital modes also recovered Kepler’s law with a fitted slope of −1.4991 against the theoretical −1.5.How does JEPA-Anything compare with other world models?

FeatureJEPA-AnythingV-JEPA 2DINO-WMDreamerV3TD-MPC2Core ideaJEPA with K orthogonal predictive factorsVideo JEPA plus action-conditioned V-JEPA 2-ACWorld model on pretrained visual featuresWorld model plus actor-critic trained in imaginationDecoder-free latent dynamics plus MPCTarget structureFactorized, recombined via pseudoinverseSingle latent targetSingle latent targetCategorical latent statesSingle latent stateDomains shown7: vision, cells, clinical, control, molecules, PDEs, weatherVideo understanding, robot manipulationPointMaze, PushT, Wall, deformablesDiverse RL domains, fixed hyperparameters104 continuous-control tasks, 4 domainsPlanning / rolloutLatent rollout, CEM planningPlanning from image goalsCEM planningPolicy from imagined rolloutsMPC planningPublic checkpointsPer-domain research checkpoints on HFYes, 300M to 1B (V-JEPA 2)PointMaze, PushT, WallNot listed in repo300+ checkpoints, up to 317MLicenseApache-2.

0MIT (some files Apache-2.0)MITMITMIT Sources: each project’s GitHub README, linked in the header row.JEPA-Anything checkpoint status from its Hugging Face card.Verified October 5, 2026.Key Takeaways OPF splits 1 JEPA target into K orthogonal factors with dedicated predictors.

It beat matched JEPA baselines on all 10 dynamics tasks.Interventional Pong single-intervention error fell 34.83%.Planning gains are environment-dependent; Hopper favored standard JEPA.Core code is Apache-2.0; research checkpoints sit on Hugging Face.

Check out the Paper, GitHub Repo and Model Checkpoints.All credit goes to the researcher of this project.Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?now you can join us on telegram as well.

Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?Connect with us The post Beyond Domain-Specific World Models: JEPA-Anything Uses 1 Recipe for 7 Fields appeared first on MarkTechPost.

Related

相關文章

量子位模型更新

最火AI崗位FDE:月薪5萬,都幹這些…

FDE(前線部署工程師)是近期最受關注的AI職位之一,海外年薪中位數約20萬美元,國內大廠也開出月薪三到五萬元。這份工作強調駐場梳理客戶的業務本體(Ontology),溝通時間佔七成以上,開發僅約三成。從業者認為,FDE與傳統外包不同,關鍵在於能否將經驗沉澱回自家產品並複用。

1 天前
MarkTechPost AI模型更新

DeepSeek Harness v0.2 為其開源代理框架帶來官方桌面應用程式

DeepSeek 已為 DeepSeek Harness (dsh) 推出官方桌面應用程式,dsh 是其開源代理框架。該應用程式隨 v0.2 預覽版一同發布,安裝檔支援 macOS(Apple 晶片)與 Windows(64 位元)。目前可作為預覽版部署使用,使用者可從 deepseek.com/harness 下載,或執行 npx @deepseek-ai/dsh web。DeepSeek 提醒未來可能會有破壞相容性的變更。 v0.2 新增內容:框架是將模型轉化為代理的執行環境,能讀取檔案、執行指令並維持計畫。v0.2 預覽版針對日常工作與程式開發進行優化。內建功能包括:預載常用辦公室與開發工具;新增插件管理頁面,可安裝、設定、啟用與停用插件;以及右側邊欄提供檔案與差異審查預覽。

2 天前
量子位模型更新

DeepSeek擴招!彈性計算團隊大量HC,尤其需要資深工程師

彈性計算團隊大量HC招人!尤其需要資深工程師。三週前不是剛招過一輪嗎?咋又缺人了。這次沒發崗位JD,直接甩了一篇DeepSeek技術分享: 《DeepSeek彈性計算(DSec):面向大規模Agent訓練的沙盒基礎設施》 現在DeepSeek的一套DSec擴展分片,大約有160臺服務器、3萬個CPU核心和250TB內存,每天要服務約300萬個沙盒。

2 天前