Sources · Updated 2026-08-21

Sources

An article source map: see which publications support the site's explanations, what role each source plays and where the evidence remains limited. The full verification method is explained separately.

这是本站的文章来源地图:集中展示各篇解释依赖哪些资料、每类来源承担什么作用,以及证据仍有哪些限制。完整核验逻辑另见核验标准。

Short answer / 一句话结论

This page is frozen as of 2026-08-21. The source and claim records below are no longer rechecked on a schedule. Each was checked on the date it names, and those dates are real; none of them says anything about what a source contains today. Nothing has been deleted, and errors are still corrected when noticed or reported. See For agents for the full freeze notice.

本页已于 2026-08-21 冻结。以下来源与判断记录不再做周期性复查。 每条记录都在它标注的日期核对过,日期为真;但它不说明来源今天的内容。历史记录一条不删,发现或收到纠错时仍会修改。 完整冻结说明见 For agents

Sources on this site are evidence inputs, not automatic truth. Official claims stay labeled as official claims unless independent validation is available.

本站把来源当作证据输入,而不是自动真相。官方声明会继续标成官方声明,除非有独立验证或更强证据支持。

Technical mappings: article, team and claim IDs / 技术映射:文章、团队与判断 ID

Article source map / 文章来源映射

Article pages are the human-facing explanations. Their evidence tables now expose claim IDs that match claims.json.

文章页负责给人读的解释。每篇文章的证据表现在会显示对应的 claim ID,并与 claims.json 保持一致。

Article Primary use Main claim IDs Main source IDs
What is a world model? Core definition, prediction, action, internal simulation. claim-world-model-dynamics, claim-video-generation-world-simulators, claim-genie-3-interactive-worlds, claim-vjepa-2-predictive-representation world-models-2018, openai-sora-world-simulators, deepmind-genie-3, meta-vjepa-2
Is Sora a world model? Boundary between video generation and interactive world modeling. claim-video-generation-world-simulators, claim-sora-minute-video, claim-sora-world-like-properties, claim-sora-not-full-public-interactive-world-model openai-sora-world-simulators, world-models-2018, deepmind-genie-3
Genie, Cosmos and V-JEPA Comparison of interactive worlds, physical-AI infrastructure and predictive representations. claim-genie-3-interactive-worlds, claim-genie-3-public-performance, claim-cosmos-physical-ai-platform, claim-cosmos-3-omnimodal, claim-vjepa-2-predictive-representation deepmind-genie-3, nvidia-cosmos, nvidia-cosmos-3, meta-vjepa-2
Why autonomous driving needs world models Traditional simulation versus learned world models, statistical mileage limits, closed-loop evidence and validation boundaries. claim-av-rare-scenario-pressure, claim-av-learned-world-model-added-value, claim-gaia-2-driving-domain, claim-waabi-world-neural-simulation, claim-cosmos-physical-ai-platform, claim-av-world-models-not-road-test-replacement rand-driving-to-safety, koopman-wagner-av-safety, wayve-gaia-2, waabi-world, nvidia-cosmos, nvidia-cosmos-3
How do you evaluate a world model? Nine evaluation dimensions, route-specific metric priorities and evidence-layer boundaries. claim-world-model-evaluation-multidimensional, claim-world-model-evaluation-route-specific, claim-world-model-evaluation-closed-loop, claim-world-model-evaluation-deployment-boundary, claim-world-model-evaluation-evaluator-is-audited world-models-2018, deepmind-genie-3, nvidia-cosmos-3, meta-vjepa-2, wayve-gaia-2, rand-driving-to-safety, koopman-wagner-av-safety, mirabench-2026, decision-centric-world-model-evaluation-2026
What have world models actually achieved in robotics? Physical-robot learning, image-goal planning, target-robot adaptation, benchmark limits and reproduction gaps. claim-robotics-world-model-real-hardware, claim-vjepa-2-ac-two-lab-planning, claim-robonet-target-robot-adaptation, claim-robotics-world-model-protocol-sensitivity, claim-robotics-independent-replication-gap daydreamer-physical-robots, vjepa-2-paper, droid-dataset-v2, robonet-multi-robot, visual-foresight-robot-control, dino-wm-v2, td-mpc2-iclr, jepa-wm-planning-tmlr, stable-worldmodel-v2
What do spatial 3D models actually measure? Measured geometry outputs, benchmark protocol and authorship, an independent photogrammetry comparison, 3D visual reasoning and generated 3D worlds. claim-spatial-geometry-measured-per-output, claim-spatialbench-protocol-bound-not-third-party, claim-spatial-accuracy-ranking-configuration-dependent, claim-spatial-reasoning-separate-evaluation-layer, claim-marble-spatial-world-generation, claim-spatial-capability-boundaries-not-established dust3r-paper, mast3r-paper, vggt-paper, spatialbench-v2, aerial-3d-reconstruction-eval, spa3r-paper, world-labs-marble, world-labs-about

Team / project source map / 团队与项目来源映射

The homepage team map is not a separate narrative layer. Its source and claim relationships are preserved here and in structured data without exposing internal IDs on the reader-facing cards.

首页团队图谱不再是独立叙述层。来源与判断关系保存在本页技术映射和结构化数据中,不再把内部 ID 直接显示在读者卡片上。

Team / project Route Source ID Claim ID
Google DeepMind / Genie 3 Interactive generated worlds deepmind-genie-3 claim-genie-3-interactive-worlds
NVIDIA / Cosmos Physical-AI platform nvidia-cosmos claim-cosmos-physical-ai-platform
Meta AI / V-JEPA 2 Predictive representation learning meta-vjepa-2 claim-vjepa-2-predictive-representation
OpenAI / Sora Video generation toward world simulators openai-sora-world-simulators claim-video-generation-world-simulators
World Labs / Marble Spatial 3D world generation world-labs-marble claim-marble-spatial-world-generation
Wayve / GAIA-2 Autonomous-driving world models wayve-gaia-2 claim-gaia-2-driving-domain
Waabi / Waabi World Neural simulation and validation waabi-world claim-waabi-world-neural-simulation
Runway / General World Models Creative video toward controllable simulation runway-gwm claim-runway-gwm-company-framing

Claim source map / 判断来源映射

This table lists the main public claims rather than every sentence on the site. The full records live in claims.json.

这里列出的是主要公开判断,不是网站每一句话。完整字段、边界和复核日期保存在 claims.json

Claim ID Source IDs Status Boundary
claim-world-model-dynamics world-models-2018 unmaintained / high Modern usage is broader than the 2018 project.
claim-video-generation-world-simulators openai-sora-world-simulators unmaintained / high Official framing, not proof of full interactivity.
claim-sora-not-full-public-interactive-world-model openai-sora-world-simulators unmaintained / medium-high Boundary judgment based on missing public environment-control evidence.
claim-genie-3-interactive-worlds deepmind-genie-3 unmaintained / high External validation of long-horizon stability remains limited.
claim-cosmos-physical-ai-platform nvidia-cosmos, nvidia-cosmos-3 unmaintained / high Platform value depends on developer adoption and downstream validation.
claim-av-rare-scenario-pressure rand-driving-to-safety unmaintained / high Statistical mileage result, not a universal deployment threshold or proof that learned world models are mandatory.
claim-av-learned-world-model-added-value wayve-gaia-2, waabi-world, nvidia-cosmos-3 unmaintained / medium-high Component and company evidence does not prove calibrated coverage or general safety value.
claim-av-world-models-not-road-test-replacement rand-driving-to-safety, koopman-wagner-av-safety, wayve-gaia-2, waabi-world, nvidia-cosmos-3 unmaintained / medium-high Independent sources support the validation boundary; current technical materials support the learned-model capability layer.
claim-world-model-evaluation-multidimensional world-models-2018, deepmind-genie-3, meta-vjepa-2, wayve-gaia-2 unmaintained / medium-high Editorial synthesis across unlike routes; no universal benchmark or independent reproduction is claimed.
claim-world-model-evaluation-route-specific deepmind-genie-3, nvidia-cosmos-3, meta-vjepa-2, wayve-gaia-2 unmaintained / medium-high Reusable taxonomy judgment, not a standardized score or ranking.
claim-world-model-evaluation-closed-loop world-models-2018, wayve-gaia-2, nvidia-cosmos-3 unmaintained / medium-high Action and inverse-dynamics evidence remains system- and protocol-specific.
claim-world-model-evaluation-deployment-boundary rand-driving-to-safety, koopman-wagner-av-safety unmaintained / high High-confidence autonomous-driving boundary; broader generalization is an editorial inference.
claim-world-model-evaluation-evaluator-is-audited mirabench-2026, decision-centric-world-model-evaluation-2026 unmaintained / medium-high Bounded to MiraBench as published on 28 May 2026, which discloses the prompt split and the per-model agreement itself; overall reported agreement is 87.8%. The position paper supplies the general principle and does not mention MiraBench.
claim-vjepa-2-predictive-representation meta-vjepa-2 unmaintained / medium-high Reported transfer remains task-dependent.
claim-marble-spatial-world-generation world-labs-marble unmaintained / medium-high Creation workflow claims do not prove strict physical simulation utility.
claim-gaia-2-driving-domain wayve-gaia-2 unmaintained / medium-high Driving-domain specificity limits general-purpose claims.
claim-waabi-world-neural-simulation waabi-world unmaintained / medium Company materials provide platform positioning, not detailed model validation.
claim-runway-gwm-company-framing runway-gwm unmaintained / medium Open technical detail and independent evaluation remain limited.

Sources and references / 参考资料

These are the publications used by the Atlas. The table explains what each source is and which part of the site it supports. Machine IDs and the dated access notes remain in the structured data.

这里展示本站使用的资料、来源性质和用途;机器 ID 与带日期的访问说明保留在结构化数据中。

Source Type Used for
World Models, Ha and Schmidhuber Paper / project Classic latent-dynamics framing for imagination, planning and control.
OpenAI: Video Generation Models as World Simulators Official technical note Sora's world-simulator framing and boundary between video generation and interactivity.
Google DeepMind: Genie 3 Official blog / model page Real-time navigable generated worlds and official Genie 3 capability claims.
NVIDIA Cosmos Official product page World foundation model platform for physical AI.
Cosmos 3 technical report Technical report Omnimodal world-model route across language, video, action and physical-AI scenarios.
Meta AI: V-JEPA 2 Official research release Predictive representation route, physical reasoning and robot-control transfer.
RAND: Driving to Safety — official publication record
Readable PDF mirror
Independent research report Statistical limits of road-mile validation and alternative testing as a supplement.
Koopman & Wagner: Autonomous Vehicle Safety Independent peer-reviewed article Testing-only boundary, ML validation data limits and end-to-end safety process.
Wayve GAIA-2 technical report Technical report Domain-specific autonomous-driving world-model case and rare-scenario article evidence.
World Labs: Marble Official product post Generative 3D world creation with an edit, expand, combine and export workflow. This page does not use the word “persistent”; the World Labs About page does.
Runway Research Company research page General World Models framing and creative-video-to-simulation route.
Waabi World Company materials Neural simulation, autonomous-trucking validation route and closed-loop article evidence.
DayDreamerPhysical-robot technical reportOnline world-model learning on four separately trained robot-task setups.
V-JEPA 2 paperDeveloper technical reportAction-conditioned post-training and two-lab Franka image-goal planning.
DROID v2Dataset paperRobot-data provenance; arXiv metadata/PDF task-count conflict is recorded in technical notes.
RoboNetPeer-reviewed paperMulti-robot data and finetuning-based held-out robot adaptation.
Visual ForesightTechnical reportVideo-prediction planning and physical robot control lineage.
DINO-WMTechnical reportLatent planning in controlled benchmark environments.
TD-MPC2ICLR 2024 paperContinuous-control and multitask benchmark evidence.
What Drives Success in Physical Planning with JEPA-WMs?TMLR reproduction studyV-JEPA 2-AC implementation and planner sensitivity, with original-author overlap.
stable-worldmodel-v1Reproduction preprintDINO-WM Push-T reproduction and robustness stress test, with author overlap.
DUSt3RDeveloper technical reportPose-free point-map reconstruction and camera geometry; table identity verified, cells not read.
MASt3RDeveloper technical reportDense 3D correspondence and localization results; a previously recorded p. 12 locator was withdrawn as nonexistent.
VGGTDeveloper technical reportFeed-forward geometry prediction; Section 5 states the fisheye, panoramic, extreme-rotation and non-rigid limits.
SpatialBenchBenchmark paper whose authors also enter a modelCross-paradigm protocol and out-of-distribution findings. Its cross-model ranking is not third-party evaluation.
Evaluation of DUSt3R/MASt3R/VGGT on photogrammetric aerial blocksIndependent peer-reviewed evaluationThe one independent comparison in the spatial 3D article; the accuracy ranking reverses with image count and overlap, in aerial photogrammetry only.
Spa3RDeveloper preprint3D visual-reasoning benchmark results; a single system cannot define the capability category.
MiraBenchUnrefereed benchmark preprint whose authors fine-tuned five of the twelve models it scoresAction-conditioned failure preservation, and the configuration of the automated evaluator that produces every score in it.
Decision-making-centric world-model evaluationUnrefereed position preprint, no experiments of its ownIndependent boundary on treating model-as-judge output as evidence of decision usefulness. Independent of MiraBench and does not cite it.
Technical review log / 技术核验记录

Review notes / 核验记录

  • 2026-08-08: MiraBench and the decision-making-centric position paper were added for the evaluation article's material update. Both are unrefereed preprints with no journal reference. MiraBench's own authors fine-tuned five of the twelve models it scores, and every score it reports is produced by its own VLM evaluator rather than by human annotation; both facts are recorded as source-role boundaries rather than treated as independent evidence.
  • 2026-07-12: Nine robotics sources were added. Fixed PDFs, experiment tables and limitations were checked; reproduction studies remain explicitly separated from independent physical-robot validation.
  • 2026-07-10: RAND Figure 3 (p. 7), Table 1 and conclusions (p. 10), plus Koopman & Wagner's testing discussion (p. 6) and conclusion (p. 9), were added as independent autonomous-driving validation evidence.
  • 2026-07-10: GAIA-2 Section 2.2 (p. 6), Figure 7 and safety-critical scenario discussion (p. 13), and Cosmos 3 Table 18 / Figure 22 (pp. 65-66) were visually reviewed for page-level locators.
  • 2026-07-10: Waabi World remained company-positioning evidence because no open technical report or independent benchmark was located in this review.
  • 2026-07-09: Waabi World was added to the source and claim layer to match the homepage team map.
  • 2026-07-09: OpenAI's Sora source returned 403 during direct access check; it remains listed as an official source with an access note.
  • 2026-07-09: Genie 3 24 fps, 720p and several-minute consistency are kept as official claims, not independent validation.
  • 2026-07-09: Cosmos, V-JEPA 2, Marble, GAIA-2 and Runway are treated as source-backed public claims with explicit boundaries.

These notes are a record of checks that were made, on the dates they were made. Since 2026-08-21 no further scheduled rechecks are performed, and nothing here is a statement about what a source says today.

核验记录的目的不是制造“已经盖章”的感觉,而是保留每次判断当时的依据与边界。 2026-08-21 起不再有计划性重查,这里的每一条都只说明"当时查到了什么",不代表来源今天的状态。

Using this evidence programmatically / 供 Agent 与技术使用者

Machine IDs, field definitions, access notes and retrieval guidance now live in one technical entry point.

机器 ID、字段定义、访问说明与读取规则现在集中在一个技术入口中。