Pith. sign in

Paper Citation Record · LEDGER

Cosmos World Foundation Model Platform for Physical AI

As of 5 August 2026, this Paper Citation Record lists 100 of 253 outbound references and 100 inbound Pith citation observations for arXiv:2501.03575.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.03575 v3

Coverage vector

measured 100 of 253 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T23:38:44.933410Z

measured 200 of 200 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 100 of 358 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:14:34.361533Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 253 outbound references displayed

  • verified exact34
  • verified fuzzy63
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

9
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b9eddad2-76cb-4889-99c8-a678007a23f1 · outbound

This paper cites SemDeDup: Data-efficient learning at web-scale through semantic deduplication.

Cosmos World Foundation Model Platform for Physical AI SemDeDup: Data-efficient learning at web-scale through semantic deduplication

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:43:31.136179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:be2f13bbb8a3236ee8d6cf63a823cf52f3124083595eb89e6de140e7ad369964

Observation d1a7fd91-439a-4324-839e-3360cddc9da9 · outbound

This paper cites Nemotron-4 340B Technical Report.

Cosmos World Foundation Model Platform for Physical AI Nemotron-4 340B Technical Report

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:38:45.226532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:27613fc4dcc908560742e39f1439139e9acdc6497e8be0fa473789e1f203efe6

Observation 76d9eb47-b380-48e9-8d31-8ffadc314883 · outbound

This paper cites Pixtral 12B.

Cosmos World Foundation Model Platform for Physical AI Pixtral 12B

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:53:29.932651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:48bdf46c3c0e37a0f9a365d58c22ba081401f939723f0ce83faf599f200b0e4c

Observation b8248dc6-b877-4311-91ff-ad33d19df6b9 · outbound

This paper cites Bbc planet earth dataset.

Cosmos World Foundation Model Platform for Physical AI Bbc planet earth dataset

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.954091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a778c4118c233c042b19a0bb67ff0da92287e5fe5ca67b049bc46f48999f3b16

Observation 6efd1fdc-cf2f-4d05-a6c2-ec65ee35ac70 · outbound

This paper cites Diffusion for world modeling: Visual details matter in atari.

Cosmos World Foundation Model Platform for Physical AI Diffusion for world modeling: Visual details matter in atari

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.960688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b2e5e939a4182d3076499a9fa2c8cf891536378fa8ecafe85911c3ca3eef65c7

Observation f0494f37-92a4-4ff7-a3d6-915ed3006afd · outbound

This paper cites Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation.

Cosmos World Foundation Model Platform for Physical AI Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.256610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:3cede24a466acb1b7181ebfa857d2fca02b41f308653aae70ca0b4f3063d09c7

Observation 9fa0fb77-f2be-4623-9000-9911990835d7 · outbound

This paper cites Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models.

Cosmos World Foundation Model Platform for Physical AI Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.273741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a790937e78444a288380e39bbfe69fbfc6ff02b7139cc8001286e57fe3089d9f

Observation cf7f9cde-43d7-4a40-921f-1846801b588a · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Cosmos World Foundation Model Platform for Physical AI eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:44:22.991611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:34b8afe7d28b58b2cb3cc6cdf21c87b0f6b159229d53111639932bdceef8fdfb

Observation 8d2efc41-4c23-47e5-89e2-3780ca54d986 · outbound

This paper cites Navigation World Models.

Cosmos World Foundation Model Platform for Physical AI Navigation World Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.293773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:0256bbc1fa39902d8b88b1c1a74b94a1b520f9d5c6eb306fa9c6fa9f78d8ae3b

Observation 44af0301-d311-4d46-a3f7-1583e5c2ae40 · outbound

This paper cites Improving image generation with better captions.Computer Science.

Cosmos World Foundation Model Platform for Physical AI Improving image generation with better captions.Computer Science

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.032345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:56d1f7dad7a9ff4c0240942cc2eb5f795e713b4e64b68b1bc653672e701d68a2

Observation 5d9fd3bb-f132-439c-a9af-8c5c96f2ee6e · outbound

This paper cites Zero-shot robotic manipulation with pre-trained image-editing diffusion models.

Cosmos World Foundation Model Platform for Physical AI Zero-shot robotic manipulation with pre-trained image-editing diffusion models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.050351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a016680382c96d769ce9cc6c37a6ff6d1144f992661c050a6bedad68b079e257

Observation 82540ea0-d2ca-4dd0-b9b6-1a4ebb59f932 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Cosmos World Foundation Model Platform for Physical AI Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.307012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:987e13ca8bea7602b1e76f73aaeaf97f7ea76f26d2d42d5b6a0d60fcdc15d683

Observation 8907886e-a7aa-47a4-95d0-fdc8a09b4772 · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

Cosmos World Foundation Model Platform for Physical AI Align your latents: High-resolution video synthesis with latent diffusion models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.077346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b638f844aa18d8d1e9705b58c26d7130e5e8f450a2d965f6f97255d0aa589429

Observation 97be606c-2830-4c56-81e2-a1217a4a21c8 · outbound

This paper cites an unresolved cited work.

Cosmos World Foundation Model Platform for Physical AI Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-10T23:38:47.103360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b63dc4a7b9c9861029145268a9e3ebc7689a5407cf256b0a8ab76b2909080620

Observation 24ff8b8c-8423-454a-abab-5bbac740306e · outbound

This paper cites Video generation models as world simulators.

Cosmos World Foundation Model Platform for Physical AI Video generation models as world simulators

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.112379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:53133b1d6fce2ebeb13c4e82548b8caaf7868bf71bba6fdea657c14116c22a77

Observation 7b001c47-7138-46fc-b888-302c1c553f15 · outbound

This paper cites Language models are few-shot learners.

Cosmos World Foundation Model Platform for Physical AI Language models are few-shot learners

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.125359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:1ec2dba142a3cb9e8acbcea18ba800e9886ca841466d9cdd5fe64174aff34c9b

Observation c510d33b-5c7b-4c81-918e-6bbdc0952019 · outbound

This paper cites Genie: Generative interactive environments.

Cosmos World Foundation Model Platform for Physical AI Genie: Generative interactive environments

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.138349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:0d87bd2695199b64c0495791e02e94088a9939e0328361d596e7147202d43da8

Observation db82a035-dd8e-4187-ba25-023dbe0e05d9 · outbound

This paper cites Lee, Deming Chen, and Tri Dao.

Cosmos World Foundation Model Platform for Physical AI Lee, Deming Chen, and Tri Dao

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.146855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:1c56a16033242dc5e49ae60cd53fe4bd70fc8e2b35b092d84252d187babb4def

Observation f45fb0e2-4feb-4837-9304-8db4035cc461 · outbound

This paper cites Pyscenedetect.

Cosmos World Foundation Model Platform for Physical AI Pyscenedetect

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.156354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:29d5dda1e33a552a6e1d73cc551edcdb450c23cbd82c971985b1c281f82100b0

Observation 082efe4e-f27d-4cdd-902e-c9c115d9211a · outbound

This paper cites an unresolved cited work.

Cosmos World Foundation Model Platform for Physical AI Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-10T23:38:47.160400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:285ddcfc4390b911caf6848a9e3bf218bd8358d17d8640db4599a3961760d456

Observation 3a08caa9-f224-41a1-a58b-ea867f4fce90 · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

Cosmos World Foundation Model Platform for Physical AI pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.166349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:1042e9b3732ef01fbd68192acf732d33308d1ca69c4352e51aaca59e8a91ff64

Observation 0d1d4298-c3d0-4f38-b305-c9920570d619 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

Cosmos World Foundation Model Platform for Physical AI GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:09:34.215368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:33ab3afbe764f198df9f8d6abccf59e2715a63b5ccad9849e901c5157242fb6b

Observation df32755f-f906-4059-9229-11da98eeece8 · outbound

This paper cites Training Deep Nets with Sublinear Memory Cost.

Cosmos World Foundation Model Platform for Physical AI Training Deep Nets with Sublinear Memory Cost

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:44:07.405726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e1f952a2b52e9622b12d64886abf1a6584ee16e9803c5fd949de8148c31816bb

Observation 610c0e48-b34e-48b6-be5f-9aab5e2955e9 · outbound

This paper cites On the Importance of Noise Scheduling for Diffusion Models.

Cosmos World Foundation Model Platform for Physical AI On the Importance of Noise Scheduling for Diffusion Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.368038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:dd6986ad323b7fc1d010e1128f969e7f2546173c1384f9609fa7663e9c55c2d4

Observation 68e5fbf5-4dff-4f0f-a559-1e7c26f17cb3 · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

Cosmos World Foundation Model Platform for Physical AI Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.194838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:bbed303f6dec76297a3eeed84befefad11ae350902c3a08acd65b4fc3db4f9a7

Observation fd0ccab1-fe39-4eb6-9969-9a718150bb5a · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.RSS.

Cosmos World Foundation Model Platform for Physical AI Diffusion policy: Visuomotor policy learning via action diffusion.RSS

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.201403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:387f5090998168af0cfa6b7c8872aed9a1559805a1125212f1026e72a2320136

Observation 74b91a0b-aac7-4316-90e4-3e26a37af673 · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

Cosmos World Foundation Model Platform for Physical AI Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.377347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:18a4340023f531c2bf2533f0755168c3ddbe6001c7c91f88c12be3f36af7a270

Observation 9a568bc0-fb85-4164-9823-5d8aa4abe361 · outbound

This paper cites The Z-loss: a shift and scale invariant classification loss belonging to the Spherical Family.

Cosmos World Foundation Model Platform for Physical AI The Z-loss: a shift and scale invariant classification loss belonging to the Spherical Family

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.394809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:862b889b47779f6993c9945188902c1a4be0ae49c0ddd3232255b6c2cebbd63d

Observation 3e634f34-4f4c-4401-822e-87f9a4f806fc · outbound

This paper cites Scaling vision transformers to 22 billion parameters.

Cosmos World Foundation Model Platform for Physical AI Scaling vision transformers to 22 billion parameters

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.218259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:54b19d53ca1d5d9111b9ffe42f5d04ca649508c275a04380eb7d9f2d7ba293f0

Observation 351ca4d4-d0e0-43e3-b8ea-cdfbfb372e8c · outbound

This paper cites Autoregressive Video Generation without Vector Quantization.

Cosmos World Foundation Model Platform for Physical AI Autoregressive Video Generation without Vector Quantization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:07:39.939813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9020ccaf393bf9074ec6308f87efd738482c497eff553f226d7880cd3c928da1

Observation f0919126-1aa5-4eea-be12-28b9dc668c4b · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Cosmos World Foundation Model Platform for Physical AI Imagenet: A large-scale hierarchical image database

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.226707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:71b2e42f52ff12b29d6f8496e00fe2e4ec712dd4c878ff44e8bf0feca23cb6b8

Observation 42020881-2bce-4ab7-86f2-b28455c74624 · outbound

This paper cites Retinaface: Single-stage dense face localisation in the wild.CVPR.

Cosmos World Foundation Model Platform for Physical AI Retinaface: Single-stage dense face localisation in the wild.CVPR

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.231215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:473ae70cde1dcb0ce3b580dc889d7c4da6ac32dd232533a8bbc6e1a0de0bce9d

Observation 131fd15d-901f-48ce-8280-e7f4fbc4bcd0 · outbound

This paper cites Superpoint: Self-supervised interest point detection and description.

Cosmos World Foundation Model Platform for Physical AI Superpoint: Self-supervised interest point detection and description

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.240577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:6cd74a01d01099ff1fa7c24a9f0fabd5d7d545a4cfc01a7714a0e97393dbc51d

Observation 5481f156-04b8-48e8-a82e-f270de850db1 · outbound

This paper cites Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning.

Cosmos World Foundation Model Platform for Physical AI Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.413511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:fc5c03f3bcc23c98b017de1f8c631d3cb7dae9b2fc371b950acfefde0e489581

Observation a97cd7aa-e6e6-4248-a7f3-ae9150f3d901 · outbound

This paper cites An image is worth 16x16 words: transformers for image recognition at scale.

Cosmos World Foundation Model Platform for Physical AI An image is worth 16x16 words: transformers for image recognition at scale

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:47.251179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:6235269de2da30ef3fa9634856a687ecaf40276aa7a0b35820c172de4e42c878

Observation f6fe4fe4-14b2-4d62-b9d3-f0e7f5c79cd4 · outbound

This paper cites Learning universal policies via text-guided video generation.

Cosmos World Foundation Model Platform for Physical AI Learning universal policies via text-guided video generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.447628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:23626106a6c723637fa75455c716c00c179502ab28238f1d16221a058e14ba22

Observation b403a2cb-1c99-478b-b0bf-44738389745c · outbound

This paper cites The Llama 3 Herd of Models.

Cosmos World Foundation Model Platform for Physical AI The Llama 3 Herd of Models

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.422498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:20b767cb25460e4bf496870a1d08a6d608100cb218aefc3dff415001e2b385b5

Observation 585782af-4b38-427a-a647-45a3e5edf30c · outbound

This paper cites Bridge data: Boosting generalization of robotic skills with cross-domain datasets.

Cosmos World Foundation Model Platform for Physical AI Bridge data: Boosting generalization of robotic skills with cross-domain datasets

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.479120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:0d0e21f26268b4e55da732f0e4cec24292575bec1952f252e1d683949dfc8539

Observation 7692dfdc-dafe-4b03-967c-a37691e03551 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Cosmos World Foundation Model Platform for Physical AI Taming transformers for high-resolution image synthesis

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.531141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ac0bb7a8023236de2bc6d9b00cdb23c285b8eb937fe646b64a57e6df50353b39

Observation 3791fcb7-9592-4788-9496-2d62610ab126 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Cosmos World Foundation Model Platform for Physical AI Scaling rectified flow transformers for high-resolution image synthesis

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.362308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ebc594a25bf419265a2b56855cd9afbf8d275bc6a22bc4fc1bd24cc66967d8ce

Observation 1ac09ec0-84e0-4eb0-bcd2-b33870d039a4 · outbound

This paper cites Two-frame motion estimation based on polynomial expansion.

Cosmos World Foundation Model Platform for Physical AI Two-frame motion estimation based on polynomial expansion

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.474704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:62a9a015e8aa174f26c6fbf21c47f896e09f6aaa797a41964afe14bc68523e7b

Observation 327fb231-5d79-474e-b826-a73383ac6036 · outbound

This paper cites Deep visual foresight for planning robot motion.

Cosmos World Foundation Model Platform for Physical AI Deep visual foresight for planning robot motion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.475477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:f7bf52b994e12c71d3a9bc2637f0403afa1e0f8b891fe90576edbbe6ae6ce544

Observation 3d79f949-f051-4164-89cf-ba8b3f6329f6 · outbound

This paper cites FLUX.1: Image generation.

Cosmos World Foundation Model Platform for Physical AI FLUX.1: Image generation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.521517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:535b7e619fd7935f8554d572e183822115de7dccbe0cfdd09fd160c8c5f949b6

Observation d9c09137-44a6-4317-a1a7-eab79132a116 · outbound

This paper cites Dreamsim: Learning new dimensions of human visual similarity using synthetic data.

Cosmos World Foundation Model Platform for Physical AI Dreamsim: Learning new dimensions of human visual similarity using synthetic data

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.366304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:2198e00ab8116ef3e1933c98efd92bb054c099b7b81d36a6b38420161078082f

Observation 41cc40ae-4f80-4429-98bb-b4268818dce9 · outbound

This paper cites Datacomp: In search of the next generation of multimodal datasets.

Cosmos World Foundation Model Platform for Physical AI Datacomp: In search of the next generation of multimodal datasets

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.517500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:c84dc9d49a100237a812c03e3f5fff82b31a1af07b948e12a917bf35dc2ae02a

Observation 3e58b3af-452c-48f5-a3fe-33af67801cad · outbound

This paper cites Make-a-scene: Scene-based text-to-image generation with human priors.

Cosmos World Foundation Model Platform for Physical AI Make-a-scene: Scene-based text-to-image generation with human priors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.548648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:ea07b64fba81c5802ffa5748c33014123bdaa7ac703c894512953567fc99ad7e

Observation 7e0c643b-a5a5-4b66-b1a7-b77ca8068e58 · outbound

This paper cites A new algorithm for data compression.The C Users Journal.

Cosmos World Foundation Model Platform for Physical AI A new algorithm for data compression.The C Users Journal

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.427828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cc001c351614fd7c152fa1f7025fc9ab216a54b9053e8901a2fbcfb7f6ae24f2

Observation a061fda2-7525-411b-b07d-d61826ecd72f · outbound

This paper cites Murphy, and Tim Salimans.

Cosmos World Foundation Model Platform for Physical AI Murphy, and Tim Salimans

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.503457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cc661cd3cce9b71acf440439473539cec7b750c4708b14b991f478bbdf314a22

Observation 09e0f4a1-c4a4-45c4-b6c3-c6c62c124c8b · outbound

This paper cites MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control.

Cosmos World Foundation Model Platform for Physical AI MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.431062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:656ecb27005b65dd986d78076aa9e50715b31beaf6bc9a471db39d64e88eb3c3

Observation 052f054b-77b5-40f2-b58d-192145f5dc0e · outbound

This paper cites Magicdrive: Street view generation with diverse 3d geometry control.

Cosmos World Foundation Model Platform for Physical AI Magicdrive: Street view generation with diverse 3d geometry control

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.499767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:14bb809f123e4a34cca627b0e2cba86196e6c00a2b2ace6397054327a2fb6fe2

Observation 917ffa15-80ca-4f3e-945d-7ca6ada99ebb · outbound

This paper cites Vista: A generalizable driving world model with high fidelity and versatile controllability.

Cosmos World Foundation Model Platform for Physical AI Vista: A generalizable driving world model with high fidelity and versatile controllability

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.508533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:caf3e7fd237b416bb6a88fe002482b0f46a7e8a40047880f3d07af9ab962b5ce

Observation 7efcc68e-0777-4923-90dc-da561c561e9a · outbound

This paper cites Image style transfer using convolutional neural networks.

Cosmos World Foundation Model Platform for Physical AI Image style transfer using convolutional neural networks

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.491965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:2e58aad2e7eee04d184027427eed587308c9f236399bdc506614d8f24fd173b4

Observation 8419c0a4-cbfe-435e-b835-8f2cc198a026 · outbound

This paper cites Long video generation with time-agnostic vqgan and time-sensitive transformer.

Cosmos World Foundation Model Platform for Physical AI Long video generation with time-agnostic vqgan and time-sensitive transformer

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.441584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:0185fec6ce29e5418b77289dd1a0de58fb113fc58a30195a4b071f84458dce4d

Observation a7b15757-ef00-4560-b4a7-59310c3b1dd4 · outbound

This paper cites Preserve your own correlation: A noise prior for video diffusion models.

Cosmos World Foundation Model Platform for Physical AI Preserve your own correlation: A noise prior for video diffusion models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.471705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:0b03771ad7f9a600070292227281f1c758e708c4bf8a979d37a23d7ad7132b9b

Observation bb242d0a-a2b3-4b03-9913-3f5b3f4abce5 · outbound

This paper cites Visual fact checker: Enabling high-fidelity detailed caption generation.

Cosmos World Foundation Model Platform for Physical AI Visual fact checker: Enabling high-fidelity detailed caption generation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.420285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:be5d9bc859bddbe680c736414b496b8ea42ef74caa1fbbab16faa882d56566d5

Observation c38e0513-05f2-4a38-8b44-ff848401c888 · outbound

This paper cites AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts.

Cosmos World Foundation Model Platform for Physical AI AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.444469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:211f16d1d5cab3b781f6f901081d49da2d87b0ef7cf0a61102b7c599c814017d

Observation f6763656-85ec-4941-81a1-d5b0e763cb35 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Cosmos World Foundation Model Platform for Physical AI Imagebind: One embedding space to bind them all

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.556492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:872ab0b25214a31c3f1b60cc00e1540749d5c2b0534289d094550de236adbaf4

Observation 20df0eaf-79b7-4907-8ccb-70d10a214247 · outbound

This paper cites Emu video: Factorizing text-to-video generation by explicit image conditioning.

Cosmos World Foundation Model Platform for Physical AI Emu video: Factorizing text-to-video generation by explicit image conditioning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.540196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:24f59c5b56b2284027724214b59a4a724be819f77f2c1708030482b4b51ca065

Observation eecd8030-5738-4296-9b79-c6da8741e23a · outbound

This paper cites Ego-exo4d: Understanding skilled human activity from first-and third-person perspectives.

Cosmos World Foundation Model Platform for Physical AI Ego-exo4d: Understanding skilled human activity from first-and third-person perspectives

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.495940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:65de2fae931a45a6379d0787f8f776c247578abce7aaf2eba221457cf7e7858a

Observation d636bb7a-7ddd-4903-a7ee-df722c13e080 · outbound

This paper cites Photorealistic video generation with diffusion models.

Cosmos World Foundation Model Platform for Physical AI Photorealistic video generation with diffusion models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.544202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:edaa90e37c68096560a8b42666b5772fb47e0820f689eea74adb95edaa97287a

Observation 71ac6de8-988a-4710-837c-de74d6a31447 · outbound

This paper cites Pre-trained text-to-image diffusion models are versatile representation learners for control.

Cosmos World Foundation Model Platform for Physical AI Pre-trained text-to-image diffusion models are versatile representation learners for control

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.483224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9faec19eb77065a49cea7c9702141c19803bef173dd525d7d0125a35104c4c9a

Observation 01b8d6ca-8679-4bb8-bc57-9d7a95f16fa0 · outbound

This paper cites SPACE:Speech-driven Portrait Animation with Controllable Expression.

Cosmos World Foundation Model Platform for Physical AI SPACE:Speech-driven Portrait Animation with Controllable Expression

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.401003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:096bcaf6cd3b0314ff4f5151e421fbba8ab2b29b45950de943cec3859cbc51f0

Observation 1cd3b051-b359-4c21-991a-c774a606a1f2 · outbound

This paper cites World Models.

Cosmos World Foundation Model Platform for Physical AI World Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:08:36.665942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:335c0ae812419df23e93e20f9504cc501770790bfd5999d0d8d8a3027f875f6a

Observation 4bf1f6d6-0a88-45e8-8516-61591075103a · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

Cosmos World Foundation Model Platform for Physical AI Dream to Control: Learning Behaviors by Latent Imagination

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:16:36.618198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:bfd592ab84a05561dc61f29d0daefaf284603b16533a9ea9b4c3ec5604f70ddd

Observation 1b88f61f-2bcd-4d2a-a1ac-a3dad34dff1d · outbound

This paper cites Mastering atari with discrete world models.

Cosmos World Foundation Model Platform for Physical AI Mastering atari with discrete world models

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.430428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cad3b719e3d33ec6503985f7e6baae61f6fc3e65161d9bd32316086ff18d71ce

Observation 122bdbad-d14d-4b14-9a82-23007c0ebf3b · outbound

This paper cites Mastering Diverse Domains through World Models.

Cosmos World Foundation Model Platform for Physical AI Mastering Diverse Domains through World Models

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:08:22.800194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a9d167234dc8aaa0f1226d11adef3cdbc0fe5244d03c05ccf72ff93bea3c4fc7

Observation edfa3c1e-a956-4442-8e01-695c33dcb755 · outbound

This paper cites Td-mpc2: Scalable, robust world models for continuous control.

Cosmos World Foundation Model Platform for Physical AI Td-mpc2: Scalable, robust world models for continuous control

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.453439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9503266be775eb125367e4436c1ae4e1ceb123128583e79d99108d620be0e112

Observation c884df7d-420d-43b2-b2dc-e9f2dbe0ed98 · outbound

This paper cites Cambridge university press.

Cosmos World Foundation Model Platform for Physical AI Cambridge university press

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.525849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:50477647430b55587630f6453921d66e51a459e6a1ddd9136023683d2f2b0d74

Observation 11282f87-2687-45a5-a525-cb7e44d96d25 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

Cosmos World Foundation Model Platform for Physical AI CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:06:24.067006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:c5f575d54f5f0028cea5c993ad514d6e5de6de4de059ebd58a24c772ca487092

Observation bba45182-89b5-48e0-932b-838544dfd566 · outbound

This paper cites Learning an actionable discrete diffusion policy via large-scale actionless video pre-training.

Cosmos World Foundation Model Platform for Physical AI Learning an actionable discrete diffusion policy via large-scale actionless video pre-training

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.391013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:78df36757e1d368c53f42438354ad1b4c30de8cd46e4bc144135c81e927ee725

Observation 0ac13db5-4cf7-4b4e-987b-26417e0e75c1 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Cosmos World Foundation Model Platform for Physical AI Masked autoencoders are scalable vision learners

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.484329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:73f1204bdc2ee72e9c9e21813cbe0483406bf2c8f60e64ac1dbc459d066bbaee

Observation 00f28e15-78dc-4604-acd8-d2a850d7f5b5 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Cosmos World Foundation Model Platform for Physical AI Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.535933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a3272294ed363a581a0cc0da03a4fdcf770a76a4fa97409cdb70e0f1a28e0ac0

Observation b959d73d-d7c9-4306-a626-96b50251e9d9 · outbound

This paper cites wake-sleep.

Cosmos World Foundation Model Platform for Physical AI wake-sleep

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.487460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a5398300c5c4299de576736292e93382725e5db7fe98a32c47b85ab8dcd23d71

Observation e9f131af-e71d-402b-8206-6ec61072c4a5 · outbound

This paper cites Classifier-Free Diffusion Guidance.

Cosmos World Foundation Model Platform for Physical AI Classifier-Free Diffusion Guidance

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.494382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:e8825729885d57521421ce656e1cdc15e352d637e8f3163a3e6b9aee330845b7

Observation c7a204ae-f586-430e-aff6-27a1dac0a726 · outbound

This paper cites Denoising diffusion probabilistic models.

Cosmos World Foundation Model Platform for Physical AI Denoising diffusion probabilistic models

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.436107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:3a29c315e8e7e10c2385faf22d9d8f4527a6ece5a338c14db53695dc483a08c7

Observation e694479b-d138-4d0d-98db-aeadf1da4a02 · outbound

This paper cites Video diffusion models.

Cosmos World Foundation Model Platform for Physical AI Video diffusion models

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.462831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:45f4261f041f19f372e7ff41bdb8a1822f8bca85e1976ae5daccc70712a7619b

Observation 8fb25a0f-7222-4691-9deb-742c3600272c · outbound

This paper cites Cogvideo: Large-scale pretraining for text-to-video generation via transformers.

Cosmos World Foundation Model Platform for Physical AI Cogvideo: Large-scale pretraining for text-to-video generation via transformers

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.560558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:8e8f9439de9b754bbe307e918d6eed20574c3f8ef1f24a30ad950e34e855a2bd

Observation 162e0e9c-1f84-464a-bd31-c3ede4134fca · outbound

This paper cites Simple diffusion: End-to-end diffusion for high resolution images.

Cosmos World Foundation Model Platform for Physical AI Simple diffusion: End-to-end diffusion for high resolution images

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.471142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:7420544430ba76915f504e6c95d866a883aeb6f8a76523532c9b7ec530cc4543

Observation 8ffd52d9-526e-4394-99e8-712bbd521c8e · outbound

This paper cites Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion.

Cosmos World Foundation Model Platform for Physical AI Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.501344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:cc433e8b48da1f612e3b6cb423f69ed620582c21525a69922baa3f042f8be686

Observation eb161c84-38d1-4100-bda3-dab8034346ed · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Cosmos World Foundation Model Platform for Physical AI GAIA-1: A Generative World Model for Autonomous Driving

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:15:10.779550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:771d101f5384a6bab9bde2c551e707de675056285a4f0dcb15dcfa5057896572

Observation 9a07989a-0508-4b42-86d2-5b744b87a199 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Cosmos World Foundation Model Platform for Physical AI LoRA: Low-rank adaptation of large language models

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.451531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:80462ffb06d6cf0a78bfd633d7f9f8f10005d350186881d3b640398423e5e189

Observation 6412333a-37af-4624-9afa-a6bef513e673 · outbound

This paper cites GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs.

Cosmos World Foundation Model Platform for Physical AI GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.524063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:a9f4c1214780bbd248eaf1c1489d4d36339ad1b82d1e1b22e04bfab88d255e97

Observation c126c200-4f63-41a8-9780-7af978d91c6d · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

Cosmos World Foundation Model Platform for Physical AI Vbench: Comprehensive benchmark suite for video generative models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T01:20:29.564785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:128df3048336afc5fae648c7cc6cd3e202bdb5b212210a6abe76bbfc8701f193

Observation c077c955-8436-4525-b0ca-a0fd153a664d · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Cosmos World Foundation Model Platform for Physical AI Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 84

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.530796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:59d33ac64ad798cdfff19b2e1861111f86f5140a78f7999de6405efa7f3f8844

Observation 5cb2648e-9301-4a93-8528-17243aacb2eb · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Cosmos World Foundation Model Platform for Physical AI DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:07:22.650116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:afe8a34706898e01f8381ac06d611b8998a0aa9056f37fa0101599ebc9d62e5f

Observation 747c2d4e-dcdc-4926-b9f2-09752f00729a · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

Cosmos World Foundation Model Platform for Physical AI ADriver-I: A General World Model for Autonomous Driving

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.553921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:23361067b4d5dec8f5a844073754086a70b130e5c61bdab43227c12731df6043

Observation 8cccda4b-da3c-42a0-a2f1-5dba18b811e7 · outbound

This paper cites Mistral 7B.

Cosmos World Foundation Model Platform for Physical AI Mistral 7B

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.559658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:63cd05c5f8d87b527513876e9cf108f98f124ea6c297a827f7c92e81de11dc9d

Observation bf0fded8-bb87-4995-8ca2-da5b34d34525 · outbound

This paper cites How Far is Video Generation from World Model: A Physical Law Perspective.

Cosmos World Foundation Model Platform for Physical AI How Far is Video Generation from World Model: A Physical Law Perspective

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:13:41.491170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:6cf4c222d6911e8720b62a10b92432fd61918db1de21627b2a96c7730fb4ac8d

Observation 77c7af6d-1fa0-45e2-a2eb-b79dd46bf9f3 · outbound

This paper cites Scaling Laws for Neural Language Models.

Cosmos World Foundation Model Platform for Physical AI Scaling Laws for Neural Language Models

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.577077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:f8503244568ed5eeace9bc27d730c11720770cb6a75861d10775611c8dec3151

Observation cb7d58f3-d100-4925-9604-bc3dbdd0df0c · outbound

This paper cites Analyzing and improving the image quality of stylegan.

Cosmos World Foundation Model Platform for Physical AI Analyzing and improving the image quality of stylegan

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.219507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:dfc0d306fb9a315a5aee0136ff75a5b0862954e13a30104222b4444f9e063c1e

Observation eaebe424-dd94-4215-af7e-f4b7ad315ba7 · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

Cosmos World Foundation Model Platform for Physical AI Elucidating the design space of diffusion-based generative models

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.236004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:99169bc2089b204233243ccf0accecc52348e44053e6f8b469646fad4705a8d2

Observation f48f9da6-31ba-4f65-bf5e-34f43a025811 · outbound

This paper cites Analyzing and improving the training dynamics of diffusion models.

Cosmos World Foundation Model Platform for Physical AI Analyzing and improving the training dynamics of diffusion models

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.241534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:5b80ddf5707d0e3c04b6af7b253d8f02e1fd88fe92851c9c3309ea30bedde5d8

Observation d0944257-c6c4-4007-ac27-6d8db6044f2f · outbound

This paper cites 3d diffuser actor: Policy diffusion with 3d scene representations.

Cosmos World Foundation Model Platform for Physical AI 3d diffuser actor: Policy diffusion with 3d scene representations

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.253128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:2f44cae9777f56e2ec92e8efe78713919d2b4bfa68d65693bf066e08ea1b01bf

Observation 1b849c64-4670-4591-82a1-601b4c16b9ae · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Transactions on Graphics (TOG).

Cosmos World Foundation Model Platform for Physical AI 3d gaussian splatting for real-time radiance field rendering.ACM Transactions on Graphics (TOG)

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.261429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:872b592f4d2fd780c7cab7361aa8a76b00d1a8548703829c5a76540623329fc6

Observation 2e255add-5a32-4427-a3be-bab393a1f6fa · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

Cosmos World Foundation Model Platform for Physical AI YOLOv11: An Overview of the Key Architectural Enhancements

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-12T13:29:16.418786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:72c6df128dbe2ba466aac0a71ce107f21259f8b495420cc978ba145d481bf22c

Observation 752045ee-252c-4ac9-8516-4dccea3b3dc9 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Cosmos World Foundation Model Platform for Physical AI OpenVLA: An Open-Source Vision-Language-Action Model

Reference 96

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.595747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9707e7c1034afc8c46158f2e5fa71c68ea78b579349b5c6765f8a28e72f25856

Observation 8ee927dc-3969-497a-ae2e-cdfc08a66b50 · outbound

This paper cites Learning to Simulate Dynamic Environments with GameGAN.

Cosmos World Foundation Model Platform for Physical AI Learning to Simulate Dynamic Environments with GameGAN

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.295756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:79465da60f70bc209e11831bdb33cb4d321d95bc4aaad5c0b6f252520cedbf9d

Observation e36484a6-c7e1-4902-b27b-112c6787ae3c · outbound

This paper cites DriveGAN: Towards a Controllable High-Quality Neural Simulation.

Cosmos World Foundation Model Platform for Physical AI DriveGAN: Towards a Controllable High-Quality Neural Simulation

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.308459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:b3dacb87dc5b5cba4fb8bb3130ce31eb84f64b3cadd7bdc0fec4f9e83372c79e

Observation c8ab99c4-d22a-4efc-8137-a9b8fda50ff8 · outbound

This paper cites Auto-Encoding Variational Bayes.

Cosmos World Foundation Model Platform for Physical AI Auto-Encoding Variational Bayes

Reference 99

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.606393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:7de67c9445ad920d8e00d75d2e41992bfc332b1678dadeb6353faea112522acb

Observation ba229324-e554-46c4-8eea-fb3ec7296d31 · outbound

This paper cites Learning to act from actionless videos through dense correspondences.

Cosmos World Foundation Model Platform for Physical AI Learning to act from actionless videos through dense correspondences

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T23:38:46.337589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:985e3663e0b0f82dd0c1ff3ba68ffa00eca610fa50a163d07b563f924a7a6bcb

Pith citing papers

Observation 87e26b48-3ac2-407e-b763-1e55d40d7074 · inbound

Latte: Latent Diffusion Transformer for Video Generation cites this paper.

Latte: Latent Diffusion Transformer for Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-13T21:45:35.796766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T21:45:35.754742Z digest=sha256:90ddc53f6f13d2085741733df9fa9d22a258f7e8ac3ee6a3c263d077af127295

Observation 4097a27e-822c-4e46-b132-57a05af61dd6 · inbound

Do generative video models understand physical principles? cites this paper.

Do generative video models understand physical principles? Cosmos World Foundation Model Platform for Physical AI

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-20T12:47:05.854160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T12:47:05.825659Z digest=sha256:f39ca7c5cf76a66680c31b509bb1761ee479277e61a9f17fcdb9b4a065c67f93

Observation 07d9e05a-eb3a-484f-b201-1502b89bf29a · inbound

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling cites this paper.

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-17T02:52:20.670876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T02:52:20.643070Z digest=sha256:189280ac06a2f6d4a18d1e960625d97a28ffaa8f112028cceccbab2c7b6afdba

Observation 6f1d706e-5be6-48ce-a7eb-88a2acd01955 · inbound

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model cites this paper.

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model Cosmos World Foundation Model Platform for Physical AI

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:02:23.962171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-19T08:02:23.002090Z digest=sha256:c874fe5913a35a21b9763fc7e0fbab41f0321dc3a07f7039444506516134f533

Observation eaad51e5-1286-4666-9fd2-e6c708b049e3 · inbound

Simulus: Combining Improvements in Sample-Efficient World Model Agents cites this paper.

Simulus: Combining Improvements in Sample-Efficient World Model Agents Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-23T02:32:26.169585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T02:28:23.113408Z digest=sha256:1f65ca830f4776aa51a61b185ebe18fd15f5cf1ad4291695d2651309bf91f993

Observation f4bafd84-306a-4ebd-80cf-94f7275ee836 · inbound

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots cites this paper.

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:09:10.112304Z digest=sha256:5109a72f85bb45b7927b83e47fa6578bf48b23fc9179c0c6a997f7e664c89548

Observation 0086e66b-c74d-4771-86c5-fc8b87d6633c · inbound

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning cites this paper.

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning Cosmos World Foundation Model Platform for Physical AI

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:47:10.244526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:47:10.146795Z digest=sha256:d6116efa1bbf9f5010282988b166851499f26f6fe5e8ac70d3c9de939fb7d426

Observation a3d5a6f2-c170-4668-8183-428a9b4e96b2 · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:05:17.269584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:5eb53986e1874058749100016db377230990b54ac53597946ead9d4748637287

Observation 6f53b649-f81f-4bcd-8611-eb03a817272e · inbound

GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving cites this paper.

GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-15T13:48:22.318514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T13:48:22.279405Z digest=sha256:a7832ee4c3fd35a4dc1a72c28b8658d5bae7744e7fc2efd86e19fa9407208881

Observation d8035884-cc36-4141-b377-9d2473332ee3 · inbound

AccidentSim: Generating Vehicle Collision Videos with Physically Realistic Collision Trajectories from Real-World Accident Reports cites this paper.

AccidentSim: Generating Vehicle Collision Videos with Physically Realistic Collision Trajectories from Real-World Accident Reports Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-22T22:27:12.531213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T22:25:57.796922Z digest=sha256:656c0461360ec1ccef5abdbe6a7a9fcc4846f4aa2f0514e243bbe4f655cd2a68

Observation b2a5a481-821d-47ca-b442-25c02053e973 · inbound

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness cites this paper.

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness Cosmos World Foundation Model Platform for Physical AI

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-14T18:42:03.311760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T18:42:02.940250Z digest=sha256:a123688fbb731b0a0d0902791455b166697bf381b828e753d17c892e146e8f63

Observation faa18e6f-0cb2-4d91-a77a-f29599f73889 · inbound

Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets cites this paper.

Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets Cosmos World Foundation Model Platform for Physical AI

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-13T16:25:00.461445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T16:25:00.365534Z digest=sha256:2fc7829cee3e7cd774f565448f68a29f13f921d1ee9023d9d32c957631b82187

Observation 727c359b-f3fa-4b70-802e-74e0fe353586 · inbound

SkyReels-V2: Infinite-length Film Generative Model cites this paper.

SkyReels-V2: Infinite-length Film Generative Model Cosmos World Foundation Model Platform for Physical AI

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-14T20:23:04.195745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T20:23:04.022599Z digest=sha256:2822895b61e2f079770821f2a440d915e0f6d757754623f5dd36bd76b48463d2

Observation fda9cd74-497c-42ec-ba96-b2fb99f30de6 · inbound

EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video cites this paper.

EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video Cosmos World Foundation Model Platform for Physical AI

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-15T15:40:29.404704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T15:40:29.361187Z digest=sha256:e8f3813de08d32e97ca9cb61213220dae9d42bc8a1237b349d368aa577a36e4c

Observation 3be5cab3-7d3b-43a4-96a2-d5a64357f8ab · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Cosmos World Foundation Model Platform for Physical AI

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-15T23:50:45.427504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:7f79e3230cf0b5c69ccebf4d4336d3c29cb9b872f5913a9dbdbfad289b5b5fb9

Observation 589dae3b-092e-4076-be5d-cc5563d90ebf · inbound

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence cites this paper.

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-22T13:11:35.653194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T13:07:11.548885Z digest=sha256:e60a70d9f4aa8e447880149d56f1bde0052c2fbb92e56f26f745329384f9472c

Observation b40e162f-bb38-4acf-a8b4-4ee07e587fdb · inbound

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild cites this paper.

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-19T12:57:17.713523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T12:56:30.634284Z digest=sha256:1e1cd072eff64ed066d752a6f7065890c6bcba3893299d63cd3267ccc607e7a0

Observation e4f598ac-8b86-4fea-86e9-ae2af88e63ad · inbound

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis cites this paper.

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-19T12:02:16.767815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T12:00:37.025335Z digest=sha256:e962d6ffb16fac8024e3b6e92455e4dfcf478e9aa8c025a3aad53b97fe8cf8b4

Observation 201bc43f-a55d-4a40-abd1-5a4ab1bd16a4 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:33:50.706975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:1fe5fa619906397ee2b043321b67f7d1f25fe85ee5b20ca90939637b8536e7ce

Observation ec3b4777-ccef-40fa-bbb0-fd13f8f9a177 · inbound

WorldVLA: Towards Autoregressive Action World Model cites this paper.

WorldVLA: Towards Autoregressive Action World Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T22:57:07.947528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T22:57:07.883617Z digest=sha256:cee8f3fba8a257ccca500d839c82af91dd94812b63dccf7a082f8e1ecd44320c

Observation f4a82a89-61d6-4790-b488-fad07dea0890 · inbound

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling cites this paper.

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-19T05:17:06.758003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-19T05:13:28.767788Z digest=sha256:d1f9ff14de4ffe18867f5ec4e48b7d314e707c6a7d615387a9f226c480ed1393

Observation 38d58f02-6377-4c88-8266-ca01c57b2ee7 · inbound

Bridging Brains and Machines: A Unified Frontier in Neuroscience, Artificial Intelligence, and Neuromorphic Systems cites this paper.

Bridging Brains and Machines: A Unified Frontier in Neuroscience, Artificial Intelligence, and Neuromorphic Systems Cosmos World Foundation Model Platform for Physical AI

Reference 184

Resolution
verified exact
local_arxiv, observed 2026-05-19T04:42:04.946578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T04:37:33.928616Z digest=sha256:0d77e2d4528a4fdd675fe5ceaeda9a833af0f2c2f761ab1678464540e67a17b5

Observation 88c3cd79-cdf0-4f59-a40d-b06cf58dcd00 · inbound

Qwen-Image Technical Report cites this paper.

Qwen-Image Technical Report Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:29:06.883874Z digest=sha256:312a7fb0aabbc33ccb28fac4083891a0b94be0c551b4f660a49054e40a4dd920

Observation a726a1b0-1197-4cf7-b00a-1ea8bc50e8fb · inbound

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation cites this paper.

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:28:41.941408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T21:28:41.904725Z digest=sha256:3882de0f4d43a7137ec7c8427fac81540d6832802d13189f5be67640661b351c

Observation 9183893f-1457-4eb2-9d8f-cea04e2bebd3 · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.695393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:8ace70ad9e46ecfbe9da4d46a08738b28c03d15505d6d98f75c0e1d4d5a5dc3b

Observation a8af67dd-e818-4905-9946-171650e69f48 · inbound

Matrix-game 2.0: An open-source real-time and streaming interactive world model cites this paper.

Matrix-game 2.0: An open-source real-time and streaming interactive world model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T22:36:53.350371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:36:31.044743Z digest=sha256:5c22779967d032b3ed0a050a377e80ea7d9f8c8770fcefb701457bd0da96f796

Observation 1daa2b1b-c9ed-47ff-9302-d7b07c4255a8 · inbound

Precise Action-to-Video Generation Through Visual Action Prompts cites this paper.

Precise Action-to-Video Generation Through Visual Action Prompts Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T19:14:34.361533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:14:34.361533Z digest=sha256:07d41a162a5b50c134d5a62e00afe18ae11ceb0c24513ef32933756b31fda7c2

Observation 0f405814-3c69-4ed3-8836-3670c0dcd93b · inbound

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models cites this paper.

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T22:11:52.772377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:10:16.645510Z digest=sha256:2118fc8a2b349972051330675074366cc0271b3a654a671e3e1c78ab3dbfeee2

Observation 51a8f960-5ebe-4255-a418-2d0284ce4141 · inbound

LuxDiT: Lighting Estimation with Video Diffusion Transformer cites this paper.

LuxDiT: Lighting Estimation with Video Diffusion Transformer Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:52:02.100197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:52:02.100197Z digest=sha256:601444fc66a3576aa160898809ca7858e6ea474c901a1d7f32b63cb97d177794

Observation 8d07c304-da9f-4247-b043-8013e0921c41 · inbound

Exploring Autoregressive Vision Foundation Models for Image Compression cites this paper.

Exploring Autoregressive Vision Foundation Models for Image Compression Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T05:36:52.435440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:36:52.435440Z digest=sha256:613bb17c2210a5e662456d0ff6c2d5afd9dc03c806ee6efda183f7ba06710d3b

Observation eda46ae9-ced9-464a-b77b-f65d436b6b31 · inbound

3D and 4D World Modeling: A Survey cites this paper.

3D and 4D World Modeling: A Survey Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T06:04:02.866781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:04:02.866781Z digest=sha256:ac841bd6e9aad921b2ef076c3a7ace9e90881f7e481f9b0944ac991d34380e8f

Observation 5a788f97-59ff-4b64-86ba-6431ab55bc65 · inbound

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility cites this paper.

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:56:24.406818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T12:55:42.679016Z digest=sha256:553e294971599b3078460c54835b2c674c6c56d8caf67144468d338ba81b4b8f

Observation 64df5fca-a99b-4165-9c84-2164f926d364 · inbound

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation cites this paper.

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation Cosmos World Foundation Model Platform for Physical AI

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:52.135555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:52.135555Z digest=sha256:33f000f4ec46077a2ab736d6a121f34f55439bde830279161b5fc0d5ecc58ca8

Observation 6453827b-e935-4b90-b87e-1381952a8a0e · inbound

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models cites this paper.

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:21:23.781604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T13:21:02.738561Z digest=sha256:6cf6fc09106536fe9a2a8cf43103d3e9ff9f3db7391f25cfb561e0baec109254

Observation afd23de3-b2cc-4395-9129-b0f4404e7310 · inbound

SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization cites this paper.

SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T11:25:33.514720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:25:33.514720Z digest=sha256:c4e6f50099a374a43f0926a6d472409a8ca71ed8285d7545dbf1a9d89be7dd8a

Observation cad30db5-7e0d-47a5-aa2c-b4e6a4c5b74d · inbound

VChain: Chain-of-Visual-Thought for Reasoning in Video Generation cites this paper.

VChain: Chain-of-Visual-Thought for Reasoning in Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T13:21:35.661138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T13:19:50.252345Z digest=sha256:b8641ad713e8a00d2b06067a60500bb1afbe2a3e09c106245aaedbf50ba2090e

Observation 62134817-4749-42ed-a948-0a647c7ae0ff · inbound

Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency cites this paper.

Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency Cosmos World Foundation Model Platform for Physical AI

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T08:51:09.155679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T08:46:16.541104Z digest=sha256:77edaf67ee8ab41534e2bfc20500eb66d4d51ad125f177092b8db5115c7a25c6

Observation cb8ef6e5-5063-4c47-b854-74fbad2a2bca · inbound

Ctrl-World: A Controllable Generative World Model for Robot Manipulation cites this paper.

Ctrl-World: A Controllable Generative World Model for Robot Manipulation Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T01:14:10.550114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T01:14:10.174044Z digest=sha256:02fa69f2868f9c15e51aaff628d0f2e734b066182899495c209f56eaa5b1eecd

Observation d6c2b98c-7420-43ca-ab21-8a9e0bd5f380 · inbound

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer cites this paper.

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:30.449145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:54:30.449145Z digest=sha256:50cacb1b84da59e6965fbdb05f7aa9e7136795c4f45947df2eea7d78b93aec6f

Observation 83f7ae15-d0ae-425b-a7b1-22c1e2a35cef · inbound

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation cites this paper.

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T03:20:48.732645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T03:20:35.021120Z digest=sha256:7255d9942d98d98b2f269b5cb9786a797e5b4810f9d9fb37bf1e02b3bd4553df

Observation 6f8d0196-3608-45c3-a9f9-cdf0d070daea · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI Cosmos World Foundation Model Platform for Physical AI

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-12T23:01:13.716640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:fac4760bb607814dbe25ff5e78664707d4f6cd012742bff247ad3ca758bce711

Observation 756e2c06-caf1-4485-8a01-b0dcfa310336 · inbound

Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction? cites this paper.

Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction? Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:20:29.138865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T18:17:53.581119Z digest=sha256:dc453bb9e20b53cfe2a178b629ff57adafe4d93bb8adfb8000ad04761d3f6953

Observation 5c57004e-ad8f-4ff7-91ab-003a995d78e4 · inbound

Saving Foundation Flow-Matching Priors for Inverse Problems cites this paper.

Saving Foundation Flow-Matching Priors for Inverse Problems Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:40:15.123284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T20:36:13.408128Z digest=sha256:c3a712e4b21e1922510b3f9bc2b2aec3293e95853537bb6419e90c1e5501c7f5

Observation a5fcd7b0-61c9-4d43-856c-684c59195325 · inbound

RynnVLA-002: A Unified Vision-Language-Action and World Model cites this paper.

RynnVLA-002: A Unified Vision-Language-Action and World Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T20:59:49.860461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:59:49.860461Z digest=sha256:101fe648d22aa6cb6ed51afbf0ded80b5d6adb89f001fcf249344632be8b2dc7

Observation 77cee43b-b4bb-4bf9-b536-b608f6e86948 · inbound

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? cites this paper.

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? Cosmos World Foundation Model Platform for Physical AI

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:02:04.183417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T20:00:20.895672Z digest=sha256:3cbc18825bac288b806cbc93fba11e426ed26e7922962cdad17c60bdf43ce2f7

Observation a7bf38da-7c00-4a99-8fc4-46f9bb600904 · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-17T05:59:08.747937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:0c400374811c7537530d4e3aa36fa0e835009bbc29dd10cb1bc426674fb11487

Observation a1cd09db-40be-48b9-9554-72809d1f5f1f · inbound

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents cites this paper.

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T20:42:52.881712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:42:52.881712Z digest=sha256:83d1d259b1e5366ef2c3e0604f0ab6dbfc8af0deda7b2504ee4f18e77ec674fd

Observation 711f6b55-66d8-4b88-a717-2556747a3ae3 · inbound

Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations cites this paper.

Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T20:23:51.494699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:23:51.494699Z digest=sha256:4c402cb97df4108df2fe67b6c787e36e105338256bf4dbb861dff9edfc57266e

Observation 2bd4df5f-9306-4039-a52f-ab575e28b0f9 · inbound

RubricRL: Simple Generalizable Rewards for Text-to-Image Generation cites this paper.

RubricRL: Simple Generalizable Rewards for Text-to-Image Generation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T20:15:25.349915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:15:25.349915Z digest=sha256:d6b9963fde0581e31b8c2a64f1c8f293b7fcbc9d9bc58c73841d15506d8429ac

Observation be46b26c-0823-4a42-b745-796776d3d43e · inbound

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection cites this paper.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.485045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:549f1081023c0d21d37fc3972bf1f2f1c4667f97c1a1376d65d8d86312e6f3fe

Observation 211eaf85-e813-4e71-899a-5357d85588b3 · inbound

Audio-Visual World Models: Learning Physically Grounded Multisensory Dynamics cites this paper.

Audio-Visual World Models: Learning Physically Grounded Multisensory Dynamics Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T19:25:08.357752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:25:08.357752Z digest=sha256:1cccc5ef53ef203714a3d3efa86ba37b3f745fc2c4578d38b65fbf1f72b6c1b2

Observation 74c3bca8-3193-417a-a7f6-fdd76c6d43cb · inbound

IGen: Scalable Data Generation for Robot Learning from Open-World Images cites this paper.

IGen: Scalable Data Generation for Robot Learning from Open-World Images Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-17T02:58:54.813530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T02:58:36.214948Z digest=sha256:448270d5d66c05dd36d6f4406fe0bd214356e06487583524d4f51ecc413279c3

Observation 73cf5fbb-b26c-4b2b-81dc-be8c27004265 · inbound

ProPhy: Progressive Physical Alignment for Dynamic World Simulation cites this paper.

ProPhy: Progressive Physical Alignment for Dynamic World Simulation Cosmos World Foundation Model Platform for Physical AI

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-17T01:08:47.981319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T01:05:08.087136Z digest=sha256:7ca41840a58bb2041ab5e85228ee7b0522743d2d0fb79e9054950047c1f6d61d

Observation 43947ee7-2fea-4c6b-bf40-2727267b96de · inbound

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis cites this paper.

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis Cosmos World Foundation Model Platform for Physical AI

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T16:49:00.987676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:49:00.987676Z digest=sha256:ee2b7df2faca0c9d4a836eeaa0a85ffbe52f942c5acc8d723cf990172a956a08

Observation 38bb136d-d8b4-4729-8a71-75e6f370fa13 · inbound

Video Deepfake Abuse: How Company Choices Predictably Shape Misuse Patterns cites this paper.

Video Deepfake Abuse: How Company Choices Predictably Shape Misuse Patterns Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T19:58:59.464724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:58:59.464724Z digest=sha256:b53d5aabefc0debe505dbe16f925d529257564ee8309625022e534e2b39805ca

Observation 858787d6-fa51-4b56-92a6-ff28ef52fa6f · inbound

mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs cites this paper.

mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs Cosmos World Foundation Model Platform for Physical AI

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-15T10:41:00.367388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T10:41:00.142543Z digest=sha256:9d4fb6e0bb765808aa954fc378a34af73b1bfbca37f6f131c45f192cec89c767

Observation ceeb5cf4-eb86-498d-adf1-6641f4e749c6 · inbound

Large Video Planner Enables Generalizable Robot Control cites this paper.

Large Video Planner Enables Generalizable Robot Control Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-16T21:28:34.020511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T21:26:32.048309Z digest=sha256:d3e3e93a7b06d16fc3e8e3a3ac60560aafea7f5b05c2c8792c48cbb31884ca06

Observation ba64bb5c-c7f1-4526-9114-ed5890ba1751 · inbound

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure cites this paper.

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T20:11:13.682105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:08:23.319013Z digest=sha256:de508604bdb28333f7fc50416103e57b5797ff3b101f44a9c38b8217c7f98931

Observation 636a5f05-1042-4f9f-945f-4be152fb0d18 · inbound

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure cites this paper.

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T14:09:40.252101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:09:40.252101Z digest=sha256:841681e4d99ef519d2a676ab732ce530c56a1a98325ecc5ca542c723342aff86

Observation 9f59abb0-699e-4953-91fd-ba706899ec89 · inbound

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models? cites this paper.

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models? Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T15:34:15.044268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-21T15:33:24.616338Z digest=sha256:9c60ebf14232336542690c424a44c30c5fca5f7e02a97e90a189181a4c8cbc77

Observation 181564d3-dc81-4f08-8f1d-d5f92983b899 · inbound

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation cites this paper.

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T13:24:42.374669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:24:42.374669Z digest=sha256:435c43b0654bb73dd7a65a2b475bc2284d4de7b0c3061fff7f0c817820686f42

Observation ccaca35a-8057-4a75-94af-df64aa2e4afb · inbound

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments cites this paper.

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T13:00:51.091098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T13:00:51.091098Z digest=sha256:7b4e595a495db805a34e5753570d46ec10396b913e46a6ae68fa905f994f096e

Observation d04d3857-195e-4b86-b348-3f59a965b0cb · inbound

Transition Matching Distillation for Fast Video Generation cites this paper.

Transition Matching Distillation for Fast Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T10:35:03.866071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:35:03.866071Z digest=sha256:bd9882bd0b88725c7f00edced298856cf587484b928d2ac98f913e23c5ede63f

Observation d4c200fa-415a-479e-957b-99c130cd7669 · inbound

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos cites this paper.

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:47:57.481566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:43:26.460480Z digest=sha256:216001e113d6df268271c2b9e25a1253951c68b8c660691fa5b47ed713b75751

Observation d2e3beac-fcce-4933-ba2f-80d6c19bdd0e · inbound

Walk through Paintings: Egocentric World Models from Internet Priors cites this paper.

Walk through Paintings: Egocentric World Models from Internet Priors Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T08:56:07.968450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:56:07.968450Z digest=sha256:38c9d48c09ad197265a241bc506dc69b69e7f3819da6aa21b4bb9a94f1895912

Observation b0bc8a33-549a-4796-8fff-16617616b692 · inbound

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning cites this paper.

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning Cosmos World Foundation Model Platform for Physical AI

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T14:50:12.999217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T14:50:12.804707Z digest=sha256:5ff09e85bbd48a51c4d9a5f4b5cf069108854e76f17cd3e61c4d09dcc5aae154

Observation 10b34efe-3635-4916-bd56-ec9093d41e16 · inbound

Advancing Open-source World Models cites this paper.

Advancing Open-source World Models Cosmos World Foundation Model Platform for Physical AI

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-16T09:07:00.937055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T09:07:00.904794Z digest=sha256:ad34c84427a7113b7906fae199eca13c7949889e2da74d04d22828439edbfdb7

Observation bd185ba4-808b-4e23-a6f8-db91cf9a7b4e · inbound

Enhancing Table Reasoning with Deterministic Table-State Rewards cites this paper.

Enhancing Table Reasoning with Deterministic Table-State Rewards Cosmos World Foundation Model Platform for Physical AI

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T14:10:13.658249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T14:05:27.593292Z digest=sha256:1e44947e077f284ac7ce43dff582fab5773d5c2f0f286ef8ab14fdb9976a9ac1

Observation 15d02536-cee8-4c96-b8d7-dc179a09a145 · inbound

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation cites this paper.

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T09:17:40.251530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T09:16:57.528410Z digest=sha256:59408e4e33db95e332d80d5202e0b2a07fcefea2526b9bbc97005fd1e1206157

Observation f1c434d1-3c70-4116-ac32-ff3e423cb418 · inbound

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation cites this paper.

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T06:13:19.913571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:13:19.913571Z digest=sha256:212142f56869a5950d5f0d589a8209f73885bd5965c9346384ae986f4759cc1a

Observation 30c17424-e16a-4f18-82c6-7058bf544a67 · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T07:07:30.035725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:b2f61cba90cba75c5f776a204dc2b156582fe81dd4a492fba8fa85c577fb22a1

Observation 52071e9b-e375-4f87-96a4-3c7465b46f9b · inbound

Olaf-World: Orienting Latent Actions for Video World Modeling cites this paper.

Olaf-World: Orienting Latent Actions for Video World Modeling Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T01:20:04.586979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:20:04.586979Z digest=sha256:8d2bcfce5300fe89c27f2c8662d2a1c40b5bd4fdf1bb23a29af260ae8873ee53

Observation 3f038c34-90df-49f2-afcb-3075ba7fc046 · inbound

Wireless TokenCom: RL-Based Tokenizer Agreement for Multi-User Wireless Token Communications cites this paper.

Wireless TokenCom: RL-Based Tokenizer Agreement for Multi-User Wireless Token Communications Cosmos World Foundation Model Platform for Physical AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T23:56:42.235745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:56:42.235745Z digest=sha256:63e684879a52b43e9fcc87c1fa874e0766d64aebd26f6f9ec8403f96077b24bd

Observation 42ff98a4-1b53-4c60-89d9-12c4795bf03d · inbound

PhysMem: Scaling Test-Time Memory for Embodied Physical Reasoning cites this paper.

PhysMem: Scaling Test-Time Memory for Embodied Physical Reasoning Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:06:33.785467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T20:05:30.324309Z digest=sha256:1863be409b29bf452141075bc822115a1cd034dce9cbc0ac16dff48e9308d041

Observation 60e3c68c-8eb5-4bd4-8e56-2a0cbc9ad40c · inbound

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving cites this paper.

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:54:10.215735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T12:53:29.732960Z digest=sha256:4e506cc98206123628170be1a82b2b616428a0fd2d4c39f5191ebc6196476ce8

Observation 0c177d14-f711-4fc1-b9d0-379d73c2ede4 · inbound

Online World Modeling Enables Real-World Inverse Reinforcement Learning from Observation cites this paper.

Online World Modeling Enables Real-World Inverse Reinforcement Learning from Observation Cosmos World Foundation Model Platform for Physical AI

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T20:06:56.472771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:06:56.472771Z digest=sha256:b76e3e65450e200278cdea9e945a797f6d30f79cfbd92f46b58a36e845e2be10

Observation 4b5bb698-928e-4482-81d8-40a46a90de06 · inbound

RoboLight: A Dataset with Linearly Composable Illumination for Robotic Manipulation cites this paper.

RoboLight: A Dataset with Linearly Composable Illumination for Robotic Manipulation Cosmos World Foundation Model Platform for Physical AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T18:56:44.628502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:56:44.628502Z digest=sha256:9b6ac299d16408e33af3f61ae7290e1fc834b45277446f01a03c5be0f8c23069

Observation a13d0ec0-85ed-41de-b0c5-b1d57bb680e8 · inbound

Margin in Abstract Spaces cites this paper.

Margin in Abstract Spaces Cosmos World Foundation Model Platform for Physical AI

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-15T13:24:52.801459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:24:52.801459Z digest=sha256:16a42303d3b69254519bd8f4b6b05160eee64fa58b3d6eabdfaa928d2d19dd29

Observation 993a34eb-7fff-414b-89ea-a63f667cd230 · inbound

SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents cites this paper.

SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-22T11:16:27.930103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T11:14:47.943242Z digest=sha256:8797aeafd81f4ed6a546e62e5d0381ae48fc9862e0cf65004a183cf01e9fc471

Observation 8925d46d-82f7-4bc2-897e-43b534052194 · inbound

PlayWorld: Learning Robot World Models from Autonomous Play cites this paper.

PlayWorld: Learning Robot World Models from Autonomous Play Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:05:54.664447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T14:05:32.112432Z digest=sha256:8f530445de0f36e4d8cdf98704c2a3abfe68b6ff7527550ba9427b45528dcc31

Observation 7ea16b4b-91e8-4be4-bb6a-aee346cb3915 · inbound

InSpatio-WorldFM: An Open-Source Real-Time Generative Frame Model cites this paper.

InSpatio-WorldFM: An Open-Source Real-Time Generative Frame Model Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-15T12:09:59.870496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T12:07:17.341611Z digest=sha256:e6a6e1dc12b6f194ca73b77dafd49ec99bbcb1c2d6d454b7cc436d06d53fdaa7

Observation 27ad37cb-fc61-4d3a-921f-c4c6297c1942 · inbound

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization cites this paper.

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T12:15:34.517448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T12:13:50.985113Z digest=sha256:87b779966cb1f05feae5fd3e88054f1d8e9cd9621960c433a7d078cd85e43d70

Observation 1864cfd5-3db4-4184-8265-e6bbd6d3739a · inbound

WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems cites this paper.

WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T11:44:09.365900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T11:40:28.002606Z digest=sha256:456f15078f0ec3c393ec4460ee2fb7d243dc592f383690fa9dcec6ab2f59dd7a

Observation 8775c1ec-38dc-416e-8e2f-812962038f1b · inbound

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion cites this paper.

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T21:09:36.040879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:09:36.040879Z digest=sha256:03700e9dc1befc1a045ddee28b3afc7a12c2b486324490386ae639b84aff59ad

Observation 87789918-47d9-4c03-a305-cf33b8d3ee3e · inbound

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation cites this paper.

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T22:13:53.383917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:13:53.383917Z digest=sha256:3568b2a0693733498787bc3a7700ab455aec53173e55b90d0b3b89e40a49db76

Observation fb390ba4-af4a-4bac-bc95-70579efcc48c · inbound

Under One Sun: Multi-Object Generative Perception of Materials and Illumination cites this paper.

Under One Sun: Multi-Object Generative Perception of Materials and Illumination Cosmos World Foundation Model Platform for Physical AI

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-13T22:08:47.493022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:08:47.493022Z digest=sha256:8e56d1a309c886aa3bcc0e1a7a63c9e2a7242b7ffbaa1063eff6694f1f647963

Observation 95c95053-f670-446f-b97b-516ee1b20c61 · inbound

LongTail Driving Scenarios with Reasoning Traces: The KITScenes LongTail Dataset cites this paper.

LongTail Driving Scenarios with Reasoning Traces: The KITScenes LongTail Dataset Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T00:33:23.359419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:30:56.156113Z digest=sha256:d44e8f51de1d72634c4903e19241ecbb834ed8d0a5a555cd05ae7ab2e959ee87

Observation 82847d3d-fa08-45f5-ba93-d6213bf26fd5 · inbound

KappaFormer: Physics-aware Transformer for lattice thermal conductivity via cross-domain transfer learning cites this paper.

KappaFormer: Physics-aware Transformer for lattice thermal conductivity via cross-domain transfer learning Cosmos World Foundation Model Platform for Physical AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T13:13:22.720572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:13:22.720572Z digest=sha256:c679149b07267dc6b451388ea70918bce7c4f1466f1ef9f27e07a878d8fdb3ba

Observation b8bb0f46-4fb4-4572-bf02-82d309fcc8d5 · inbound

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models cites this paper.

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:36:42.100191Z digest=sha256:7ddf06082147c16889127bf0f7feee6bd602506efe4828d5e0a7cb0b568e5572

Observation 414af9e3-47d4-4398-a455-fc112d2d59d6 · inbound

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models cites this paper.

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T09:42:23.808691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:42:23.808691Z digest=sha256:69a60453a585eb86315e4f63536246f030c72a247a5abf936a2964662c8a38e0

Observation 3a21de28-d7a1-4674-b5f8-d8760200fb4c · inbound

A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens cites this paper.

A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:19:17.796377Z digest=sha256:b541b838d1498ee3d2fe48ab71ca36b79cfd1d2f22f85852ed8909c6bbbc5018

Observation 7b79b170-9f89-42c0-bfb4-b88a004fb762 · inbound

SEM-ROVER: Semantic Voxel-Guided Diffusion for Large-Scale Driving Scene Generation cites this paper.

SEM-ROVER: Semantic Voxel-Guided Diffusion for Large-Scale Driving Scene Generation Cosmos World Foundation Model Platform for Physical AI

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:32:52.012404Z digest=sha256:a8835a25671900fb2e94c5004e53ff69e1d19b64c170af30144b4163d02d511d

Observation 4f91832e-0a56-4bdd-9bd3-c941dad62228 · inbound

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models cites this paper.

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:38:47.256647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:04:24.131372Z digest=sha256:1c50a0fcc62e7cdad145e4b2cfe58a00c76aacce71401099940fd0a013c11202

Observation dcf05559-e7c4-4e33-ac40-5b49004e2221 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations Cosmos World Foundation Model Platform for Physical AI

Reference 171

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:05:51.482297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:ef2aa0bfa9527afcdeecbc884f0db14577e799dc992bc0dacc0c7c8fe8b7ba44

Observation bf1be9f9-7169-4c40-a272-b521a248a317 · inbound

MoRight: Motion Control Done Right cites this paper.

MoRight: Motion Control Done Right Cosmos World Foundation Model Platform for Physical AI

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-11T06:26:01.040529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:38:03.776766Z digest=sha256:bd88161df1ffcda77c2abda18269ffd3f56751099f9cfb2137417ae88d4d4a09

Observation 26da545e-bb01-43d4-a47e-59faffd08ded · inbound

ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks cites this paper.

ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T06:01:02.673365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:49:33.822264Z digest=sha256:60ac7ddb4212930b9b85309f9878745a4a11899055cf97045420be9fc93feea9

Observation 0c9d35fa-75ec-41fc-b44c-5129633d8fb2 · inbound

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning cites this paper.

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T07:25:59.374639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:12:08.970164Z digest=sha256:e25a857ef5d17e47ad50e615e84c6cb766b0fed0894cf4530433e41725401497

Observation 192c2629-4c58-4c29-b2ae-2a7586f78c84 · inbound

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence cites this paper.

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T00:03:53.609175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:03:53.609175Z digest=sha256:aa05f0d91b0847abdb708f7618808c9d938a2c6eeed86fae9d7e6dae26dc53c8

Observation 23fedcdf-2c3e-4404-8402-568bcf0a6be3 · inbound

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics cites this paper.

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:10:55.625340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:15:37.338442Z digest=sha256:d703b4199b72cc4aecf9dcdfc54678d8e0c7491499863bf423cbec3dab916bde

Observation 77d45919-e4c3-4ad8-821f-51a23bee6436 · inbound

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics cites this paper.

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T09:19:56.623509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T09:17:53.731701Z digest=sha256:612bbdf5e4a86e8e74f6e05e3d6b8f525abbbc14decfa6c63257a0a2fc4c0502