Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T05:06:02.556817Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2608.07409.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T05:06:02.556817Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 24cbaa79-2054-4393-8f54-e63325486322 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc7a71d-ba13-4532-8ecf-be3043235c94 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling Learning and Leveraging World Models in Visual Representation Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30bd77c7-f4ca-4cbb-a6b0-8ed961ac26bf · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 060cbd3e-4b45-4538-a678-032f4d4a006c · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f78f59b3-f56f-4c3c-bb46-2df38e4b4c4f · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6a2148f-a133-4997-aaa7-aa526c13740e · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling Implementation Details We use a ViT-Small/16 encoder (15M parameters) for control and ViT-Large for large-scale image/video representation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5865df4c-7a22-4d94-84a9-4a385d0569b1 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling World Model on Million-Length Video And Language With Blockwise RingAttention
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fda6f588-d0ca-437e-99fd-4408a6368685 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling Mastering Diverse Domains through World Models
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fb21eff-7dbb-464a-8316-875e9d85aca7 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling World Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f1d49ef-b69e-4614-9827-9b04cfb27af0 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling OpenVLA: An Open-Source Vision-Language-Action Model
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 352b986b-b00f-4e1c-950a-b0af7760ac81 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling Back to the Features: DINO as a Foundation for Video World Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d969f65-d119-40c4-8b92-833dea55a436 · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da6c6179-14b6-45d5-a77f-5e42090c050b · outbound
UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.