Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T03:16:53.178921Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2607.14187.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T03:16:53.178921Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
56 of 56 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1483f19f-b5a1-418b-8d45-1f748282ee72 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Cosmos 3: Omnimodal World Models for Physical AI
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fe4b1d2-fd68-4961-9478-07f5cdd54d21 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70ee18af-e177-4d1f-9c81-642953eb68a5 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Revisiting Feature Prediction for Learning Visual Representations from Video
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86b44128-0ad2-4f01-bd20-3da95425aaf3 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 441af304-fadf-4ee3-b89e-8502d1c10f96 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bb7f226-a638-4647-b3c9-cfb9cf2cc548 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Junhao Cai, Zetao Cai, Jiafei Cao, Yilun Chen, Zeyu He, Lei Jiang, Hang Li, Hengjie Li, Yang Li, Yufei Liu, et al
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a61824be-2977-4df6-9356-7133a8dc1da0 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68faf5a4-ee68-4e77-89b1-77a87726cd0d · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b8cd26f-09cc-4d5a-9b3c-577f5a6a43eb · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Rynnec: Bringing mllms into embodied world.arXiv preprint arXiv:2508.14160,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f293a8c0-a783-42f7-aead-801df76ed549 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Rynnbrain: Open embodied foundation models.arXiv preprint arXiv:2602.14979,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a2bd0dc-7117-4fdc-95d2-d04bc7d30eed · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Emerging Properties in Unified Multimodal Pretraining
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2b96bf6-d4d6-4a6f-94f3-38a77bbca105 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Rethinking video generation model for the embodied world.arXiv preprint arXiv:2601.15282,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80c327f4-c8ec-48bb-b4fc-cc9e65c1ca7a · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae0e4585-db95-4289-9304-68fa09bea061 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination MolmoAct2: Action Reasoning Models for Real-world Deployment
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdbeaf69-9a4f-4259-9131-c8080434fb10 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Galaxea Open-World Dataset and G0 Dual-System VLA Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a3fb7e9-2d9b-463f-8bcc-297c6f6a4f2a · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62f09888-3c15-4368-ad0b-20dc82fe5cfb · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination GenEval: An Object-Focused Framework for Evaluating Text-to-Image Alignment
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd55c7f-8b3e-4ec8-8f42-5aa9b0665447 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Ctrl-World: A Controllable Generative World Model for Robot Manipulation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0a9d072-836c-442c-93ce-1a0ff3a5bab1 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0aac4a1-4c52-4a07-8567-fe51c20ab1d0 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ef8783-cc27-4d87-884f-0ce71a4da012 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a66bdc8-57aa-487b-9a80-58a44eb08f65 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbfd2e79-e0a3-4186-baae-1dd0db7044a4 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46844f05-3ffb-441e-b5fb-d811455e5487 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 532ddc54-a7fd-4dee-b38a-6a75134d972c · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 827494e7-2d97-42c6-a2e7-265c38e680a0 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination VideoPoet: A Large Language Model for Zero-Shot Video Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e67efccf-9c77-47e4-b0c7-f19c09994fc3 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination MolmoAct: Action Reasoning Models that can Reason in Space
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bff5aa85-5343-4384-84ec-f98fa6fa5973 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Viewspatial-bench: Evaluating multi-perspective spatial localization in vision-language models.arXiv preprint arXiv:2505.21500,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38d59eae-cda7-4910-88f3-e3c8c5314bd5 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Causal World Modeling for Robot Control
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9ba8fe4-bfa8-403c-af0e-cca74d46f49d · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Mm-act: Learn from multimodal parallel generation to act
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4dd2bde-4167-4e29-b1dc-c66c53065d69 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6466870e-29bd-446b-90c4-0c726a26244c · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination URL https://www.biorxiv.org/content/10.64898/2026.05.01.722168v1
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b23715cf-5277-4c16-9e62-edd407dc59f3 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f63dfd1-1926-48cd-a770-f80e3f759670 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination GPT-4o System Card
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3ea2a64-7297-40c0-9390-389db261c6e5 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37db1346-387e-4609-9120-c79a43eb0db7 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination EgoMe: A New Dataset and Challenge for Following Me via Egocentric View in Real World
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4dd7ce6-2cbd-4d6f-a17e-5507dff19f90 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination RoboCOIN: An Open-Sourced Bimanual Robotic Data Collection for Integrated Manipulation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c86811f8-6d5f-427d-a644-7354c8e4db39 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination OpenAI GPT-5 System Card
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42c0d821-160d-4137-aaf4-e782ffa90d8d · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination BAAI RoboBrain Team
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64d1afd2-ab83-4882-81a2-d358a7b7d2ad · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Gemini Robotics: Bringing AI into the Physical World
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b676f83-b117-40c3-ba69-e0484e85f341 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2512f4c5-febf-4f33-917f-de470cb3ce5c · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Wan: Open and Advanced Large-Scale Video Generative Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9378d785-8a77-4277-9a84-06c3b3a4c87f · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9cb534f-f35d-4b70-b06c-c97799f76cc4 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d8f654d-eeb6-4763-bd63-eea21781b577 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination A Pragmatic VLA Foundation Model
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76dec642-57e2-4091-aa9d-cf732843a2e0 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination World Action Models are Zero-shot Policies
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1eafb56b-242d-48e1-a279-b236400e7160 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58d9205b-0127-4f60-bcdb-574a6b05700a · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6a6cb12-87f5-4a87-933f-14bf1666022c · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db948e56-df8f-40b5-b381-53089309c917 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6eeba8aa-c59b-42dd-b91c-2fd9e9c3d1bb · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Sat: Spatial aptitude training for multimodal language models.arXiv preprint arXiv:2412.07755, 3,
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 037afa4d-f0de-4652-86ed-d45f604371b3 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0af90b43-9b58-4c5a-bcc3-e10ae3acd14c · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Qwen3-VL Technical Report
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73623032-e266-4c79-a3e4-206c17193e77 · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2d0bc79-c207-4f12-ab3d-3e6fbcf929af · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44690fa4-1f37-466d-95ad-61a70d875c4f · outbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination RynnVLA-002: A Unified Vision-Language-Action and World Model
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.