Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T00:23:03.822300Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2607.15054.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T00:23:03.822300Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4afc75ac-6588-4b67-9378-2ffc5ef46892 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f649197a-b618-408a-87f8-e57038564a1e · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Grounded 3D-LLM with Referent Tokens
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8b896e1-377a-4bf6-b5de-c6c188cc415e · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4675c47a-7b5a-4ab3-b24f-83a706477a0e · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05707b40-b89d-40bf-8e03-39d0298d5072 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding GPT-4o System Card
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87e64744-d2a1-4358-87b8-233278902a8f · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Adam: A Method for Stochastic Optimization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e88b47d-566b-43b7-9ad8-e1851429c6fc · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Thinking with Geometry: Active Geometry Integration for Spatial Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f189430-0882-4be4-8ef6-5b005ae1de66 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Depth Anything 3: Recovering the Visual Space from Any Views
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a7fe89a-a414-4b7c-9969-fa9de197a021 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Trace anything: Representing any video in 4d via trajectory fields.arXiv preprint arXiv:2510.13802, 2025a
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf8cbbc5-f5d1-4423-802f-21ab9d50271a · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding SpaceR: Reinforcing MLLMs in Video Spatial Reasoning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3b06a84-d677-4a0a-9a70-d483a1a880cd · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edf82e13-76f9-433c-902f-45a4153ecd22 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cac0151-ae00-4379-8ced-26b44dd6e06d · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding LLaMA: Open and Efficient Foundation Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21e08e6c-d6e8-46fa-afc1-3e2ea1de8f1b · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Wan: Open and Advanced Large-Scale Video Generative Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce1e85c9-3eed-47d0-9f27-d9e51c7d7277 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f91d52c3-b8c2-47b3-a065-f8004863312e · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a94d5c6e-44f2-4f55-b3f5-195fd52981b0 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7437ff82-5dab-4f3b-bb6f-57ef258a53b1 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding From flatland to space: Teaching vision-language models to perceive and reason in 3d.arXiv preprint arXiv:2503.22976,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a305829a-724f-4163-949f-21405c8dc6f3 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding SpatialStack: Layered Geometry-Language Fusion for 3D VLM Spatial Reasoning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a174ab-bebd-457d-9885-22696f41c3b0 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Long Context Transfer from Language to Vision
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ffbde95-a8de-47a6-b3f6-8568ef4d81d3 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Multi3drefer: Grounding text description to multiple 3d objects
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7f2432c-0cb8-4974-abe2-f3bc6d064307 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Unifying 3d vision-language understanding via promptable queries
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c891bda5-50ec-4552-b0e0-f6f66ad68479 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding LLaVA-OneVision: Easy Visual Task Transfer
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4ec22de-4f42-4007-8a85-dc08415a53dc · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Seeing through imagination: Learning scene geometry via implicit spatial world modeling.arXiv preprint arXiv:2512.01821,
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ea6e02-6a1d-4020-adf5-02c703efe42a · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Qwen3-VL Technical Report
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71d197b1-e280-443c-b1c5-c20763002696 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding 3drs: Mllms need 3d-aware representation supervision for scene understanding.arXiv preprint arXiv:2506.01946,
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ba8b9c3-c4ab-43e9-9f7c-f875897ca14e · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a027346-4e82-4a16-966d-da6358e1aa4e · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c4a1fa3-70d6-499e-b42b-d96f56b54eb2 · outbound
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Improved Visual-Spatial Reasoning via R1-Zero-Like Training
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.