Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T02:42:09.922300Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2605.09719.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T02:42:09.922300Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e92edeac-4e94-4f76-87c6-b4391a0fadbe · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Learning transferable visual models from natural language supervision
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ddb2c839-ccc8-48d3-a68d-d30bdf3bd1c2 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Flamingo: a visual language model for few-shot learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 525f2561-b90f-43f9-97c3-16719d2ae51b · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Superlora: Parameter-efficient unified adaptation for large vision models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05ddfdf3-00fb-458a-aa00-d13e6dd877d0 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT LLaVA-3D: A Simple yet Effective Pathway to Empowering LMMs with 3D-awareness
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5adcbd4b-9be4-4c71-b461-ad3354c9de6d · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Palm-e: An embodied multimodal language model
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5be164d3-752e-442c-96f6-7562dc3ef2d8 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT RT-1: Robotics Transformer for Real-World Control at Scale
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 03af8fc4-5078-4c72-9a37-4e7a4568dae6 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Energy and policy con- siderations for deep learning in nlp
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a0717ceb-85e3-4d71-b6f8-a4c976ba5f8b · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Mofa: A model simplification roadmap for image restoration on mobile devices
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b70e9e08-eb19-47bb-9a73-cce51385b84a · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Learning efficient vision transformers via fine-grained manifold distil- lation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 867bcbf9-6916-4fd2-aa8f-67e6d9da22f3 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 662e8702-89fd-43c8-b470-af7fa7a165fa · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Knowledge Distillation in Vision Transformers: A Critical Review
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0ea23fa1-5629-47a3-9ea8-95a88aaecce3 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Distilling the Knowledge in a Neural Network
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 925184fb-6836-4ae4-8527-70562bd4d48b · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Knowledge distillation: A survey
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e25fc2f8-f1a3-4f17-9d91-5120e8c4284e · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Vggt: Visual geometry grounded transformer
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 618d2b07-ca86-4709-ae9d-ece36449725b · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Pf3det: A prompted foundation feature assisted visual lidar 3d detector
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6a03fc50-7737-40db-89a7-ae434cdd3350 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Visual instruction tuning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 20a36fce-7480-4358-a812-70cbe8a1c51c · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 91293638-1e97-4309-b372-4bcd8c441f33 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Vqa: Visual question answering
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation edae610a-5608-49f8-b3ab-43fe9fb6b450 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a3bdaee0-b542-432f-995c-81c4087057a2 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT 3d-llm: Injecting the 3d world into large language models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5ddb7169-7dcb-4369-a3f1-e10f6a854a03 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Distilling vision-language models on millions of videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f48b52ab-4f4b-4fb0-a155-076208245d16 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Chain-of-thought prompting elicits reasoning in large language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 040dcedb-fba4-47a6-bdfd-92ea217dbded · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Large lan- guage models are zero-shot reasoners
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bd00bbb6-2003-42be-9397-05e27e3274be · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Scanqa: 3d question answering for spatial scene understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7826bfbb-6c62-4c52-a68a-ac00e0869521 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT 3d-sps: Single-stage 3d visual grounding via referred point progressive selection
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 48d7229e-af7a-4dea-adcc-ac08d9138960 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0978cd05-e33f-4d50-a9ac-d118983477c2 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT LLaMA: Open and Efficient Foundation Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c31dc4fd-7d8e-495e-8093-c25fb4d29b3f · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7e9b0384-ddb9-42dc-a6f0-6f7487a71550 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Multi-task learning using uncer- tainty to weigh losses for scene geometry and semantics
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 305e958b-9484-47a3-9910-51df8bdcba37 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Scannet: Richly-annotated 3d reconstructions of indoor scenes
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba99407b-ae00-4aae-b5f8-0f22e57927a7 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT Midi: Multi-instance diffusion for single image to 3d scene generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 25791731-a687-4d60-9f62-bb59acbb3940 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0bc4c46f-7df8-4b1d-a268-eed16a45a5f9 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT 3dsrbench: A comprehensive 3d spatial reasoning bench- mark
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 649bd923-e9ca-44a5-82c5-65922f10d106 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8df35154-8f87-4997-9cb5-5fc8610a04a6 · outbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT PaliGemma: A versatile 3B VLM for transfer
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
No inbound Pith citation observations are available.