Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 51 inbound Pith citation observations for arXiv:2506.17218.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:03:11.626570Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 36665cc5-a23c-48f6-bd39-a2d07728ebc6 · inbound
Artificial Phantasia: Emergent Mental Imagery in Large Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6a920d1b-defb-48d5-99e2-44aef54f2b2a · inbound
LaRe: Latent Refocusing for Multimodal Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7755be9a-e860-4f09-9ecd-3c582b502929 · inbound
Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c9dfc7df-df83-4685-b22c-9e0930945a01 · inbound
Forest Before Trees: Latent Superposition for Efficient Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c462f9b0-d339-43ff-b7b2-c46c9af7496d · inbound
MentisOculi: Revealing the Limits of Reasoning with Mental Imagery Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c0febcc-f470-4614-8e5e-a846544225a0 · inbound
Towards Explainable Industrial Anomaly Detection via Knowledge-Guided Latent Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e4899f4c-3a74-4efb-9bb9-cff07c02b3f7 · inbound
Thinking with Drafting: Optical Decompression via Logical Reconstruction Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation da6a259a-1b97-4c95-a63f-63120205b382 · inbound
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 268
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de5ec7b8-5e89-464c-b21a-d2d6639880a3 · inbound
Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 570e609f-3dc3-41ae-9f50-edcad274807a · inbound
V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 612ebab2-b9f0-4a81-80bb-7c8312b62360 · inbound
Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d03206f0-f22e-4813-ad06-6ceae2bd69f5 · inbound
Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 458f7de4-7dea-4b6a-852e-a8888eeb12bf · inbound
Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d4c4efca-738a-4fc9-aaf6-3766dfcdea06 · inbound
Visual Enhanced Depth Scaling for Multimodal Latent Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4c6ad6c9-fdf3-43df-8778-46a1964bd261 · inbound
Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ab543ed3-ed1c-4606-ab82-51888183bee4 · inbound
HyLaR: Hybrid Latent Reasoning with Decoupled Policy Optimization Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation eaf0e5d4-9274-4335-b35c-4c0a6de31ea1 · inbound
HyLaR: Hybrid Latent Reasoning with Decoupled Policy Optimization Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e606c722-48a3-4a3e-a93c-928e86959529 · inbound
Using Machine Mental Imagery for Representing Common Ground in Situated Dialogue Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e55f0731-9dd4-4a79-b85f-7c009246b541 · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 23a8ebd4-8ab7-409d-981b-149f75048b9b · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9a002af4-fc4d-42c0-aece-bc49f199e477 · inbound
Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b3257744-8fbc-4791-9ae0-3a7d3ed086b0 · inbound
Retrieve, Integrate, and Synthesize: Spatial-Semantic Grounded Latent Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 74cd7e14-0902-4f29-aff7-f5af4f4c9fdb · inbound
CoLVR: Enhancing Exploratory Latent Visual Reasoning via Contrastive Optimization Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d0b8133f-092f-4eeb-af0b-6605c8d66833 · inbound
CoLVR: Enhancing Exploratory Latent Visual Reasoning via Contrastive Optimization Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5f99935c-c814-4684-883d-47d9657b2ae0 · inbound
UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1e03b941-5181-467c-84b0-33d1b30f8bb6 · inbound
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1990d930-630b-401f-bcc9-866ca4e740da · inbound
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e38a023-09e2-4746-815b-27852730357b · inbound
Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 80d05957-b336-4b20-81ef-ef27be586518 · inbound
Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4003f055-b9ec-4812-9534-a002f7a4233c · inbound
Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cfd3c977-46db-48ce-b02b-29fbdf57809d · inbound
Semantic-Enriched Latent Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation eb7fdc95-373b-4b52-b6ed-90d17d125bc0 · inbound
Semantic-Enriched Latent Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ead86fe9-3639-44a2-b16f-34c923b51c50 · inbound
STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bb8b15b3-3e80-4378-9f3e-ad26b684b386 · inbound
ReGuLaR: Relation-Grounded Latent Reasoning for Large Vision-Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6ba625fa-0b2a-4d40-a6da-45ce7edfe9b5 · inbound
DeepLatent: Think with Images via Parallel Latent Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 52cd2762-eb65-4afb-ad5e-50ad477b734d · inbound
MUSE: A Unified Agentic Harness for MLLMs Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4a2d7c9e-0a34-482d-b5d2-cb01938830fa · inbound
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 63ddab03-9136-4c40-b3b6-1604641457a1 · inbound
UniCanvas: A Diffusion-base Unified Model for Text-in-Image Joint Generation Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f4828936-c2d8-4ae9-b5fb-d69f471cdd29 · inbound
Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ce7a9a48-e4f8-463f-9f72-d77fae6558b2 · inbound
CVSBench: A Comprehensive Benchmark for Cross-view Spatial Reasoning and Dreaming Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4566a277-a066-4f04-b0c2-2e3f3257905d · inbound
Latent Visual States for Efficient Multimodal Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9a33ec06-f553-4f50-86d2-08e2a6d7f46a · inbound
V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a49e6b72-f388-4693-968b-e6cf3c29b957 · inbound
Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f9abef83-1b04-44f2-82ef-22b99f6802f1 · inbound
Latent Noise Mask for Reducing Visual Redundancy in Multimodal Large Language Models Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 28cff490-528d-4d2c-8319-8c703119c9e4 · inbound
Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2e019a78-9093-467e-877b-216c3b32a019 · inbound
ProLaViT: Learning Progressive Latent Visual Thoughts in Structured Latent Space Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ee2f962-6b5e-4826-b422-3518eb5e0a6b · inbound
APIVOT: Adaptive Planning with Interleaved Vision-Language Thoughts Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 17d753cb-9dbb-475e-9f20-82b549c88feb · inbound
Latent Memory Palace: Reasoning for Control as Autoregressive Variational Inference Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2dd301e0-1288-4f03-aad9-e60e3a6b54d8 · inbound
OPLD: On-Policy Latent Distillation for Multimodal Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27386c4e-2c67-4476-a2b4-570bcd9b295d · inbound
LUT: Latent Utility Training for Visual Reasoning Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9f77a4a-2a24-479b-8e50-eb7f1c101a61 · inbound
Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.