Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:02:43.614624Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2608.04726.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:02:43.614624Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
17 of 17 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7a482709-b5dd-4cb3-a056-9dfd3d727272 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db58b1e2-504e-4452-9dae-b83961accf95 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Li, B.; Zhang, Y.; Guo, D.; Zhang, R.; Li, F.; Zhang, H.; Zhang, K.; Zhang, P.; Li, Y.; Liu, Z.; and Li, C
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9133902b-5ab7-40d2-9911-fdede827cc13 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Lin, H.; Liu, Z.; Zhu, Y.; Qin, C.; Lin, J.; Shang, X.; He, C.; Zhang, W.; and Wu, L
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 073a3ff1-8751-4f9b-9003-9930983f4ffa · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning arXiv preprint arXiv:2601.21821
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5082ee0b-8a44-460d-9ebe-fc329006bb18 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4571943-360e-4046-9080-7be6d136e523 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 182f00bc-48d2-4ba2-9971-1b52146dab92 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcdee749-e364-4976-8046-a77cfebff434 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Zheng, C.; Liu, S.; Li, M.; Chen, X.-H.; Yu, B.; Gao, C.; Dang, K.; Liu, Y.; Men, R.; Yang, A.; Zhou, J.; and Lin, J
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e59df4-17db-4dd9-b4fb-0c5a31f7e731 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Group Sequence Policy Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 393aa8f8-c4fc-49da-9c84-44ecef1a648a · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa9de672-90ac-483d-bee0-3e604dfe204a · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning InInternational Conference on Learning Representations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1780b0e9-a7ba-4a10-923c-93daf34ea924 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4a02599-d57b-4a32-83b1-4d2d1094bfc0 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Kimi-VL Technical Report
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f726666-04ca-4bbe-834b-124876c1cfb1 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b70000d-798d-4848-b40c-80934cbca12a · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43834796-eb56-4b6b-85e4-d0ba82abc155 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning Assran,M.;Duval,Q.;Misra,I.;Bojanowski,P.;Vincent,P.; Rabbat,M.;LeCun,Y.;andBallas,N.2023
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e1ffec3-0e69-4bff-89f1-4d1b2d5f4022 · outbound
When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning arXiv preprint arXiv:2602.09483
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.