Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T16:45:17.022062Z
Paper Citation Record · LEDGER
As of 3 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 3 inbound Pith citation observations for arXiv:2604.04172.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T16:45:17.022062Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-02T06:30:47.504484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T16:12:23.659005Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-30T12:16:12.909640Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3aad4050-c983-46f0-9cdb-2582f8018c3e · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 29923d39-5047-4c56-a629-c4fddb464879 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models SridBench: Benchmark of Scientific Research Illustration Drawing of Image Generation Model
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 637a6ca9-398d-4514-b9a2-aa414119ae6e · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Figure Captioning with Reasoning and Sequence-Level Training
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 3c45bc01-ce71-4cf9-90a9-a5378ebe2afe · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 4906e079-76c2-4fdb-bc59-428e16635311 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models TIAM -- A Metric for Evaluating Alignment in Text-to-Image Generation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 9ac896bd-787d-49b9-8aa1-0384152d2871 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation aa87a366-a820-4db3-9be1-bdd36d720a5b · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models SciCap: Generating Captions for Scientific Figures
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 1a8df55f-d682-43a3-9150-19411977c5c1 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models mPLUG-PaperOwl: Scientific Diagram Analysis with the Multimodal Large Language Model
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 1a663a3f-b1ac-4b2e-aa3d-8ca67ed77482 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models GPT-4o System Card
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 6ea16a10-3e8f-4df2-98b2-5816654ae642 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation f9e1b439-4b0a-440e-8fe7-509dc4e9a0e0 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models DVQA: Understanding Data Visualizations via Question Answering
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 57e0cd0d-85d4-4dc5-8cb7-eb0fa457a6e5 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models FigureQA: An Annotated Figure Dataset for Visual Reasoning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 36d86ebd-6a76-4eb8-bf23-5d56f63093bb · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 5ab9584a-fef9-4d1f-ba7b-4957d8e3abe5 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 4241a15d-582b-4c18-aafa-a8dd0b191c77 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models MMSci: A Dataset for Graduate-Level Multi-Discipline Multimodal Scientific Understanding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 728949d0-95c7-4b12-8945-2e66611f446e · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Aligning Large Language Models with Human Preferences through Representation Engineering
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation a68b9895-21cb-4fc4-bbda-60bd4415c2d0 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation c1c364f7-8a10-4067-b9ba-964d218cacd4 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation bd49d785-8948-4577-a935-5beb3f527ee1 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation f1cfd909-86c2-4065-8e3c-e448d9621a5c · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models DINOv2: Learning Robust Visual Features without Supervision
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 8d6ccf26-197f-4b07-a06d-602920319e59 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 670a2808-f012-4078-851e-45ada37ec5d8 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation ffe3a154-5aaf-4027-80b5-847567de1cfd · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 6a1702af-bec8-460d-b80e-41888bc8672b · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models SciFIBench: Benchmarking Large Multimodal Models for Scientific Figure Interpretation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 99457d57-9e9b-488e-9493-2aa71d9cee3b · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models StarVector: Generating Scalable Vector Graphics Code from Images and Text
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 87f3003e-6cf7-43a5-9a1f-829d500ca49d · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models OCR-VQGAN: Taming Text-within-Image Generation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation a2666c8e-65e4-4d7e-bc8c-bf595c51b326 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models FigGen: Text to Scientific Figure Generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 8fcb09b1-f074-4784-a7d7-675820ef6211 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 77b4005e-1dd4-46f5-b1b1-8be724f3edf7 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation a725aeec-e143-4fca-a19f-070a6fa01b0b · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 8aeb5649-b9f0-406c-a534-cf6ff8eb1f1b · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 841567f4-5b03-4a87-a5bf-e8314b9bf0e4 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation a2252e7b-776c-41c1-a6b9-0db6ed906358 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 6d1625b7-d971-470e-97a3-00b0d46a3d21 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation e977d784-24ad-476a-aef0-f7058ec1b1fd · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 2098109e-c795-4eb8-a2f0-e990eeeb7b7c · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 73f37f0d-40e9-4022-b32b-978c601e84b5 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models SciCap+: A Knowledge Augmented Dataset to Study the Challenges of Scientific Figure Captioning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 0b982cf7-55e5-4ecf-82bd-c1a2c83e607e · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation d9f239f0-318a-4107-9de3-8ec44bafc90b · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 83ad4aca-ed07-4ca4-be35-c520170bfa20 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models You are an expert on translating academic writing to visual specification
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 2487533b-bfc4-4ebc-8995-757631700859 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models Autofigure: Generating and refining publication-ready scientific illustrations
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation aaddee79-f288-4f75-8b22-6b81650df147 · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models online" 'onlinestring :=
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation f235db4d-e6cf-4d13-9104-83fca1482a0d · outbound
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models write newline
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 079c2a9e-6acc-4a07-b591-945fb24f61b2 · inbound
SciForma: Structure-Faithful Generation of Scientific Diagrams GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87fa53f8-e43c-4cdb-ae0d-e5e9fbe6fa81 · inbound
SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39b48dcb-1e7a-4f9f-9980-a454bbd1f682 · inbound
SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.