Pith. sign in

Paper Citation Record · LEDGER

An Empirical Study of GPT-4o Image Generation Capabilities

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.05979.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.05979 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:46:02.746041Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:49:51.431619Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 05e6ee68-f58a-4b09-ad8e-0aaced55cb8c · inbound

TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation cites this paper.

TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:09:10.636658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:09:10.636658Z digest=sha256:51e1c2a2983fb665c0cac6fb118b924c6f793495f8d5c0ac70fe6c52cedfe6a7

Observation 144adc4e-f03e-4738-885a-6ecade1076d2 · inbound

Preliminary Explorations with GPT-4o(mni) Native Image Generation cites this paper.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.746041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.746041Z digest=sha256:5382ee44a472a463bdd3bd5fccda0aacc404adc8901441b2fab7d508739917d7

Observation 45a68c89-357f-459b-8636-d0fa70b96755 · inbound

Emerging Properties in Unified Multimodal Pretraining cites this paper.

Emerging Properties in Unified Multimodal Pretraining An Empirical Study of GPT-4o Image Generation Capabilities

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:23:41.909986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T16:23:41.854132Z digest=sha256:03e6390ae9018203feb15f2bfed61199d8cde5bf2e7c2eef88866d271f527af0

Observation dd15863d-8cf8-4044-a176-65211f606d62 · inbound

MV-CoLight: Efficient Object Compositing with Consistent Lighting and Shadow Generation cites this paper.

MV-CoLight: Efficient Object Compositing with Consistent Lighting and Shadow Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:34:19.368948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:34:19.368948Z digest=sha256:e421a30e143da060b4f10c88000f3d8f1ea5ef9c856059cac97c7e880295e6a9

Observation 957fd8ea-0870-43da-8c62-b3105c881098 · inbound

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation cites this paper.

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.456631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:28:03.456631Z digest=sha256:0ad26d9646913e8969221c712aaa157e3acf4a668a3fd96bcca7579a459e6c17

Observation eb24732f-0bf0-418c-b2d8-fe4667ea7b09 · inbound

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks cites this paper.

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks An Empirical Study of GPT-4o Image Generation Capabilities

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:57:08.065260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T05:55:09.188048Z digest=sha256:be8a19c1733f68fecec6855fc950608c162b134295b47ac642a61775fc0f6584

Observation 408d4dc0-97c8-4740-929f-1fc45933843d · inbound

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models cites this paper.

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models An Empirical Study of GPT-4o Image Generation Capabilities

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:11:52.762645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T22:10:16.645510Z digest=sha256:34203b56dd43636ebd702acf40ada758b4f2ac30fc194c7e1a85edfd35cc5157

Observation 632f58af-6a82-4904-98ac-23309dbaabd0 · inbound

HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images cites this paper.

HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images An Empirical Study of GPT-4o Image Generation Capabilities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:26:22.000134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T17:26:02.023978Z digest=sha256:2add984890abcae6645ddfd4379b67de2eabf3a0bb22541dec607c26f5b859ab

Observation 261287a6-2821-4e2a-92aa-90e776afcca7 · inbound

Structural MRI Synthesis for Alzheimer's Disease via Conditional Diffusion on Anatomical Masks cites this paper.

Structural MRI Synthesis for Alzheimer's Disease via Conditional Diffusion on Anatomical Masks An Empirical Study of GPT-4o Image Generation Capabilities

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:29:02.745471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T22:12:46.169295Z digest=sha256:6d95215cb8543af29deebdff28381b8ae12a517de39ee0736120fdc724d1ad15

Observation b9179e33-676f-4962-ac36-4aa3b5babbdd · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:49:51.433341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T04:55:42.018348Z digest=sha256:98ad7b395ce4fede8c3237d6b70a9386da18800b48673b3f88a85d941bec3001

Observation b3abc640-9e19-4156-8321-7ce8459d9e97 · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T11:44:54.717393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:44:54.717393Z digest=sha256:cd7cd971c15672bd6b9bbdde92083232964f867d9f7840f4dc82126a495ba8bc

Observation 089415a7-d51b-4f27-8f3c-90aa1c8a50ef · inbound

AI-generated Images Challenge Visual Trust in High-risk Scenarios cites this paper.

AI-generated Images Challenge Visual Trust in High-risk Scenarios An Empirical Study of GPT-4o Image Generation Capabilities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T08:52:47.347766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:52:47.347766Z digest=sha256:5ef02d1cc1a79ea4a4d7008763d3dc2f14428d619987129ff797d695ba09ceca