Pith. sign in

Paper Citation Record · LEDGER

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

As of 6 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 25 inbound Pith citation observations for arXiv:2507.21033.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21033 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:05:39.722236Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:41:13.116725Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:44.737452Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1e13080-fddb-45bd-9fee-6173d472ece9 · outbound

This paper cites Qwen2.5-VL Technical Report.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.658388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.658388Z digest=sha256:3b62a1c012bf1173a770cf6156368711a24362e93e2c11ef280d0f733605c20c

Observation c000a00f-84f8-4056-affc-74cc710698c7 · outbound

This paper cites GPT-4o System Card.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset GPT-4o System Card

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.678552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.678552Z digest=sha256:0e3c5fbf71f5451edc7fb1ec3aa71621cba0877b8b40ace2194ffe9c9e6df328

Observation 32fc32f2-8c12-40a3-b483-b9d867bb167a · outbound

This paper cites UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.686427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.686427Z digest=sha256:0b05ab13bcc0fd66db111b112707233948775e029f988d07b187f52f32c9bfa5

Observation 063b186a-9143-4448-89ef-493090b7bb25 · outbound

This paper cites Flow Matching for Generative Modeling.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Flow Matching for Generative Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.690728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.690728Z digest=sha256:bfa394ace37cac49cbeee1e176780d1ee84a9441e6c419f818b0896e73c35ec1

Observation fe926506-216b-4320-b9fb-1b1ec0734e3c · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Step1X-Edit: A Practical Framework for General Image Editing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.699110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.699110Z digest=sha256:d927bafa6939acb827db2e4514b34eba6a928119fdcd681b9d033e0beb64f891

Observation 2240c72c-18a1-4a4a-b969-a0ada78022ec · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.702682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.702682Z digest=sha256:6dbed4766a18a7864083306971fa1f0fce82ef63ef657c2913fc735b08d2af51

Observation 2d3e8b29-c081-4409-a199-3f34a5ce3a5b · outbound

This paper cites SeedEdit 3.0: Fast and High-Quality Generative Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SeedEdit 3.0: Fast and High-Quality Generative Image Editing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.710714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.710714Z digest=sha256:df1d8708ee6816a2a59ff12d3a461bafdf3a315c1318f84d7f6a23c641925704

Observation 74e7d532-c840-4003-b8b1-7aa5177ed517 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.715086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.715086Z digest=sha256:2d0b77daceba7ec40c7e822c83b6b452e66e41163642076f033737b04a81439b

Observation a2ebc351-07dc-4a79-84cd-95e12a8ff3e2 · outbound

This paper cites $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.718864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.718864Z digest=sha256:c75865357e835a1014e65fc2745cf0454963be562d32f9aaaa3236d480c09b93

Observation 2b14b3b4-0542-41cf-880a-1293cd42c1b2 · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.722236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.722236Z digest=sha256:50842d0cdbc98c9a8baec540a2e0c0ece342dfb0e46a179c87fbeedbda43b9fa

Observation 309f22f9-9241-4f0d-b7d2-f8eccc15f32e · outbound

This paper cites ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.665891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.665891Z digest=sha256:0c28b47ff3050e44322e5cdc97ba03cf39009e560482868e93bd66819576057d

Observation d2ad8c8e-b29d-4508-a2da-1a63cf7db32c · outbound

This paper cites SeedEdit: Align Image Re-Generation to Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SeedEdit: Align Image Re-Generation to Image Editing

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.706618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.706618Z digest=sha256:9ef19d7ddc7c481ffb75c6493a7a1d5ac24b6e943a63a7a579ebfe46db8dcba4

Observation 35079285-d3c6-4ba9-b891-e25033bbb6e7 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901,.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.662514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.662514Z digest=sha256:49c21c1fbf537c823617bc7c7b4ce25fd05d42377b930fedfe0e42f0b1c30f89

Observation b9ae3ed5-2307-4878-96dc-04bfe799296a · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.682539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.682539Z digest=sha256:3cc00d1362dd11fd10dd55c07de4b08ccf3133d6bdf3b4b360d8702ef1a51027

Observation 90a86f81-562d-40d1-afcb-7c216e4df9a0 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.669870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.669870Z digest=sha256:7dc79fd8f27d5aae6c0382d644f99ae051c2f89e5d718a8c2dcab9131f18452d

Pith citing papers

Observation 16cd742a-e808-4f03-b47f-51c663bed9a7 · inbound

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning cites this paper.

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:52:07.965603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T05:47:19.552825Z digest=sha256:5022bc50adf48f29eaabdc10e3fc4a1b7d063fb9588fff55c83e1ab856b9cca0

Observation 52408d33-734c-434f-824d-073db230f35b · inbound

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning cites this paper.

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:41:13.116725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:41:13.116725Z digest=sha256:bffb05c1ea9617601bb23b29cd0de7aad7e2d5bcf2480944ac17cf3c4a29161d

Observation ccef5db5-15eb-4134-9cfd-857a8cfa91a8 · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:16.254186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:16.254186Z digest=sha256:8c22c79b8eecdb123887f2b29a4d359b5cc12fa5df8072088ff433a01b245aaa

Observation 48524113-4174-4be8-8849-bf52cb11b5bd · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:08.215305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:08.215305Z digest=sha256:bf0af94555f1ff6418f85ea5530bf4d6bfcd7319cf0b0c6041c5df210986951b

Observation 4c23dc36-ce14-4211-a05d-f86970344e92 · inbound

Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation cites this paper.

Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T15:39:40.585648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:39:40.585648Z digest=sha256:dba9bba1c2d15448a735f65e60bb5b7f9ca8494fc9de0505c71825d9948e651f

Observation ac43947d-a470-4f97-b4ea-fd7ac54aa1fb · inbound

Emu3.5: Native Multimodal Models are World Learners cites this paper.

Emu3.5: Native Multimodal Models are World Learners GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:12:13.583455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T01:12:13.426640Z digest=sha256:c056b594710d11b10c91e8aad162f7cfe49326f5caecef6ec2216823f42d11f9

Observation f30cd611-26b5-44df-9794-0973834a8c51 · inbound

Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models cites this paper.

Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T16:21:37.664980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:21:37.664980Z digest=sha256:f385ba84e161f351f07148a3aa52f1f9910e4b9044b9a22f88932c981aee49c5

Observation 7bdf3801-08f8-4b70-944e-3ce1704759fb · inbound

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation cites this paper.

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:11:11.494014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T19:10:47.425041Z digest=sha256:f4d13a5aa3a8c0dbc33194959550873e9bddbe2a64d695ff979a2e5f0afe436d

Observation f74e0d7c-be25-4471-8ba8-b45fe48e28f6 · inbound

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing cites this paper.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.474554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.474554Z digest=sha256:d2cd887e8641b75826488949ad73ff1100b7ce29293d91d7f318b3d971aedbe9

Observation a13baf45-5f93-43db-a4e7-2387d978ae3b · inbound

Under One Sun: Multi-Object Generative Perception of Materials and Illumination cites this paper.

Under One Sun: Multi-Object Generative Perception of Materials and Illumination GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-13T22:08:47.493022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:08:47.493022Z digest=sha256:9379185e132d5ee0f7aa5b5776d93d26c49137889e77cb027881f55c0ba488a2

Observation aefbd785-8e73-4f20-ab8a-72683fc2803f · inbound

SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing cites this paper.

SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:00:50.312722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:23:50.614589Z digest=sha256:1cdd5cde456b1112f68cac2ed2d59aade8f41e690dd089c3ea69e8619ff8664b

Observation cd5cec37-e215-440a-bf19-4592609a2b1e · inbound

RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details cites this paper.

RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:20:53.908150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:11:43.172296Z digest=sha256:88aeb52fa7254120d1c12ea016db3fc19ec3c76e387007937468c2b713ae94fe

Observation 0bd309f4-2e51-49ea-b1e5-d849e921ec5c · inbound

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation cites this paper.

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:57.551819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:39:59.791758Z digest=sha256:b25722238bc84dca147874a016e034f62110270500458b4791dc72129e02b837

Observation b2185e1f-8287-4024-a27d-7a576bb61fd4 · inbound

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance cites this paper.

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:06:00.083616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T16:15:23.578176Z digest=sha256:69e30c14b55f870370720ae24a16d4fecc62409459e478c28e5b925cc833c2ed

Observation 27b246a9-c1b2-4da7-8f1d-e3d35af300b5 · inbound

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing cites this paper.

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:21:55.104500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T07:21:41.483427Z digest=sha256:54ecf6c3534e5731716827288d5bfc41000afce2baac20b1c61243f68cf8eb0d

Observation ed312bff-e852-417c-8819-982019823f64 · inbound

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation cites this paper.

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:55:43.713579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T17:57:08.606559Z digest=sha256:628c4b612e512ccd368da6f705b585d4a722eef4c6727a00aef19c79822012f4

Observation e3eddba3-60ac-422b-a8cf-c18f06e73ced · inbound

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation cites this paper.

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.686633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T08:15:58.020894Z digest=sha256:e3aeff83b83c265633566575c132b511dc8e52fc87e3630b76ec2bf1429c4566

Observation a5f3e129-d018-4598-a752-d6b87a31c88e · inbound

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality cites this paper.

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 143

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:07.965833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-08T15:04:41.518195Z digest=sha256:eca055cd8e2964504d029e0d5e28170b9b447fd6a4598c4f5795bcd5893e1f4a

Observation b3805d8d-e440-43ad-8e55-c33dfb307e93 · inbound

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis cites this paper.

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.029644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T21:41:10.852265Z digest=sha256:f4c4b5ce3dd7005c3ee31d2ddd52eb8d151fdaa74933dd0bc1cae0db62a104aa

Observation dedc9623-590f-480b-8b92-1112562b4e66 · inbound

TextSculptor: Training and Benchmarking Scene Text Editing cites this paper.

TextSculptor: Training and Benchmarking Scene Text Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:19:39.343635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T05:16:43.756525Z digest=sha256:585399b16e2aa28e409171db902c5d6bcfce5cd2ab3592b6727dbdb8405e3c45

Observation 24c23ea0-e9ad-4a9e-a638-d45473cf0782 · inbound

Bernini: Latent Semantic Planning for Video Diffusion cites this paper.

Bernini: Latent Semantic Planning for Video Diffusion GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:41:10.427081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T06:39:47.124605Z digest=sha256:977d2efafb3f301e589ce3e8b668293b309bfd0e36d9ffcb0d7bad4efd5dccec

Observation df9272dd-e106-4133-8111-c198da1b2927 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:44.738743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T09:08:25.661515Z digest=sha256:413fb10a3211a2f5f47614e63b3d950df3fe38912de25ee9a62b4cb21994dedc

Observation 1c38d0f6-d031-425e-8674-539fbdf5c845 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:19:02.587384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T23:15:09.253879Z digest=sha256:136173d8875566bab2e99de74dad6e27c0f2fea87f4c953eb70bcf9403721500

Observation 8d95ddc0-def7-4859-a94a-c6d646f64f28 · inbound

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry cites this paper.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.773307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.773307Z digest=sha256:2f782baa2a0220f19dae61f971885f9327d56cc2d45fcd3c07f91aed59bc6dcc

Observation e875a0a6-b8b5-442d-8947-11f2f9821327 · inbound

Illuminating Visual Identity in Universal Multimodal Embeddings cites this paper.

Illuminating Visual Identity in Universal Multimodal Embeddings GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:43:17.318655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:43:17.318655Z digest=sha256:f4d61c6780828253e5b26d932be17901d349035db3f53f0d6e0e1aa8bc86a47c