Pith. sign in

Paper Citation Record · LEDGER

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation

As of 7 August 2026, this Paper Citation Record lists 88 of 88 outbound references and 0 inbound Pith citation observations for arXiv:2507.07317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07317 v2

Coverage vector

measured 88 of 88 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:48:46.673046Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

88 of 88 outbound references displayed

  • verified exact2
  • verified fuzzy36
  • unresolved49
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 73803670-10a8-400a-90f2-086b1f14f79f · outbound

This paper cites GPT-4 Technical Report.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:36.859682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:36.859682Z digest=sha256:be6e0b1fa0bab5f18699136828553fd7405e7d010d739e8031c730024431299e

Observation b3b8855b-32cd-4c55-9187-c9a8286ff7c7 · outbound

This paper cites Cos stable diffusion xl 1.0 and cos stable dif- fusion xl 1.0 edit, 2024.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Cos stable diffusion xl 1.0 and cos stable dif- fusion xl 1.0 edit, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:36.939717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:36.939717Z digest=sha256:1acdd4438c1ee02d71ab0bafdfc6691ec904537b3ffc1eec2f616e31cc4f9a58

Observation 0443baac-97af-4ca7-b4a8-32a38837091b · outbound

This paper cites Claude 3.5 sonnet model card addendum.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Claude 3.5 sonnet model card addendum

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.088224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.088224Z digest=sha256:e547cd1c5f2fa636df656974cbfc67ad8acb07c65a6f7588e8a3ca93e6d36f4a

Observation 7760d3d8-8267-4825-8be7-572708b72096 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.185937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.185937Z digest=sha256:733ea1fc0415d13c462c7895fa9c226a61f6266dd8d05e97508c642d1279b16c

Observation 1fd3bd14-795b-438d-9757-48d709760987 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.334872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.334872Z digest=sha256:e7f3c7835f7cafb5a939c0c473ec1d269daf4f50ce4e0871d10915b3016cbfe3

Observation 1a00ba9d-c536-4cc8-8ecc-4ce29ebee588 · outbound

This paper cites Text2live: Text-driven layered image and video editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Text2live: Text-driven layered image and video editing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.444474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.444474Z digest=sha256:5b1a50a5fbe323b17a28b0be79ae1c5e53d6688ff9e316ed94f77f49cb80ec20

Observation c1d90854-d996-4922-ac60-437f3c9f41bb · outbound

This paper cites Introducing our multimodal models, 2023.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Introducing our multimodal models, 2023

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.535105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.535105Z digest=sha256:94a9dee2ed6c03a257240e16e09cec8b04df5adb60a098dbfa21d5d225327516

Observation 6062fa7b-067f-4ecf-a83c-1ba8d949ddd2 · outbound

This paper cites Is clip the main roadblock for fine-grained open-world perception? In 2024 International Conference on Content-Based Multimedia Indexing (CBMI), pages 1–8.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Is clip the main roadblock for fine-grained open-world perception? In 2024 International Conference on Content-Based Multimedia Indexing (CBMI), pages 1–8

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.672421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.672421Z digest=sha256:dbedf08baf2c6a9eef0beb24fe5319cbbd083133665d35a244057f72d01ea28a

Observation 55afb521-809f-451e-8042-e60ee9701ddc · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation In- structpix2pix: Learning to follow image editing instructions

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.770834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.770834Z digest=sha256:90116b34c37001c7533fc5365d3b7f6783cb5ee0d90d74b1d87099362e9c8fe6

Observation 030b1cbf-c569-4ae2-b528-c72b49bda671 · outbound

This paper cites Lan- guage models are few-shot learners.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Lan- guage models are few-shot learners

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.883358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.883358Z digest=sha256:70bd910864c4a4df7a610bd9d59f5ed164ab19d470dd1dfa95c32adfc325460f

Observation 31cdb809-1e0f-4bba-8d2f-ac017d213150 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Emerg- ing properties in self-supervised vision transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.978585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.978585Z digest=sha256:6ab725845e73c9dc421ca4b6bf810bed91ea256daaa33fadd9219223009a6939

Observation 1c8db78e-3087-448a-b247-82dc018016e8 · outbound

This paper cites MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.097897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.097897Z digest=sha256:71582fe5ca0a7d725c8555d6898a3cadd0641a4f201c122c4dc48a12d2161267

Observation 318e331a-086b-45b9-a626-cd80b3d6c8e9 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.227993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.227993Z digest=sha256:bbcf4c6394e857fbf5de4f56e158ea6710c4916dd47bc69dfdd1aacb1f645ec0

Observation 841ff0b5-0a4b-4a14-81b7-9e66f2ce5a7b · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.322722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.322722Z digest=sha256:ca07a07948d8f5e8c03f6dca8e2f94642448027be7bdf60f91564081a93c114f

Observation 00cdd155-fda2-4c85-922b-7ba276b3128e · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.482789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.482789Z digest=sha256:8b22ed6e3279a66f3e87c6a2070fa3e7757667572a520394303eccd453ca6706

Observation 5e89a477-5d13-41d2-a3c4-40be01ee4a4d · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.623306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.623306Z digest=sha256:3616477e642cfce7dd56069ef75d1f892cc7b6cdb2ff972e7f40d68b137421b2

Observation fcd48602-d109-4347-a20c-69e931a1a218 · outbound

This paper cites The epic-kitchens dataset: Collection, chal- lenges and baselines.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation The epic-kitchens dataset: Collection, chal- lenges and baselines

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:57.287543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:38.694374Z digest=sha256:2ece2a346e5d702f26c42fd51a1f3d8c5c04de66d84a682d910df972ba4c23ef

Observation 3df92554-ed0c-47eb-9ff2-99fdf29e2844 · outbound

This paper cites Diffusion models beat GANs on image synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Diffusion models beat GANs on image synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:57.107238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:38.805037Z digest=sha256:d7694fdf7b62f9b5f195ebd2b6fabd394cee6b8d2798a1935256c801e357e454

Observation aa04f5c5-a8c4-414d-b9ae-cc4f7c45248b · outbound

This paper cites Dreamlike photoreal 2.5, 2023.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Dreamlike photoreal 2.5, 2023

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:56.865429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:38.903563Z digest=sha256:c4cd0c52c665a719f727d4f56bb078faf9c3da8bc0574cf66a5e222470c37ddc

Observation 315354ec-94f7-4a86-b2a4-6c612b76d4d3 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.987435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.987435Z digest=sha256:c363337ce9cc60fd803c0840785bb7909305d0c1f9d1b0ca1c6ead2d3f573c85

Observation bf0f9e38-56d8-4007-b7e4-53875d7f8dfd · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.067547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.067547Z digest=sha256:6e4ef88d44cb032c134acb81729528bfb5792d255f1537b4ae6ef613c14c1a52

Observation 44dd0e9c-a8d4-494b-8314-94d6cda05bdd · outbound

This paper cites SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.186905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.186905Z digest=sha256:2f3f3e0ba2c6b2112e29f5f332da23b592b234a1a64de1edf41406e0eab9c3eb

Observation 812faeec-4c2f-4a18-a3a6-57115f07b1ec · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.256568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.256568Z digest=sha256:fd7185a9575c5576a79f6869ddf5bb835606c28d96a345ca89605669284c0846

Observation 939bf505-b7a9-4d98-b509-d71f4ba618f1 · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation The” something something” video database for learning and evaluating visual common sense

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:56.712513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:39.416699Z digest=sha256:019d970fff22bfd2e984b0dd436870a85bb377c223aaeeb59e6487768b3175b1

Observation 7e212353-a222-4a26-a374-efc1aa6b8390 · outbound

This paper cites an unresolved cited work.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:48:56.489778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:39.526651Z digest=sha256:50778fc1b328bc83570100d1d765ebe1af82c5f3915e7318ba656aada80e1aa7

Observation 45219a64-2c0e-4161-becb-94c209e23cef · outbound

This paper cites Multi-Reward as Condition for Instruction-based Image Editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Multi-Reward as Condition for Instruction-based Image Editing

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:48:47.999415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:39.649531Z digest=sha256:d4d456184ab676e2f2cb45dd0aef3072055bb240a19069da20c346e325619cfe

Observation 2c2f0c50-0fd6-480d-ad51-af2ed04f2177 · outbound

This paper cites VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.761303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.761303Z digest=sha256:f40bd554cf5315c342b08eb36163a85a0dedad08b063a0c36283365052bedbba

Observation fc613fd7-3279-4765-9322-bd541e9a29f7 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.908378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.908378Z digest=sha256:01a04c960ad8eec8de2260d9379589fc24f072742606c45b44e092cb9b163dd7

Observation cc1f3b99-1f29-40b5-a64d-2094ddebe56a · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.054243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.054243Z digest=sha256:25396419204408f092f704eb87fbd10550eef73c1e45a5b9bd30e618075afc7a

Observation e422a841-340b-42c6-9422-4d3695b107ad · outbound

This paper cites Lora: Low-rank adaptation of large language models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Lora: Low-rank adaptation of large language models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.154473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.154473Z digest=sha256:93f74d9ee247a287d85ad353f98fbd74c709a8315d979413bc0e133b7b8bf52d

Observation 75f18027-de1c-4400-b27d-273d352bf943 · outbound

This paper cites Action genome: Actions as compositions of spatio- temporal scene graphs.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Action genome: Actions as compositions of spatio- temporal scene graphs

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:56.262111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:40.217599Z digest=sha256:459258e154e1ee1cd8cb8e83b71a4b456b7bdab1f12d91767fe0d07efc622e61

Observation 2560f2a4-585c-465e-96cb-ecc9f473ae27 · outbound

This paper cites GenAI Arena: An Open Evaluation Platform for Generative Models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation GenAI Arena: An Open Evaluation Platform for Generative Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.299542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.299542Z digest=sha256:bfa3909ae48ee4b1f3bb37c07c62162931d471f2b06e1b5ead9d3558273cb1f5

Observation e91b6027-b560-4ee7-960b-ed6ed8549437 · outbound

This paper cites Clevr: A diagnostic dataset for compositional language and elementary visual reasoning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Clevr: A diagnostic dataset for compositional language and elementary visual reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.981983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:40.392429Z digest=sha256:850e933f3acf34f74caad5758e7e8d4e2eb3ce9835eac35d12f24576d13a4fb7

Observation c37c56fe-5ab7-4df3-a94d-6b5ca9c7359c · outbound

This paper cites What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.774617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:40.508645Z digest=sha256:1364a50d715cce31e1128441013031dc1a3395afdacd0bf51fd7af4a4816c941

Observation 85d71823-4eb0-4e36-af0b-940643c377bf · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image genera- tion.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Pick-a-pic: An open dataset of user preferences for text-to-image genera- tion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.524988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:40.651286Z digest=sha256:19b96fac1701e32cc1aaf77a62c514d31969064f818d3170646b0fc2abed7d29

Observation 13b85a26-e7f0-47f4-996b-90bc6c1e1ba5 · outbound

This paper cites Learning Action and Reasoning-Centric Image Editing from Videos and Simulations.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Learning Action and Reasoning-Centric Image Editing from Videos and Simulations

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.296234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:40.763102Z digest=sha256:874ed07c7264e49c30f3d0ddfeb2770f62634a7e878f8316094692e9651145c8

Observation a3bd7ea0-470b-434b-9038-2b132119e827 · outbound

This paper cites VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.925409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.925409Z digest=sha256:0ff8879c9859bda5f4f3bf195acd121f6f863e3fdfd894f5a72060ce687cf0c8

Observation e38527a1-a279-4413-bc17-0dbf56168060 · outbound

This paper cites Imagenhub: Standardizing the evaluation of conditional image generation models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Imagenhub: Standardizing the evaluation of conditional image generation models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.134883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:41.094239Z digest=sha256:9c71ffe58381181d17fa267dfc71a68f58a12c893551f4ebef97a04aa6e71f75

Observation 273f83b8-2f55-4422-b866-b56ff5e11ddb · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation LISA: Reasoning Segmentation via Large Language Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:41.170936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:41.170936Z digest=sha256:5c9300b7a1724a1259e1e114e56444555741b5e11d03b1483cad85a1a64148fe

Observation 5d425981-fd93-4cfc-9610-296a2aac90b4 · outbound

This paper cites Introduc- ing idefics2: A powerful 8b vision-language model for the community.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Introduc- ing idefics2: A powerful 8b vision-language model for the community

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.842646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:41.328554Z digest=sha256:9fad47bcc536d23c6d715ee7b2e1d629cde2cc3946ab223c66e764406ba9dd28

Observation c2fa4195-ca75-476c-bac3-732febbd8811 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation LLaVA-OneVision: Easy Visual Task Transfer

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:41.472557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:41.472557Z digest=sha256:395d72a9f823cb4b4528f830b69ab4fb2330c61e4fdee1608faf84738de67c8c

Observation 70baf567-1bc0-4c63-bd42-dea7ad167215 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:41.543773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:41.543773Z digest=sha256:af4df5f7e1b95e8da849f49a194453254abcc3389d2fe9f3ef1ad3fd5d89274f

Observation ffb0ec51-6caf-4aa3-89b2-77cf5dd41369 · outbound

This paper cites ZONE: Zero-shot instruction-guided local editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation ZONE: Zero-shot instruction-guided local editing

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.663638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:41.666684Z digest=sha256:345e97ca3793b94d71a6d0965c0785c8f9fc0c8c252de0b19ae95a4f6e0cb092

Observation ed81f119-4d62-4a1e-b5d2-643bc5d08a86 · outbound

This paper cites Rich hu- man feedback for text-to-image generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Rich hu- man feedback for text-to-image generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.408687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:41.778736Z digest=sha256:30500f271aec189e83813cc080bae396ac029ef063d1a5e65f0e3fe829238ec8

Observation 35157127-5dbe-4d91-a02a-ebbc8c33f484 · outbound

This paper cites Evaluating text-to-visual generation with image-to-text gen- eration.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Evaluating text-to-visual generation with image-to-text gen- eration

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.181411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:41.932198Z digest=sha256:d12dc4f3d028d119736099ea77a9eb119a20c1d551bde9520bd4222f43d8a379

Observation fe292af6-2fba-4015-b442-285d7662e961 · outbound

This paper cites Llava-next: Im- proved reasoning, ocr, and world knowledge, January 2024.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Llava-next: Im- proved reasoning, ocr, and world knowledge, January 2024

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.009132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:42.044640Z digest=sha256:0ce0928e91b6a6af521f0cd58371033981221db62a72888253e4d0cdcbaf9008

Observation de4403b8-8010-4dce-935d-617b596dc5a4 · outbound

This paper cites Visual instruction tuning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Visual instruction tuning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.803553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:42.110125Z digest=sha256:3f4976bae4d66b0b37472221473d274c3591c279e9e7223a8c5c5b3a4afe9a60

Observation efd8b911-2f75-4d3a-8a9e-9659053125a4 · outbound

This paper cites Visual instruction tuning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Visual instruction tuning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.239498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.239498Z digest=sha256:dbd0b6378e117dc8c9a6d58db767c9173aceb2d9230fa5234cfea68fa33976c3

Observation dacaa74e-4b55-4d74-a232-6d874d6d3146 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.388687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.388687Z digest=sha256:aa6a38b621c64147ea187832f9ea5affad83758b9ad62b542deb37e0fa61b220

Observation 07d900e7-ac67-43a5-9e58-24830fffbc03 · outbound

This paper cites I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.511233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.511233Z digest=sha256:059972c491adb8e4be816a7c68653510b2596e490dbd82cc393c3c397d657bf0

Observation c4b4b26c-8708-43fa-a6a8-971199d3f9ae · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.647440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.647440Z digest=sha256:eeaba2605e3863281ba1fcac2b3a874b01c408ae8627651dfc2786d977e5aa18

Observation 3c18578f-3c62-4eee-9cae-6545d5f94e31 · outbound

This paper cites Watch Your Steps: Local image and scene editing by text instructions.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Watch Your Steps: Local image and scene editing by text instructions

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.622904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:42.778364Z digest=sha256:617bca4962aa6544fd972121e72b12bf45d5edee82dad1a54434f70ccc4d9bd2

Observation 9a4772f2-fbac-487d-803a-c6815fa42fc7 · outbound

This paper cites Null-text Inversion for editing real images using guided diffusion models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Null-text Inversion for editing real images using guided diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.365096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:42.853041Z digest=sha256:bc9ea78fd2ec732e0e3c99f8aacbf67b9793f2e501f74fe85cc2ae0f8ea8984e

Observation 601f0582-615a-48d7-b0ca-9a285694f71e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation DINOv2: Learning Robust Visual Features without Supervision

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.928679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.928679Z digest=sha256:18d19d3ba182fee96e659451f5910b7ce162dc56cae7bc1f7d4e1c1693f856f8

Observation 229a0c2d-1188-4b68-92e6-46532039745b · outbound

This paper cites Zero-shot image-to-image translation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Zero-shot image-to-image translation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.165587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:43.018862Z digest=sha256:0c9f55976398f778142e9c08a16a2b8eeb0f01a39553e9f27e703fbec9f8553f

Observation b3080117-a3f8-4b4a-8d7d-6d700189f94a · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:43.156247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:43.156247Z digest=sha256:907c0ce46c40d7456aa27a17226b14ec61a79b87d6d395724824be6cc1bf151b

Observation 09854faa-21f6-4fb5-8c77-519bf6872ec0 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Learning transferable visual models from natural language supervi- sion

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.881947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:43.294518Z digest=sha256:f5abf911693dcca01c09740e82de85defb2d5613f778897dd8b25ef418f75900

Observation 09d01c24-5c1d-4661-992e-ce783d53ac0c · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:43.438852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:43.438852Z digest=sha256:1873178b3ece1f010108ed57180a8f913250a48143c732be8a754c681807f613

Observation 4e3a6838-2556-48b7-badd-5e638a794f9b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation High-resolution image synthesis with latent diffusion models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.623565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:43.638757Z digest=sha256:0a56213a1e8556f1203f3f2c755e3299f5e412d15811df10b51f43786a1aa902

Observation 35f71e6b-2bfe-459a-b7b6-b6c4089570f1 · outbound

This paper cites Emu edit: Precise image editing via recognition and gen- eration tasks.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Emu edit: Precise image editing via recognition and gen- eration tasks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.368993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:43.698204Z digest=sha256:6e33a24fa9610815cd5265dbc85f6376923241375aadc0f44f753f889f5460c2

Observation b32463c9-32f5-4501-ae26-d93768503d1b · outbound

This paper cites Aria: Advancing multimodal ai.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Aria: Advancing multimodal ai

Reference 61

Resolution
verified exact
raw_fallback, observed 2026-08-06T18:48:47.318442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:43.804495Z digest=sha256:ae63d5755b39d19a59532d01c2201104710fb6650ee9e8268e3a7293660be877

Observation 2152a224-2206-4fba-8722-4e3ed131b37d · outbound

This paper cites Denoising Diffusion Implicit Models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Denoising Diffusion Implicit Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:43.954569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:43.954569Z digest=sha256:7f20b220c7df42f3197914c27104d6d14c4fb9e666183de047ece1b1962ecb40

Observation 08b2c90e-c4ba-43fd-98e1-1110031dbea4 · outbound

This paper cites IE-Bench: Advancing the Measurement of Text-Driven Image Editing for Human Perception Alignment.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation IE-Bench: Advancing the Measurement of Text-Driven Image Editing for Human Perception Alignment

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.077282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.077282Z digest=sha256:f8ed38db867637b0cb28b0260d37a8b77ee4fbe5ca8e97841baaeff024c2a24a

Observation 445f8659-fb5b-4566-8c39-0f50ed8a2670 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.180428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.180428Z digest=sha256:33e045237ee9f7b624cc0de4cafacee9d7b16831ddfa22dd70350eaa0c7fe3f9

Observation 05dfb526-9c39-40de-8b2e-3edbf472e13b · outbound

This paper cites Plug-and-Play diffusion features for text-driven image-to-image translation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Plug-and-Play diffusion features for text-driven image-to-image translation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.143028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:44.295558Z digest=sha256:b5f6b9e3738e92427649d40544c9ceb6f292e6a9b47b88dde5c8e85c16a0de41

Observation 55527c56-02bd-4408-8d9d-e4e5ecbb3a97 · outbound

This paper cites EDICT: Ex- act diffusion inversion via coupled transformations.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation EDICT: Ex- act diffusion inversion via coupled transformations

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.934383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:44.405595Z digest=sha256:652c8f847b91fa042c60869631d3cd77365d5ca888863fe1b439af8ea8b11022

Observation ee6d5465-21ad-4c2f-b1f7-cbfec2c7a076 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.469725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.469725Z digest=sha256:e9a14bbccc8802b4ebfd065367c0cc9cc03182f5b3f889d5ca679b0fe51089b1

Observation a542c4b8-016b-4404-9502-a97b63c1b5ea · outbound

This paper cites Cogvlm: Visual expert for pretrained language models, 2023.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Cogvlm: Visual expert for pretrained language models, 2023

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.522758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.522758Z digest=sha256:f3163d6bc4451392979c1a3996b1bc3619a7be8a531adfd237505203b5d5dc8e

Observation 6c547196-5242-45f6-ad04-474f1adf368d · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Image quality assessment: from error visibility to structural similarity

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.661975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:44.613567Z digest=sha256:442f0cd9f5b82d80fab13f1ba502314b9c38537539d1c8fc7ab38099e9d7e377

Observation ec506cb9-3f42-4de6-b3f2-5ae60bd1a379 · outbound

This paper cites OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.706303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.706303Z digest=sha256:3db1c7efc5f81ebc3ec37b41d8521ea31baf5422ccd70ae56625575005eef60b

Observation 2df7148e-531d-4d25-9337-322a52b82641 · outbound

This paper cites A latent space of stochastic diffusion models for zero-shot image editing and guidance.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation A latent space of stochastic diffusion models for zero-shot image editing and guidance

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.457038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:44.848440Z digest=sha256:58f8b0b9ce8a4d137ef81267b5a5f37f547f6cb2a768a086f887006712b6ee36

Observation a72c81cc-f3fb-433b-a8bd-549b68e2626c · outbound

This paper cites Uncovering the disentanglement capability in text- to-image diffusion models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Uncovering the disentanglement capability in text- to-image diffusion models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.269087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:44.968134Z digest=sha256:afa3bb85ab610dd1fef1c4035c1f946ed95827da60c478b1c9242b026ee94f2e

Observation 98f7c310-02ba-486c-9cc2-3360905cbdde · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.124627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.124627Z digest=sha256:12022f926fa1ae4985a6f1cdb700d6c362400a07f463b613639932d3b2cca0bc

Observation 5767bcff-6c33-422a-91c1-7e7ad09062b5 · outbound

This paper cites Multimodal large language models make text-to- image generative models align better.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Multimodal large language models make text-to- image generative models align better

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.082936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:45.244020Z digest=sha256:ddc247064b58b5d7fdf3ab7ad683a87b432f06a497aed945639ab495fea939a4

Observation f2883f65-6823-4ff0-ae40-905a7a59639b · outbound

This paper cites Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.322166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.322166Z digest=sha256:eeb0accc32890832ec6d623d4b4633680da2e446c0574d8a33e66f433db3561d

Observation a236bf43-d83d-42a8-8ba2-a8ba8ae1c26f · outbound

This paper cites Human preference score: Better aligning text- to-image models with human preference.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Human preference score: Better aligning text- to-image models with human preference

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.816393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:45.426857Z digest=sha256:0e215554a49afb7ba5a8f6223fe4620f31836ea4ac767417da992b7683ce41ae

Observation 7bc553e3-d597-4a0d-96a5-3c733c4c4341 · outbound

This paper cites Imagere- ward: Learning and evaluating human preferences for text- to-image generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Imagere- ward: Learning and evaluating human preferences for text- to-image generation

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.561184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:45.562550Z digest=sha256:877e62c143dfb2e86f215ce96fc3b7a1cde2c6f007ab0cd18f671aadbfa84a6e

Observation 3e5e715d-3bf9-467b-be0d-808afe6ccbc4 · outbound

This paper cites Inversion-free image editing with natural language.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Inversion-free image editing with natural language

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.697618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.697618Z digest=sha256:2a098655384640ea9a01bac434b98804102b8834d9a5c5111ea5fcd4f8961293

Observation 9666feda-0088-42e3-b9bf-84424452a916 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.861490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.861490Z digest=sha256:8847c03be9eec011bfd3512dea74e50e934f8adecdad0be2ac949d2a2ee320bd

Observation a4d73f97-fff5-47e6-9079-cf3839a4e9af · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.288946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:45.909446Z digest=sha256:ad960fbb83f27f960f1cbb8e8f735c58079b4620e292a65d633cf40c5d888d78

Observation 48d636dd-8c85-446d-acd7-e7f440c730b8 · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.958345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.958345Z digest=sha256:daa2fa26847373b678c8a9fd6c30becc9eca77edc6b2ad72ef45b0bac666ee86

Observation a979365d-1191-46f0-8491-d82a85753118 · outbound

This paper cites Long-clip: Unlocking the long-text capability of clip.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Long-clip: Unlocking the long-text capability of clip

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.019536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:46.055557Z digest=sha256:aeadeb578176f90dd666ae4f071a7679bcc0efca8fa767124acd0e829d7f544c

Observation 75fd2958-d867-4973-a9a6-ac2e722471aa · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.728766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:46.208380Z digest=sha256:b3002dbb028060439a11b502c89e6a3e430c4659915bbd26be12c655697b1745

Observation 03a1bcbd-f75b-4f0d-a28f-4bbefc6f4f15 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation The unreasonable effectiveness of deep features as a perceptual metric

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.519550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:46.313062Z digest=sha256:ae8f793f68e5c22e7755fc52c09ad521611d2d3cf896f86d16d79b2f8a980616

Observation e991b1d4-3cab-4cec-a433-a5ba5a8b1f55 · outbound

This paper cites Learning multi- dimensional human preference for text-to-image generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Learning multi- dimensional human preference for text-to-image generation

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.280332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:46.376211Z digest=sha256:99fe52fb4f4fdacd6342884c0f34135ba1fdbb7745a0bbd33d1b65e572d13ebf

Observation aaa52f07-c2a3-4c94-b459-c7435d1050ce · outbound

This paper cites Hive: Harnessing human feedback for instructional visual editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Hive: Harnessing human feedback for instructional visual editing

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.021181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:46.496584Z digest=sha256:7ce809ee70efb42bfce7e829be122d9c6a985d947ed70b2f4ef7ce3414535996

Observation 1fef0756-b515-4779-be69-9de6e264fb0f · outbound

This paper cites UltraEdit: Instruction-based Fine-Grained Image Editing at Scale.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation UltraEdit: Instruction-based Fine-Grained Image Editing at Scale

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:46.618116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:46.618116Z digest=sha256:44b0a6bf896efa6d172a106681c7b516a5cc0969fce1dd432f10efc6e76ed4c6

Observation 1f9888cd-41c0-49be-a770-d0ce2f9a8846 · outbound

This paper cites Can you rate how successful the edit instruction [IN- STRUCTION] has been executed from the first image to the second image with a score from 0 to 10?.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Can you rate how successful the edit instruction [IN- STRUCTION] has been executed from the first image to the second image with a score from 0 to 10?

Reference 88

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T18:48:48.700573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:48:46.673046Z digest=sha256:714ab22e84c49935aee6feb4f424fd3983a8532bb7b78b8b8e28c802667acae2

Pith citing papers

No inbound Pith citation observations are available.