Pith. sign in

Paper Citation Record · LEDGER

InsightEdit: Towards Better Instruction Following for Image Editing

As of 13 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2411.17323.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17323 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:19:22.592578Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-16T16:07:53.054355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T16:07:53.165809Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 760af4c4-abb5-40fd-949b-68587d0a0794 · outbound

This paper cites Blended latent diffusion.

InsightEdit: Towards Better Instruction Following for Image Editing Blended latent diffusion

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.336623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.358642Z digest=sha256:8793987b09580f9293821967ff0cd3e518798d624e955da39fa0e942cfee918c

Observation 5667b81a-71fb-4832-9207-d985f4552312 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

InsightEdit: Towards Better Instruction Following for Image Editing In- structpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.317771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.364636Z digest=sha256:8893a0f9e9455143bb05d1ea3495966637cb9f8b88dd555b867480fcddc59240

Observation 1200f067-85c1-4936-b5c0-b6728a3f9478 · outbound

This paper cites Coco- stuff: Thing and stuff classes in context.

InsightEdit: Towards Better Instruction Following for Image Editing Coco- stuff: Thing and stuff classes in context

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.370164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.370164Z digest=sha256:195208b8347d05e12a6414e8da22f35492dcffd9b7a5be630b18c736dd34c237

Observation 03de9598-47b0-4693-b29c-66309b4f3ff9 · outbound

This paper cites Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing.

InsightEdit: Towards Better Instruction Following for Image Editing Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.376187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.376187Z digest=sha256:9878f119a608979fa7d8ea1ac4bdd4d9b3ec3ef6d27fd1ccda82a14ac8011244

Observation 2c38f988-f88e-420d-910a-9b79d30b4127 · outbound

This paper cites End-to- end object detection with transformers.

InsightEdit: Towards Better Instruction Following for Image Editing End-to- end object detection with transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.383137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.383137Z digest=sha256:f8578d14c33028ef824894e7593ccd4db49d84db7cefd985ddec4d2340cf33c6

Observation 19b28b7c-5657-45e1-a1cd-e182da48246c · outbound

This paper cites Learning to Follow Object-Centric Image Editing Instructions Faithfully.

InsightEdit: Towards Better Instruction Following for Image Editing Learning to Follow Object-Centric Image Editing Instructions Faithfully

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.391172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.391172Z digest=sha256:b561965daa2334bed79e7cdae6cde6f43f528174dc9ae33ac0110a5413b50428

Observation 882868f8-e7e3-4a11-9b74-55ee7ab7ecd5 · outbound

This paper cites Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts.

InsightEdit: Towards Better Instruction Following for Image Editing Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.397476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.397476Z digest=sha256:d8812095a38afb8b6fb8993b33d19eec98e65f227461ade12046f6da1ca2c96f

Observation b6fd0b5c-dc44-48d3-9e4e-2c0151863f78 · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

InsightEdit: Towards Better Instruction Following for Image Editing Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.402699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.402699Z digest=sha256:b288376265c359a5258c069e2a8843a53a75b627597b316b46dfdff4420f4431

Observation bb97db91-9e4c-4659-a3ff-27323f5ddce0 · outbound

This paper cites Instructdiffusion: A generalist modeling inter- face for vision tasks.

InsightEdit: Towards Better Instruction Following for Image Editing Instructdiffusion: A generalist modeling inter- face for vision tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.253477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.411977Z digest=sha256:527dea63b7df1e630bf33e2f3eeffa8199b1d77ea86d1cd84245a7fa483dea8b

Observation d7007744-1940-4c6d-ace5-9b3b6649b8a3 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

InsightEdit: Towards Better Instruction Following for Image Editing Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.417571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.417571Z digest=sha256:3549bbe0d1ec28cee77833f9b65029e1e8cca4ebdce1415ed4f9d4d3e856c662

Observation a66cee22-bf40-44ed-b424-ba9352f5317f · outbound

This paper cites Image quality metrics: Psnr vs.

InsightEdit: Towards Better Instruction Following for Image Editing Image quality metrics: Psnr vs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.422518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.422518Z digest=sha256:7658ccf4a353d6b8d4b5785c8e9b57eb04f9aedf1cbe803d7335747df72ee82a

Observation de62810b-1005-42e4-a081-900ba0658e5e · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

InsightEdit: Towards Better Instruction Following for Image Editing LoRA: Low-Rank Adaptation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.427569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.427569Z digest=sha256:bc6ec154a8019362a790f38033a3855abaf8de6f05c7d8500da2c8bc23b5e02b

Observation 6f051b66-ec4c-4448-bd43-30e5e500137a · outbound

This paper cites Diffusion Model-Based Image Editing: A Survey.

InsightEdit: Towards Better Instruction Following for Image Editing Diffusion Model-Based Image Editing: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.433319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.433319Z digest=sha256:ddc632a28afc601dcf87ef86764679b0cd61b791c0336feeda60917ab7281946

Observation de0a4650-16b1-486c-a064-2d003056840a · outbound

This paper cites Smartedit: Exploring complex instruction-based image editing with multimodal large lan- guage models.

InsightEdit: Towards Better Instruction Following for Image Editing Smartedit: Exploring complex instruction-based image editing with multimodal large lan- guage models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.219965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.438521Z digest=sha256:b8c7b31e2f3bed6d816529d457fb9397f4fca78c2c2e71f542258bc7ac81fbf3

Observation fe37e4f9-7969-402c-b9a2-e67348baf9a0 · outbound

This paper cites HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing.

InsightEdit: Towards Better Instruction Following for Image Editing HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.444615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.444615Z digest=sha256:688afc7b7de08ac09ed6bb42c467e9f63749f2fceaf3b96c67030745e47cc445

Observation 61fb3233-c4ad-4751-9c00-c3c0b97c250d · outbound

This paper cites BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion.

InsightEdit: Towards Better Instruction Following for Image Editing BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.448901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.448901Z digest=sha256:8fdd0bd3e33afb0df04c0b8b5cdcf30ca5af0a8cbc5322fac937cfa4c469fd5d

Observation ca56dd4b-0d50-43bd-945d-05172560eff7 · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

InsightEdit: Towards Better Instruction Following for Image Editing Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.453180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.453180Z digest=sha256:bbfca2f797037cd6f0720a3233ce284d4ae7db385e5869e0075e4d17222821bb

Observation 1f1a3217-d4b8-461e-8caa-767ac1591bf3 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

InsightEdit: Towards Better Instruction Following for Image Editing Adam: A Method for Stochastic Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.457748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.457748Z digest=sha256:a5d896b58b6018b8b061bd11b460b95437accd395dc5001a12b8f1b6659f6605

Observation cb616302-0503-4f2d-86d8-46387ef6a077 · outbound

This paper cites Gen- erating images with multimodal language models.

InsightEdit: Towards Better Instruction Following for Image Editing Gen- erating images with multimodal language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.193354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.462634Z digest=sha256:fadc904fbb12eb14320fff8d3a502a9a1d933778318acd29cf15b0d36a694390

Observation 8da70949-c9f3-4256-9c3e-a6a79f2ac986 · outbound

This paper cites VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation.

InsightEdit: Towards Better Instruction Following for Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.467898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.467898Z digest=sha256:b6069a9d115bc718902f2b826f832c32342b08088e9363b9097762600e172c73

Observation 2f37db29-59a6-42d9-bc99-caed9a1506e3 · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

InsightEdit: Towards Better Instruction Following for Image Editing LISA: Reasoning Segmentation via Large Language Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.472727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.472727Z digest=sha256:5bf8fc7f68d5179890e1074c506cf575a07e9eb5727269593bfbd90fa4a6b1e8

Observation 0f4ef9b0-c56e-4ca3-80ab-7b70d7955d68 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

InsightEdit: Towards Better Instruction Following for Image Editing Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.478124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.478124Z digest=sha256:f65237f0e5f50361e179446cd1d36699c57ac1694c006bbb776de25d091dcc19

Observation c1e1e345-5fde-4f4a-8013-4cd084f15ac4 · outbound

This paper cites Microsoft coco: Common objects in context.

InsightEdit: Towards Better Instruction Following for Image Editing Microsoft coco: Common objects in context

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.482309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.482309Z digest=sha256:83ef23d9526b505f417c1a5cbde42cbb03be188e183cbb55405c74871a3855a6

Observation 7733db1f-c23f-44dc-97a9-151a93ffcf1c · outbound

This paper cites Visual instruction tuning.

InsightEdit: Towards Better Instruction Following for Image Editing Visual instruction tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.487643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.487643Z digest=sha256:7b1f9d38d8add7c693bbde7a1f0a026a5f690798573c67b8487348d481f790a7

Observation cfcc8c28-9204-46af-a22e-70b99b42fe9d · outbound

This paper cites Language Models are Few-Shot Learners.

InsightEdit: Towards Better Instruction Following for Image Editing Language Models are Few-Shot Learners

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.492943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.492943Z digest=sha256:b8e50feaee3d701d0e2f8b6ed3856bba71e41dd09aa8107fcfcc48a44bf0abed

Observation fdf02c19-15fd-49ec-acfe-6e52678a7234 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

InsightEdit: Towards Better Instruction Following for Image Editing Learning transferable visual models from natural language supervi- sion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.497473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.497473Z digest=sha256:b909860e68abda9eec5fd71dde8d2c994c7b0684661965a69b591ccc3bd957a1

Observation d6c01bc4-8e7d-469c-ae3a-cfdc09c86c80 · outbound

This paper cites Grounded sam: Assembling open-world models for diverse visual tasks,.

InsightEdit: Towards Better Instruction Following for Image Editing Grounded sam: Assembling open-world models for diverse visual tasks,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.501781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.501781Z digest=sha256:86a7d607da4e75fcce61d6d457649d2a50089408908c98d3802764f06bfe3884

Observation d628a74b-b5be-4599-840f-ade024b74f16 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

InsightEdit: Towards Better Instruction Following for Image Editing High-resolution image synthesis with latent diffusion models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.512037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.512037Z digest=sha256:c4dad46e8395f5d8e41830df1b5caee39448a7a6e83e901be023f60ff072ae50

Observation 5a35c3bd-5b08-45d0-ab1c-4a4b750fa8ea · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

InsightEdit: Towards Better Instruction Following for Image Editing U- net: Convolutional networks for biomedical image segmen- tation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.516863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.516863Z digest=sha256:ca348551bb745641779c3f89557aaab8284d4abe34d07fe614e0b935d278675b

Observation cf115904-528c-4411-8cb5-77acd804dc56 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

InsightEdit: Towards Better Instruction Following for Image Editing LLaMA: Open and Efficient Foundation Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.521176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.521176Z digest=sha256:914a5dd3ded30606f5779276349ccbdf90ae74e521942eeeb5d4f176c200435c

Observation 32efd130-9350-476c-a03e-facd9d1d5c7a · outbound

This paper cites Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting.

InsightEdit: Towards Better Instruction Following for Image Editing Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.097887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.525213Z digest=sha256:2f0453752e6948fc85718fed0a6881a972e1ad811b92b3d3e75f3cd9fc5938c0

Observation 3d182635-52bc-4e1e-9513-7d3cad92a339 · outbound

This paper cites Smartbrush: Text and shape guided object inpainting with diffusion model.

InsightEdit: Towards Better Instruction Following for Image Editing Smartbrush: Text and shape guided object inpainting with diffusion model

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.069036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.529202Z digest=sha256:09cfe5927022c9589316475e54a40d97465cd0823b8c19c2d1dee36b2fc87234

Observation f447d16f-197e-454e-8e7a-a1533ad4324d · outbound

This paper cites DreamInpainter: Text-Guided Subject-Driven Image Inpainting with Diffusion Models.

InsightEdit: Towards Better Instruction Following for Image Editing DreamInpainter: Text-Guided Subject-Driven Image Inpainting with Diffusion Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.533203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.533203Z digest=sha256:07d2dff12048aff83f39dc6e0eccbe7fff348382eddcd5f6d3f8a218931e88d2

Observation 86b2cd0f-58d6-44a1-8edb-27a4cda6c50f · outbound

This paper cites Qwen2 Technical Report.

InsightEdit: Towards Better Instruction Following for Image Editing Qwen2 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.538501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.538501Z digest=sha256:944a411070a10baa96a15c98934523efe2adc3eb197de2ece7fa9ebd09f014c9

Observation 34aaf49b-0adb-4be0-87b3-f493d10e4670 · outbound

This paper cites EditWorld: Simulating World Dynamics for Instruction-Following Image Editing.

InsightEdit: Towards Better Instruction Following for Image Editing EditWorld: Simulating World Dynamics for Instruction-Following Image Editing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.543846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.543846Z digest=sha256:0b5544a959ee6f88dae93fee7668c99d883018e9cf0d413158a559d72e7a1687

Observation d00bebc9-12d3-4492-97cd-ffaa4de39556 · outbound

This paper cites LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model.

InsightEdit: Towards Better Instruction Following for Image Editing LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.549062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.549062Z digest=sha256:311aeacd65b1b5e774b5f744a511a0a653f482da5157ec1d4048ee8ae03ef002

Observation 3ccb8b6a-4558-49b4-9845-1b0de044f3f3 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

InsightEdit: Towards Better Instruction Following for Image Editing IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.555889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.555889Z digest=sha256:6e41e58a09d8ddb4bdc8d7a33ed10d0d0bf92661644f9c1d6e28e9f5be976af9

Observation 224856ad-fe9b-461a-a225-04d0c5c05e68 · outbound

This paper cites Modeling context in referring expres- sions.

InsightEdit: Towards Better Instruction Following for Image Editing Modeling context in referring expres- sions

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.560892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.560892Z digest=sha256:7ebb4f3e6c34d517d22c5064e69dc89c034176f4b94ef95a11b9ce4f7104a005

Observation 1c677321-181c-41d1-9f1e-0cfacc1aa4e2 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

InsightEdit: Towards Better Instruction Following for Image Editing Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.044724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.565590Z digest=sha256:e9f8f7215558f1e5814b41e48a88a59eda03961559a2a9f2857d4b265374984c

Observation a4dde6b5-45de-46ff-98b7-1eea7f7d18fc · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

InsightEdit: Towards Better Instruction Following for Image Editing Adding conditional control to text-to-image diffusion models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.570581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.570581Z digest=sha256:7f9f466f816f39ed1aa74280c210a75a60cc972c42042c23d86e2c8523741109

Observation da9bea9e-ba98-4c84-87c3-b640a70532c4 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

InsightEdit: Towards Better Instruction Following for Image Editing The unreasonable effectiveness of deep features as a perceptual metric

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.574739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.574739Z digest=sha256:d3779ea7f498efb3e66a9f48ae59c0f5afd46a192777b42cb3218d021980a001

Observation a95793dc-a990-4046-84a9-7c30c9ca6501 · outbound

This paper cites Hive: Harnessing human feedback for instructional visual editing.

InsightEdit: Towards Better Instruction Following for Image Editing Hive: Harnessing human feedback for instructional visual editing

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.006566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:19:22.580148Z digest=sha256:f21d4ecbf2457d2f9e63538b3cbe659f7f301875036fb6056403bd98ecf11d9e

Observation 4ed2c6d5-5dfa-4d7e-865f-9236dc97d98c · outbound

This paper cites UltraEdit: Instruction-based Fine-Grained Image Editing at Scale.

InsightEdit: Towards Better Instruction Following for Image Editing UltraEdit: Instruction-based Fine-Grained Image Editing at Scale

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.587711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.587711Z digest=sha256:b89374cf35c473c027897b8f5af23db26b1c486a7776a872091b5d5530e32b78

Observation c5f2084e-7948-4f04-a07f-57d1120e7dcc · outbound

This paper cites A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting.

InsightEdit: Towards Better Instruction Following for Image Editing A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.592578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.592578Z digest=sha256:11617e2c0b60aa8f00eb7cc9d3a6c9e8ba73832a192dcd1c6d37cdf6cc29e60f

Pith citing papers

Observation 17882952-f7a9-410d-b29f-8dd72fff0933 · inbound

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer cites this paper.

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer InsightEdit: Towards Better Instruction Following for Image Editing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:07:53.167573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T16:07:53.054355Z digest=sha256:4e6842089501076a0be5d9d95ebb97f7488615d87e66402a333e772ab771f142