Pith. sign in

Paper Citation Record · LEDGER

R-Genie: Reasoning-Guided Generative Image Editing

As of 8 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 2 inbound Pith citation observations for arXiv:2505.17768.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17768 v2

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:39.681727Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T21:17:04.368521Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:19:13.598382Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e49fde3-c98f-42c9-adc3-28511d6d7a2d · outbound

This paper cites GPT-4 Technical Report.

R-Genie: Reasoning-Guided Generative Image Editing GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.237102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.237102Z digest=sha256:9c7ea2500ce925a5e546b70d12e9a6a9210c0cdbfacbf299c04cfcf7ffd3fb56

Observation c7170823-5090-4293-a498-4f0b8a3d8b03 · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

R-Genie: Reasoning-Guided Generative Image Editing Instructpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.293706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.293706Z digest=sha256:65bee1789f435392ce44aae7fb0ba252bb2c25bde1c8b4bf92f01a6e75f5d70e

Observation cf8fdaf7-4fae-4a8c-bbdb-27597834ea0e · outbound

This paper cites Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis.

R-Genie: Reasoning-Guided Generative Image Editing Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.401485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.401485Z digest=sha256:d6a7cabf53a7fa8fbff85bee2d5f1e181ae07e10671cfb92b24587b8d0a31f28

Observation 0d226bed-734c-4ee3-b098-aae642972e61 · outbound

This paper cites The revolution of multimodal large language models: a survey.arXiv preprint arXiv:2402.12451, 2024.

R-Genie: Reasoning-Guided Generative Image Editing The revolution of multimodal large language models: a survey.arXiv preprint arXiv:2402.12451, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.488330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.488330Z digest=sha256:46583a3c5a92059d051c945cdce809e041626cffb214c6a914723fa4eb379759

Observation b2158e25-d588-4187-81aa-6ce1386dbc29 · outbound

This paper cites A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT.

R-Genie: Reasoning-Guided Generative Image Editing A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.591816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.591816Z digest=sha256:504df40b8fe3e8ee63c204b0d848dfafe5111af0c9dcc2d4996a33cd36e77b54

Observation 977c83a1-fbaa-4502-b6d9-7d1d5c7197f3 · outbound

This paper cites Diffusion models in vision: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(9):10850–10869, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Diffusion models in vision: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(9):10850–10869, 2023

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.742882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.742882Z digest=sha256:d86bb69a20dc2fd9b4f299e21d0a8e491832486b7773001e1305fa8ce642f486

Observation be32fd37-cc73-4585-8623-2b3b7073afb3 · outbound

This paper cites GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing.

R-Genie: Reasoning-Guided Generative Image Editing GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.827824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.827824Z digest=sha256:7aef4b7787414b4b860a3b51184aa31191ae0c9a7a86e804cf3fedd64848187e

Observation cf3dccb2-5a83-4a76-98ec-f6d6d703cb7f · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

R-Genie: Reasoning-Guided Generative Image Editing Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.920045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.920045Z digest=sha256:8aa48be2705342cd129675ab5c3da0fec74841b450c65c3290b6452ed76da6a3

Observation b4afb409-feb4-4b0d-a8df-5cf72c265488 · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

R-Genie: Reasoning-Guided Generative Image Editing Blink: Multimodal large language models can see but not perceive

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.011707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.011707Z digest=sha256:3ddffce53b564b08ab45a28740a6b0f7d90757141e6e4b6da3d587694548c06d

Observation ff78ff11-3867-42da-a0fb-0097c1696894 · outbound

This paper cites Exploiting clip self-consistency to automate image augmentation for safety critical scenarios.

R-Genie: Reasoning-Guided Generative Image Editing Exploiting clip self-consistency to automate image augmentation for safety critical scenarios

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:44.073402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:35.110852Z digest=sha256:d84886286f06a96f93e242b9245f68fc737218f49a373fc7f3c4162319a2ea6a

Observation 15b2ccd7-63dd-4da4-a229-9687a5561ce0 · outbound

This paper cites Image style transfer using convolutional neural networks.

R-Genie: Reasoning-Guided Generative Image Editing Image style transfer using convolutional neural networks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.203450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.203450Z digest=sha256:b6d716d43b03da858f21cf67f7b94c9eb4c86fcc91c25bd88323df2442b43c8a

Observation d75c3803-9d2b-445d-8e6f-fc17c60aa1e1 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

R-Genie: Reasoning-Guided Generative Image Editing SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.306595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.306595Z digest=sha256:7f9ebf23ac680b7bd48c80d81f5bb2e317c124e69d38582c05fcd598b5a3d46f

Observation c51a1187-448f-44f3-9fc8-075f158e1007 · outbound

This paper cites Instructdiffusion: A generalist modeling interface for vision tasks.

R-Genie: Reasoning-Guided Generative Image Editing Instructdiffusion: A generalist modeling interface for vision tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.392031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.392031Z digest=sha256:326d5ca7ed2863bb75e935d52ec90106dffc89cca55d80f148e7c32ff7bbca62

Observation e455fd12-1278-4f08-b577-670c702786a6 · outbound

This paper cites Artificial general intelligence: concept, state of the art, and future prospects.

R-Genie: Reasoning-Guided Generative Image Editing Artificial general intelligence: concept, state of the art, and future prospects

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.884110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:35.491457Z digest=sha256:9da04ced6b63c798b806d55f929e22a777fa1f1429fdad6834acc17a38c0ae6a

Observation 2ac678eb-afe3-46e9-bae6-b78988a74608 · outbound

This paper cites Diffusion models in low-level vision: A survey.

R-Genie: Reasoning-Guided Generative Image Editing Diffusion models in low-level vision: A survey

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.736396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:35.585365Z digest=sha256:4857b2549a30427be76ad2af3c99454f2a3f70fde151c068cdab50d7e83b196c

Observation 3c97427e-f89d-430a-b3f8-3ba0007d5ea4 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

R-Genie: Reasoning-Guided Generative Image Editing Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.684189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.684189Z digest=sha256:a2c01d5b0feb576a577284db340362fd21b4955a1d504bd3687a43552ce4f97b

Observation 5c16cacb-53c3-41a1-9127-d693984b9095 · outbound

This paper cites Smartedit: Exploring complex instruction- based image editing with multimodal large language models.

R-Genie: Reasoning-Guided Generative Image Editing Smartedit: Exploring complex instruction- based image editing with multimodal large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.753105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.753105Z digest=sha256:34a622e16162f8cade954166069314c060b7db9751935c8cc4afb40c077bc5b8

Observation 1b3064e8-3f9c-4ffa-9e35-eb2a98b81c08 · outbound

This paper cites Image-to-image translation with conditional adversarial networks.

R-Genie: Reasoning-Guided Generative Image Editing Image-to-image translation with conditional adversarial networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.845993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.845993Z digest=sha256:5d785a9537ced51121f55253f7720871885d248a60297e3348097f7b93960b86

Observation 4fb701a8-a7fa-4fd5-adfd-b33b36b63ab6 · outbound

This paper cites UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding.

R-Genie: Reasoning-Guided Generative Image Editing UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.925099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.925099Z digest=sha256:65f1af85ca5951ab4cd061ee6bd013a0f0ce6db7cfe85cf39390b777bb86a0e5

Observation 32b68690-cada-4a52-9224-4df1407d3dd2 · outbound

This paper cites A style-based generator architecture for generative adversarial networks, 2019.

R-Genie: Reasoning-Guided Generative Image Editing A style-based generator architecture for generative adversarial networks, 2019

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.992923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.992923Z digest=sha256:1976f19ad4390649aa5f1ff7c3ebde31798e0321a473a4d02d6fc0063ec1ce8d

Observation 460a7837-d519-4da3-af38-7a93203fac1b · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing Imagic: Text-based real image editing with diffusion models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.094232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.094232Z digest=sha256:15557122fe00994fbf6592437b7556828d4fd6b3de3654a16cac871cebb6beee

Observation 22507636-45af-4598-92d5-cb1de7016103 · outbound

This paper cites Referitgame: Referring to objects in photographs of natural scenes.

R-Genie: Reasoning-Guided Generative Image Editing Referitgame: Referring to objects in photographs of natural scenes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.194091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.194091Z digest=sha256:a8a48e246330df41ba22e06d900221df71e3ae4549fed70d20960a22b3a4ec1e

Observation 42958a15-5507-41a5-b6e9-be6b67421afb · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

R-Genie: Reasoning-Guided Generative Image Editing Lisa: Reasoning segmentation via large language model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.300104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.300104Z digest=sha256:225cd46b3b2b6a7ca3c7660685f629863c92d7181032f35bd3ed55bf3fe8a5cc

Observation 7d436560-1fad-4610-b587-a7041d49c8f4 · outbound

This paper cites Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks.

R-Genie: Reasoning-Guided Generative Image Editing Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.393741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.393741Z digest=sha256:7f880e9abfff1b08563f4032d854575b63d9dbe3726035d64dcd338044dc4f15

Observation d79587cf-fdb0-434e-b252-53dae09a525f · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

R-Genie: Reasoning-Guided Generative Image Editing Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.491967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.491967Z digest=sha256:6c40b133bb0177f59fc68cac36df306e1a2d99a85718efbcb62bf2579ddda364

Observation 43e0ad56-47b0-41eb-a547-704a6601ec3b · outbound

This paper cites Textbooks Are All You Need II: phi-1.5 technical report.

R-Genie: Reasoning-Guided Generative Image Editing Textbooks Are All You Need II: phi-1.5 technical report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.613898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.613898Z digest=sha256:c1faf3c9480fe47d096cc2a47f90d053d69161c823456de82ebbba7da80a19a6

Observation 4985be64-915d-436d-b285-27fdc134fecc · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.712924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.712924Z digest=sha256:4e282dbb4a9c2f3cddb9cc62e16cdc5015d25d81e166ec66bad762929de66cf9

Observation a9ff9498-5718-4a7e-b70b-46fd765e5f77 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

R-Genie: Reasoning-Guided Generative Image Editing Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.841001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.841001Z digest=sha256:5bf6fe49296e41934be79a03c9a6e7ec236a2cb9be119dca766f56124b5011f5

Observation c9305ab5-734a-45ac-a432-87e72263a1d1 · outbound

This paper cites Decoupled Weight Decay Regularization.

R-Genie: Reasoning-Guided Generative Image Editing Decoupled Weight Decay Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.939976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.939976Z digest=sha256:965c2be7060545b99e0ccd52eeab90eab2501cf514c44037090f2041cb37c3ca

Observation 86194af1-dce4-4632-822e-cf70268ff6a0 · outbound

This paper cites Adapedit: Spatio-temporal guided adaptive edit- ing algorithm for text-based continuity-sensitive image editing.

R-Genie: Reasoning-Guided Generative Image Editing Adapedit: Spatio-temporal guided adaptive edit- ing algorithm for text-based continuity-sensitive image editing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.467782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.034702Z digest=sha256:86317bca1e153f741408af1abfbc136c6ee74d1c054fafe9a19381c781115b52

Observation 70e3498a-ca8b-46c5-874b-7a91e9805c8f · outbound

This paper cites Hd-painter: High-resolution and prompt-faithful text-guided image in- painting with diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing Hd-painter: High-resolution and prompt-faithful text-guided image in- painting with diffusion models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.197517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.168546Z digest=sha256:810bcaf9f3e85b23d05357a461c3d39d4498a191ec98cac368ff288ba6253f9c

Observation 38069fa0-33f8-4dd1-8217-5f39252bb6ef · outbound

This paper cites Toward verifiable and reproducible human evaluation for text- to-image generation.

R-Genie: Reasoning-Guided Generative Image Editing Toward verifiable and reproducible human evaluation for text- to-image generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.064661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.293505Z digest=sha256:754dae842fefaa63521f382bb4c6ad3a4c68c47c3383db36ba3e6366202afe03

Observation 1b8b0ad5-1da0-41be-97c5-cd0e7ca57bdd · outbound

This paper cites State of the art on diffusion models for visual computing.

R-Genie: Reasoning-Guided Generative Image Editing State of the art on diffusion models for visual computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.910819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.381702Z digest=sha256:2f03229ab3050723039cc400f239fa784c6ec3cb92266edc0893f9fb9b9d59b4

Observation 5cd6a7fa-891c-406f-a07e-9c06bd783107 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

R-Genie: Reasoning-Guided Generative Image Editing SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.468840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.468840Z digest=sha256:3155c22c026a655172a432a700e42e39cf18172547251278aa5376b78a790c2c

Observation d9baa1b4-04bc-4cf8-9107-e0d08b3bee61 · outbound

This paper cites Learning transferable visual models from natural language supervision.

R-Genie: Reasoning-Guided Generative Image Editing Learning transferable visual models from natural language supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.532693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.532693Z digest=sha256:3321821c0e7eef72062f97dd38a6b1e0430b1a375246ed4ca2e95ed04b75bf0a

Observation b898e07b-5ab6-422a-8318-be7f4687578d · outbound

This paper cites High- resolution image synthesis with latent diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing High- resolution image synthesis with latent diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.617434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.617434Z digest=sha256:bed3b3a8be53efc5daef59739164548907ca1a95f6315613ee400605830adb08

Observation b4d59a39-b2c1-4378-b9de-9a6658a2243a · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022.

R-Genie: Reasoning-Guided Generative Image Editing Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.684125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.684125Z digest=sha256:cc3a12e8b9b4aefc93f8e71a501f4f953f454993b9e3e5fd4eb6ea81811fbbe6

Observation c9915cba-c2d2-4b06-b156-7260feb27a4b · outbound

This paper cites Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022.

R-Genie: Reasoning-Guided Generative Image Editing Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.751701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.751701Z digest=sha256:fc4c09c08a690a2abc807340b1f1033fdbf4a6d6c6a5b84505fcba750b891603

Observation 2f5098da-430d-43f3-a364-3a9caccbfbe2 · outbound

This paper cites Imagdressing-v1: Customizable virtual dressing.

R-Genie: Reasoning-Guided Generative Image Editing Imagdressing-v1: Customizable virtual dressing

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.714990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.815535Z digest=sha256:108c0adc0a45c40b5c69ec0940fdc552417ba4c02a48cec118f3c38d4f751c15

Observation 419e9b11-f379-40ac-9f57-57a65fe94717 · outbound

This paper cites Imagpose: A unified conditional framework for pose-guided person generation.Advances in neural information processing systems, 37:6246–6266, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Imagpose: A unified conditional framework for pose-guided person generation.Advances in neural information processing systems, 37:6246–6266, 2024

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.515874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.849871Z digest=sha256:7bf618403a5e6414a4258e65ff7d8ef9eca564fd15c76ba26278e02dcb5f5c2d

Observation b7595ff7-0276-4a9c-886d-bf90c2dca095 · outbound

This paper cites IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design.

R-Genie: Reasoning-Guided Generative Image Editing IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.926613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.926613Z digest=sha256:1f065965582d1d88717c0af7da1683350794a4170f5cad1c90a8f52fd163e922

Observation 6acb34f7-d88b-4490-a42f-01f0586fc529 · outbound

This paper cites Learning by planning: Language-guided global image editing.

R-Genie: Reasoning-Guided Generative Image Editing Learning by planning: Language-guided global image editing

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.335156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:37.994473Z digest=sha256:305476f3de269a3fee8a8ac8bd82ea1c32d05853934083255f20d5efed03a915

Observation cc0374c5-1128-41a2-bad8-4d4286421a60 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

R-Genie: Reasoning-Guided Generative Image Editing Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.045323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.045323Z digest=sha256:ccca74efeea38621fa179aba6f106faa47b599c20c89f355a10f4d404675a01b

Observation 1b059a4c-f89d-47ba-9326-617b3c2a808f · outbound

This paper cites MetaMorph: Multimodal Understanding and Generation via Instruction Tuning.

R-Genie: Reasoning-Guided Generative Image Editing MetaMorph: Multimodal Understanding and Generation via Instruction Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.112559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.112559Z digest=sha256:ec65a2937eae71e5a8369fa6293abb4e816cc4b52e794845bbfe313dce3017e2

Observation e760d758-4fd0-4e17-a9a7-09e5309e7dc5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

R-Genie: Reasoning-Guided Generative Image Editing Emu3: Next-Token Prediction is All You Need

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.186434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.186434Z digest=sha256:bdceca174b2f3a7f9bf5f68cd72f6cfbe8d468c8fc251e74eca76438b1c72888

Observation 3dd118ec-a2ff-429b-9674-0d4c83321c4e · outbound

This paper cites Gpt4video: A unified multimodal large language model for lnstruction-followed understanding and safety-aware generation.

R-Genie: Reasoning-Guided Generative Image Editing Gpt4video: A unified multimodal large language model for lnstruction-followed understanding and safety-aware generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.109004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:38.282879Z digest=sha256:9c6ed519095e1799930aa0f8d83a1e574ed5c922c502a684c476a05542701bb2

Observation e8a3a7c3-cc0f-4f3e-9171-1d975eb88dbe · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.353174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.353174Z digest=sha256:d478ecf1d227bcb907d7b9f2cac34a0c7a2ac7fe190a3709fb690e8bc191c588

Observation a4b8daa1-7e9f-4be8-a68e-7a2be09186f1 · outbound

This paper cites Next-gpt: Any-to-any multimodal llm.

R-Genie: Reasoning-Guided Generative Image Editing Next-gpt: Any-to-any multimodal llm

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.429401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.429401Z digest=sha256:0db3d4647be5ebe01f0bb3dc51e15f2d4328a694947f702f9688e8c064bbc8a4

Observation 150bdce3-bef3-4d05-8836-e2e2562762bf · outbound

This paper cites Multimodal large language models make text-to-image generative models align better.Advances in Neural Information Processing Systems, 37:81287–81323, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Multimodal large language models make text-to-image generative models align better.Advances in Neural Information Processing Systems, 37:81287–81323, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.812068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:38.504431Z digest=sha256:69dfa2483c67f462fe8fbe0fd7bb8d26af099a1eeb08350488f25ecdc282da0d

Observation f7e366e3-c78a-4304-b641-8e4daf21fe8e · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.577390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.577390Z digest=sha256:1184c42294d02dc915216559a44a28ba0b9fd39346c2ff5eb7448f09fe23f557

Observation 92923d8f-0a4a-44d5-91a3-6886fe48aca4 · outbound

This paper cites Omnigen: Unified image generation, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Omnigen: Unified image generation, 2024

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.672977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.672977Z digest=sha256:1c00d55b3160c802eccee3a8bd93f8ff4ae13e3da25964d254d2f15e29584a17

Observation e6e38987-118a-4238-a13a-c2ec7d61fce6 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.745637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.745637Z digest=sha256:4ac0fe8a776ccbe7417c47e477421696c153dcdb5f332b8d3f69509538dcb1a7

Observation d8119d7f-4d20-4d14-a2f9-776d8f273181 · outbound

This paper cites Smartbrush: Text and shape guided object inpainting with diffusion model.

R-Genie: Reasoning-Guided Generative Image Editing Smartbrush: Text and shape guided object inpainting with diffusion model

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.547686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:38.819025Z digest=sha256:c82a190263b4e3f13b6dea291aea695c59f43a99e5772cb36cbee6b4ff2b95ac

Observation 0b3c697b-ac14-4018-9871-1329f287f494 · outbound

This paper cites A survey on video diffusion models.ACM Computing Surveys, 57(2):1–42, 2024.

R-Genie: Reasoning-Guided Generative Image Editing A survey on video diffusion models.ACM Computing Surveys, 57(2):1–42, 2024

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.878723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.878723Z digest=sha256:59ce339c7a3b17ef7fd3c3488bc334ff60a79186226069a5dd26229b89c71717

Observation 0e16215f-03e4-4b49-b64e-f0983e3f2e6c · outbound

This paper cites Progressive instance-aware feature learning for compositional action recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(8):10317–10330, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Progressive instance-aware feature learning for compositional action recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(8):10317–10330, 2023

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.306913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:38.971508Z digest=sha256:1ea6831d070753a907d499d90036e781fee584d8cb73b1ce01addebaea82d830

Observation 505c5eb0-70e6-448d-9fd8-011f92aaa11a · outbound

This paper cites Higcin: Hierarchical graph-based cross inference network for group activity recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(6):6955–6968, 2020.

R-Genie: Reasoning-Guided Generative Image Editing Higcin: Hierarchical graph-based cross inference network for group activity recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(6):6955–6968, 2020

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.071809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:38.976681Z digest=sha256:ba4fd5cebb4b843788abaa2f8927042add239130f16ada88f38ac34756429132

Observation 03c58007-bb48-455c-a611-874a8cbf01c7 · outbound

This paper cites Mmginpainting: Multi-modality guided image inpainting based on diffusion models.IEEE Transactions on Multimedia, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Mmginpainting: Multi-modality guided image inpainting based on diffusion models.IEEE Transactions on Multimedia, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:40.826518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:38.980292Z digest=sha256:09a927d5c24ecb66d5098621bd8cca576b196c92b80b089051f8a088b6c19905

Observation c5a12dd0-34ab-4ec3-b017-c0bec96a2688 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

R-Genie: Reasoning-Guided Generative Image Editing MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.040135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.040135Z digest=sha256:cc73e0f7de7d7b9f54f684abc03b1ecb15c7f3cabf812e92de2e47a05f6eff6e

Observation 04e15a43-4534-4a46-aa4a-0ae9a0120d08 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.182063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.182063Z digest=sha256:7f6388cc6d089d789116afd8b51dba7c2faa4b7f11a36b8f64d8f099b04e174f

Observation 744f85c0-e23a-432b-b543-091f43e88fd0 · outbound

This paper cites Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing.

R-Genie: Reasoning-Guided Generative Image Editing Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.293055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.293055Z digest=sha256:2d2e41b3b0f8a8a0ae3022ea1883ec05dbb6175921c9d56513a4ba90b5fa6b28

Observation e97933a5-dccf-44e7-b652-638c0c3e53c5 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

R-Genie: Reasoning-Guided Generative Image Editing Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.426759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.426759Z digest=sha256:800048dd7b9270075f972757a04d19ad5926590895a952ba7fc9b785f6269293

Observation a28de295-3447-4a9d-933b-e887f09de757 · outbound

This paper cites Unpaired image-to-image translation using cycle-consistent adversarial networks.

R-Genie: Reasoning-Guided Generative Image Editing Unpaired image-to-image translation using cycle-consistent adversarial networks

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:40.499769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:39.554364Z digest=sha256:94e8c4ae93a17b36d34f46137b7cd52bae48049e6d3df327b388c7f45006884b

Observation 8c12e8ab-acc2-421e-b714-4bd3c6ea149b · outbound

This paper cites an unresolved cited work.

R-Genie: Reasoning-Guided Generative Image Editing Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:46:40.170631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:46:39.681727Z digest=sha256:db39a58e17fbb6d58a63272fd1cccdc709b34c61bb443f3f97e2a75f2f1dc98e

Pith citing papers

Observation 7363a002-f8f5-4e45-8f99-ee6d57bd8526 · inbound

ASTRA: Let Arbitrary Subjects Transform in Video Editing cites this paper.

ASTRA: Let Arbitrary Subjects Transform in Video Editing R-Genie: Reasoning-Guided Generative Image Editing

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:16:14.314804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T10:13:10.141426Z digest=sha256:607b93b70ba1b60fc24e1b3d43382ac597712186f88029abfbd3962c0e2740f4

Observation 847748c4-8358-48af-8c51-90af5092ed71 · inbound

ProductConsistency: Improving Product Identity Preservation in Instruction-Based Image Editing via SFT and RL cites this paper.

ProductConsistency: Improving Product Identity Preservation in Instruction-Based Image Editing via SFT and RL R-Genie: Reasoning-Guided Generative Image Editing

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.600228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T21:17:04.368521Z digest=sha256:2ca0d5e9532cfcb58fcec064f78bfb0dec7a687f668bee9a677f94e4dad8231d