Pith. sign in

Paper Citation Record · LEDGER

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent

As of 9 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2508.20505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20505 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:10:14.568791Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact1
  • verified fuzzy38
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 873047e1-45d2-4384-968a-1bdc6c5eb475 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.221159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.366413Z digest=sha256:41ceffda7c63b9e08b808583be52385d29e70c62b4607a641d8c5b2413cafc0f

Observation 6dd68923-635b-4985-9099-bf37db89baa0 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent In- structpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.210951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.370756Z digest=sha256:cfdda519d787c605ef5b537ca38d720105ea48cbc0b51ebec39db706ecbcdaf0

Observation 2c8724ba-6f65-414e-bccd-300728a1672a · outbound

This paper cites Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.201010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.375562Z digest=sha256:49af47aca3595c8e7d12d98cbf902fba1b944770dd72b6b3766e2fc6acfbd12d

Observation b6af6a11-5de4-4a0f-960a-b7c45fcfd6de · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Emerg- ing properties in self-supervised vision transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.379319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.379319Z digest=sha256:770df3a6a6c3171d9e8da2930923d6624b97c57174f8b71f6444e7f4dcfe591d

Observation 9d4a29e4-45e6-493d-b2b2-3f4834ffbefc · outbound

This paper cites Diffusion forcing: Next-token prediction meets full-sequence diffu- sion.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Diffusion forcing: Next-token prediction meets full-sequence diffu- sion

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.186609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.383783Z digest=sha256:ec61b289493ae32a53bb60ad960f91f5090919066318f758a1e9b519fa44db55

Observation b74fd36d-8b90-4247-8966-49a76ef9ae26 · outbound

This paper cites Pixart-alpha: Fast training of diffusion transformer for photorealistic text-to-image syn- thesis.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Pixart-alpha: Fast training of diffusion transformer for photorealistic text-to-image syn- thesis

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.176450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.387814Z digest=sha256:9ee1c4fb8c7a8a0a0bc51ce7c12cb79fd989e9d56d05c843cf6c3d7b700f8521

Observation 0bc4d044-4d0a-45d6-9088-17efc8330089 · outbound

This paper cites Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.391535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.391535Z digest=sha256:a92b805aa7daba5ab05780b95137b34cf253eac305de77b050f52cdd0312e7df

Observation cf3d2ddb-af56-4f52-b153-98ae7e90ac46 · outbound

This paper cites Turboedit: Text-based image editing using few-step diffusion models, 2024.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Turboedit: Text-based image editing using few-step diffusion models, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.166902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.395872Z digest=sha256:9b2762ed7e2cc35f7166d936e093681e6e83f41c38b6fe2ccddaf632f53be54e

Observation 75025253-130a-49df-aa85-3c698b23b423 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.156246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.399244Z digest=sha256:1a7e0e93ba4fb192a0bc025bacb9231db0eeefb50f5bf498879e705b9e83e550

Observation 741dd28f-0c68-444b-9116-163a8804632a · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.144431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.402557Z digest=sha256:b26f7b85c1a26d81d8eab4de03fd4b61c849df6d8041bbc4b7b4c8e7ea995c51

Observation 7155195f-817f-4bbc-ba7a-318e3916e328 · outbound

This paper cites DiT4Edit: Diffusion Transformer for Image Editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent DiT4Edit: Diffusion Transformer for Image Editing

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-05T15:10:14.696710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.406081Z digest=sha256:1b2bea124c954db5417e888b67177201457861fea9866e2f57a62f028268eb27

Observation 8207fb46-71d6-4e34-ae97-fb96b68a107c · outbound

This paper cites Guiding instruction-based im- age editing via multimodal large language models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Guiding instruction-based im- age editing via multimodal large language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.134167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.409650Z digest=sha256:ba11a83f4fcc7b119c29ffb62a55558e9e7c82357ea7085f4ad42ed454d83022

Observation 7d481ae5-0271-4f94-8ed8-24a62f684f64 · outbound

This paper cites Generative adversarial nets.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Generative adversarial nets

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.124614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.413283Z digest=sha256:5d12a5f74febdf06707952633cd5527104358a4f2a3298ee56cd5163b6fbca71

Observation a181feeb-481e-4705-8572-6dbf397e1820 · outbound

This paper cites Prompt-to-prompt image editing with cross attention control.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Prompt-to-prompt image editing with cross attention control

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.115136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.416974Z digest=sha256:e053c68871ac3f5293300921fa83a3116c7ae11dbb3b6febe7ccf8d1c610f838

Observation b847dcd4-d429-4a0c-96c1-bb86e4a46a90 · outbound

This paper cites Classifier-Free Diffusion Guidance.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Classifier-Free Diffusion Guidance

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.419892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.419892Z digest=sha256:2a80a965ad4fdb5cd6104019ba760af3238791c580730762d1aaaa105af3e707

Observation b11cba70-3b05-4cd3-92b9-b6b4d1a7e0a4 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Denoising dif- fusion probabilistic models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.104620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.423659Z digest=sha256:f6202387b831a811cc6191d42ee88e6fa11e79e3d40999cea129f5050fb823d5

Observation 65405db1-cb6c-4a2e-8f63-67a43f94a2da · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Lora: Low-rank adaptation of large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.094425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.427085Z digest=sha256:6f966fe9cc4000885c0f46eeb4ee6e33769c91027335f745e72490552b10e6f2

Observation 7f0df120-c809-473f-ae42-02916aaf6471 · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Animate anyone: Consistent and controllable image- to-video synthesis for character animation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.083347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.429903Z digest=sha256:e204229aea09296ba5f27df01b4c631e6ca3e73a67553f78dab6bf6b35040b9d

Observation 95e2efd5-f7c6-47f4-933d-ddfe32c05289 · outbound

This paper cites Smartedit: Exploring com- plex instruction-based image editing with multimodal large language models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Smartedit: Exploring com- plex instruction-based image editing with multimodal large language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.072441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.432942Z digest=sha256:aabb64ac6b376e531c35d76a9990c84d0229d29701a0f79156997af0ab71a485

Observation f80ef37f-7f00-4b8c-bbd6-66e41f8970d4 · outbound

This paper cites Pnp inversion: Boosting diffusion-based editing with 3 lines of code.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Pnp inversion: Boosting diffusion-based editing with 3 lines of code

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.062003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.436996Z digest=sha256:6880b1c9c607995951cc721dec667e07523ffefcfc1dda633ec61bf2821f9c52

Observation b9052845-83d0-4a8c-a8ca-46880f4ab115 · outbound

This paper cites Multi-concept customiza- tion of text-to-image diffusion.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Multi-concept customiza- tion of text-to-image diffusion

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.441184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.441184Z digest=sha256:d48b693977c46d968d511d734e7089f9bea686329f9bc444d481208c63dcf01c

Observation a2908287-a3f4-4a3e-bff6-c29aae809b22 · outbound

This paper cites an unresolved cited work.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:10:15.044026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.444737Z digest=sha256:727185c656485869f1b6fa95d1701fdf829e007e78e3493ea7700a7e098d91e2

Observation 785190e3-9f2f-46a0-a2a2-93f336272988 · outbound

This paper cites Autoregressive image generation without vec- tor quantization.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Autoregressive image generation without vec- tor quantization

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.034343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.448780Z digest=sha256:e23abb08cce30c6799a01b84868e049dd0e8edcb564134a1437e1d92a452a557

Observation 72347049-456b-4cb5-b024-c6d207095713 · outbound

This paper cites BrushEdit: All-In-One Image Inpainting and Editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent BrushEdit: All-In-One Image Inpainting and Editing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.451707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.451707Z digest=sha256:f73b19a9c077c1092158d1d902ff86f56332f1cc923c72cb595b02889d1aea4d

Observation 04b12cb7-9f11-4a84-a4a2-1844c2f942c1 · outbound

This paper cites Flow Matching for Generative Modeling.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Flow Matching for Generative Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.455855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.455855Z digest=sha256:25817a40d3310aa3cc0d4676e064321a964e85c5ace858a8ec5805b77acdf3f6

Observation 21883f9d-c811-44b3-ba15-f7a42704f773 · outbound

This paper cites Towards understanding cross and self-attention in stable diffusion for text-guided image editing, 2024.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Towards understanding cross and self-attention in stable diffusion for text-guided image editing, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.023343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.460066Z digest=sha256:8d93130f614851fdc39d5517aa1683b09f4ebb93b2e89659bc44716b080ba4b1

Observation 0cdeaad7-b23b-4092-94d0-d1dd9802bd57 · outbound

This paper cites Towards understanding cross and self-attention in stable diffusion for text-guided image editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Towards understanding cross and self-attention in stable diffusion for text-guided image editing

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.013665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.463689Z digest=sha256:f08ecfefcd18ff8fa044f740675a326ca7cdab9b6ec1fa2881dcfa5d1f3eee11

Observation 7ff4a2cb-6245-4c9b-9a94-0da6fa1ad59f · outbound

This paper cites Decoupled Weight Decay Regularization.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Decoupled Weight Decay Regularization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.467655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.467655Z digest=sha256:c0cf0443f1a755b053c7f06640070e0481ca384d4fa646ecbb54724a7e1a88e3

Observation 5095fcc8-9436-436a-858c-46da89a7462d · outbound

This paper cites Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:15.004001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.471018Z digest=sha256:7f5dee82859d237b7a50372c4b6103864e8e3b59748f1e36af41160ccdb83725

Observation e0d28f0d-feed-4503-b1ce-f0345b4f129f · outbound

This paper cites Scalable diffusion models with transformers.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Scalable diffusion models with transformers

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.993046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.474297Z digest=sha256:45da5da7a8a9a2ed88f7a48b1b245c4fb13fbad5c4c0e466ea3affb5bb391030

Observation 3795b9d1-02ae-483d-bb8c-57e180178676 · outbound

This paper cites Scalable diffusion models with transformers.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Scalable diffusion models with transformers

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.982645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.477499Z digest=sha256:24f31833d3147b619f540a9b31db37b21b2ebf4865c0ace3fdebce0d899d77f8

Observation 176a27db-6ae6-45ba-99bf-dd33bcfdbd1f · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.972046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.480384Z digest=sha256:1fc2a5fb70e524540f0add318c63123d0143e26f4e9ab75e9208d1cc3fa0ee51

Observation 6b3c34b5-3ba5-4928-9f67-15d8cc939a22 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Learn- ing transferable visual models from natural language super- vision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.961052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.483586Z digest=sha256:092681ab6ebb44fe752860b509a29d486bb9d5fdfaefed733470685bb0af7643

Observation ef7fd576-e89c-41ce-9a2b-4b07bfb0470b · outbound

This paper cites Zero-shot text-to-image generation.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Zero-shot text-to-image generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.949979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.487851Z digest=sha256:46e28b7d0730e7ad40883e2a512d3aa503a87fd1cfbe6004d0752d375aa02a4f

Observation aa758a2c-7ccb-4d69-8e70-66afa941e34b · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.491517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.491517Z digest=sha256:cd2713c527c15be93c0917be7bf3c0ffa4ba8719c23be259698bc80b8788876a

Observation 3359af09-24a4-4f54-8224-87a62e8c1d5a · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent High-resolution image syn- thesis with latent diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.495189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.495189Z digest=sha256:604f5bc17554ff584dc65355384529abc5097b7dc851e4fd0d66277dc4e9e6f9

Observation 790d3e38-52ca-44c5-83d9-b002e363691c · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent High-resolution image syn- thesis with latent diffusion models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.931624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.499744Z digest=sha256:9953ab22193cfbed60a027ba240dfcf48db222389c6bcd2b51c5db1eedb7df46

Observation b8d947b9-dfbe-422a-b13e-f8894dc8e72d · outbound

This paper cites Semantic Image Inversion and Editing using Rectified Stochastic Differential Equations.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Semantic Image Inversion and Editing using Rectified Stochastic Differential Equations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.502670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.502670Z digest=sha256:b187be106aa4e0f1a34364198628dc80479f06b97410ea68504523e50e76b5f0

Observation 33112e1e-f964-4b74-9887-4dd99c61a60d · outbound

This paper cites Laion-5b: An open large-scale dataset for train- ing next generation image-text models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Laion-5b: An open large-scale dataset for train- ing next generation image-text models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.846170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.505793Z digest=sha256:f287ef00138373e059c37d984aec3015e5df59bc93d2d1f3e3dbc0608b97fa13

Observation 906c8b8f-47f8-4a34-a8b7-a9994a281c00 · outbound

This paper cites Emu edit: Precise image editing via recognition and genera- tion tasks.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Emu edit: Precise image editing via recognition and genera- tion tasks

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.835840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.508700Z digest=sha256:59b44e8972e1834e1c9a7dc35925834034145abbdd0ce8c418615088c1019ad9

Observation 5599616e-50a4-4cfc-a11a-a5e193d26f57 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Deep unsupervised learning using nonequilibrium thermodynamics

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.824680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.511512Z digest=sha256:83318b72b780ab70f09b338178288197429a5a0a5f5174f0be18beefc73c151d

Observation 7d4406f6-4a5e-466c-94f3-ffa9a0780f0e · outbound

This paper cites Postedit: Posterior sampling for efficient zero-shot image editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Postedit: Posterior sampling for efficient zero-shot image editing

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.813397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.515010Z digest=sha256:44e229b0e0ca36859fe891c7839b8ed06b928dd57a1f1e57620146bfb89ec13f

Observation 8de88ad9-579e-4f37-a296-3287fb0bcd45 · outbound

This paper cites Visual autoregressive modeling: Scalable image gen- eration via next-scale prediction.NeurIPS, 37:84839–84865,.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Visual autoregressive modeling: Scalable image gen- eration via next-scale prediction.NeurIPS, 37:84839–84865,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.802201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.518963Z digest=sha256:543de80e2d67b3ac2586faf2545626282f0d4b611221897f9af4841369495b4d

Observation 3454e542-3ad7-4b0c-bb35-6577416bc056 · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Plug-and-play diffusion features for text-driven image-to-image translation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.522030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.522030Z digest=sha256:11666a3156b8fead62d72ceec148ea31165e037b73faf20e57cb9f95b349d492

Observation 2b5c21ab-2d1d-4977-8d09-d1988a10df65 · outbound

This paper cites Taming Rectified Flow for Inversion and Editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Taming Rectified Flow for Inversion and Editing

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.525885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.525885Z digest=sha256:2ecee1accc156f5f22724a173fb3b85ec471f02d2527413879807b13e7cb7fa7

Observation 986bfa1b-284c-479d-846c-14e29a296e5b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.529340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.529340Z digest=sha256:456d0e0c4ff967e6ae11f72602c009a2ad716025fa5aba8820c405fd367f6739

Observation b3d1845f-16c1-4fdc-98c1-692f79eb023c · outbound

This paper cites Imagen editor and editbench: Advancing and evaluating text-guided image inpainting.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Imagen editor and editbench: Advancing and evaluating text-guided image inpainting

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.783609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.533590Z digest=sha256:2a8861fe162a38dbbb41252a0bfdbdb6017cb46ae3f1333b552d5fcacf9585c8

Observation a35bd1ac-47c3-4e41-91ea-16f59a137aa1 · outbound

This paper cites Bovik, H.R.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Bovik, H.R

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.772704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.537735Z digest=sha256:f51dff06d253292117312e77096b160945ce24e7179723e0286657971cfc489f

Observation 10261733-2f2b-4e5a-8d5e-29c77663f06b · outbound

This paper cites OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.540887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.540887Z digest=sha256:3502347f4234540212c1e5815bf07a154cb269bbbea60125a3fb55e2634a0c90

Observation ea8b8b9d-bfa7-4506-97ba-63d9b40ee474 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.544847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.544847Z digest=sha256:695ade4e5c0560f562d8811b935373a7371aaf3f561b235e7d5ebfff92c2cfea

Observation a97d97f2-d2c2-406a-851c-e3ada674ea4a · outbound

This paper cites Anyedit: Mastering unified high-quality image editing for any idea.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Anyedit: Mastering unified high-quality image editing for any idea

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.760524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.549242Z digest=sha256:5ab92da73a5fcf05fa788dbebbb0347aca09d5475cb0511513335cc8bff1cf59

Observation 59170675-8587-4e07-912f-c276f55ca17b · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.748728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.552155Z digest=sha256:e3bb265173a867d5a507d413af24cf9eb7c04d46c59187b4b08888c173cefee7

Observation 697a6fbe-9b2b-454d-be59-2af1f795eb53 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Adding conditional control to text-to-image diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.738437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.556503Z digest=sha256:1ef35e619d8aed590abdb8797db56673c007c9ef9003eed64b03ea3f65ca7472

Observation 405d20ab-3b90-4baf-9fdf-8bb96f082cc2 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent The unreasonable effectiveness of deep features as a perceptual metric

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.728070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.560703Z digest=sha256:415d757fb4263724d81a2ee43cd79276517c9d33079913c12e814b4d77385f42

Observation 3f34ab02-3de9-4167-aea1-5f5d7b4fcbd9 · outbound

This paper cites Ultraedit: Instruction-based fine-grained image editing at scale.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent Ultraedit: Instruction-based fine-grained image editing at scale

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:10:14.717571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:10:14.564510Z digest=sha256:dc7e8cbf49e7033a73f0fa3e62b38bfdb90e82ea6e9fefcc6cac29f6caac742f

Observation 98a3c2ee-c03f-4248-b490-3022c66f35d2 · outbound

This paper cites KV-Edit: Training-Free Image Editing for Precise Background Preservation.

Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent KV-Edit: Training-Free Image Editing for Precise Background Preservation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T15:10:14.568791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:10:14.568791Z digest=sha256:84f131538d7c0ff350f69c24d3659ab53b1b97e067fee2fcc8600c4575b66fac

Pith citing papers

No inbound Pith citation observations are available.