Pith. sign in

Paper Citation Record · LEDGER

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion

As of 21 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 2 inbound Pith citation observations for arXiv:2412.14462.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14462 v2

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:16:57.637400Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:09:51.868938Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T09:16:00.656580Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact2
  • verified fuzzy22
  • unresolved29
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6f155b2-5a3f-4372-a191-26a9b628d6de · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.430298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.430298Z digest=sha256:216f81a3f3c61a2bfa6fd0792cf7f80e2b07139bc5f8f62731abc1aa63ae0d6e

Observation 69499289-9e4c-4fa7-af4b-c45114a0b52d · outbound

This paper cites Image inpainting.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Image inpainting

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.135309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.435128Z digest=sha256:32345d952dc3f524c0bdcd3f35d0d887847b648f4f42ea110f2eb8307cff0d4b

Observation 11afce0d-b5c1-4610-868f-a69773afd9bd · outbound

This paper cites The ecological approach to visual per- ception, 1980.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion The ecological approach to visual per- ception, 1980

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.124811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.439653Z digest=sha256:3c85cc3147521d403cc76611ffd6d4e05b0a498f91e28602f4126961f64e3dc1

Observation 3c106799-0f45-454d-9144-529e7fb71b28 · outbound

This paper cites Hallucinating pose- compatible scenes.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Hallucinating pose- compatible scenes

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.113782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.443563Z digest=sha256:bc65103eeb3e805b5e53257e6ce93c401769549ef61f7dffcff854cf7cc72c50

Observation e4537298-bf6b-41f0-a47b-487955b738b6 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion In- structpix2pix: Learning to follow image editing instructions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.102684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.448232Z digest=sha256:f210e8bfcf180e827489cc8377216c7336505113ca0529b2f20b9041cf121042

Observation 2ccadfcf-cea8-4557-b598-c376e9e7ea92 · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.452699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.452699Z digest=sha256:a8a6a94714011f5aacd44a90b835d1ce4837ae917c538318adb79909c8cef649

Observation cb1ca916-cdfc-4c9b-855a-997d0ed25d09 · outbound

This paper cites Scene semantics from long-term observation of people.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Scene semantics from long-term observation of people

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.092059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.457233Z digest=sha256:a61f0180d5092a4ae9fc161a9f17a6fe7d646b9e671a058cba7b6c942e161113

Observation de74c44e-c779-45b1-a2b0-c8e3d9efbb10 · outbound

This paper cites Diffusion models beat gans on image synthesis.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Diffusion models beat gans on image synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.461342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.461342Z digest=sha256:3d8c935d753db6930d968f72e788950b7a4070aeed016ab9a1923aa856e637db

Observation b1aa3854-d4f5-4a41-aa8b-ec09b5d921bd · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.464744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.464744Z digest=sha256:9969f9a3cd2961660d58d8ba6abbc75c214e41aece9af92b80b70cb52d1b1446

Observation 3c77f26e-41ef-4274-b05c-bcc0e3718eca · outbound

This paper cites People watching: Hu- man actions as a cue for single view geometry.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion People watching: Hu- man actions as a cue for single view geometry

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.075469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.468138Z digest=sha256:9ea2e4dfbe1b450297df4b07c98944588aa72bb3b6ef62bcab740a0f130c317c

Observation 2f4b7fe3-06ed-49d4-a465-ecda590609a2 · outbound

This paper cites In Defense of the Direct Perception of Affordances.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion In Defense of the Direct Perception of Affordances

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:16:57.816767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.472041Z digest=sha256:e4a64af96031aec84f2d713f3122091f1069a043833b25266bdb2b9c0b3b78a1

Observation 54ec3fec-b875-4fe5-82d4-87d6135075a0 · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.476563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.476563Z digest=sha256:51a18bc6c15021115f1ed91724314715e62d0698695c74452cc922f99795b9ca

Observation 7c713bcf-1ec5-490a-bd8a-3a163d16ac0e · outbound

This paper cites Detecting and recognizing human-object interactions.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Detecting and recognizing human-object interactions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.480994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.480994Z digest=sha256:75bba0783ffa046af11223591beee4121f7d5306cb8860e67007294f5c303ea9

Observation 90045db6-2852-46ed-bc4c-b65bf134c1a2 · outbound

This paper cites Generative adversarial nets.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Generative adversarial nets

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.484492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.484492Z digest=sha256:35e1c84e578ef3eb221e6b4e0ebcc667638123f7f041705d5f9354c017401c57

Observation 0e11d079-48f4-4082-a92b-3860a96c3751 · outbound

This paper cites From 3d scene geometry to human workspace.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion From 3d scene geometry to human workspace

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.053801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.488097Z digest=sha256:e9fe933f6d63337c94d437efb548bd6cd8634ca96d01c53891febc07cc3ed03b

Observation daf0d248-7263-4c58-af48-58b859b3b0d5 · outbound

This paper cites Deep residual learning for image recognition.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Deep residual learning for image recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.491751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.491751Z digest=sha256:d4aaf60a37fcacabed03535e69c5d43af491309e0d753af0f70180e99c763236

Observation 24649402-5757-4334-b93b-f09c4d7a27f1 · outbound

This paper cites Classifier-Free Diffusion Guidance.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Classifier-Free Diffusion Guidance

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.495732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.495732Z digest=sha256:9fe425537f97011d23501d83a874116801ae44e1764d8118b963b9d27e067e6a

Observation d53be710-0841-46a9-b365-ee505b71046a · outbound

This paper cites Video dif- fusion models.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Video dif- fusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.499696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.499696Z digest=sha256:31e5c35ad38cfdd6932a3221d4dc0c834efcd6ecb0b02fb7db613402f281f311

Observation f568c950-5746-49fd-bae2-e219e63ecd7b · outbound

This paper cites Zero-shot Generation of Coherent Storybook from Plain Text Story using Diffusion Models.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Zero-shot Generation of Coherent Storybook from Plain Text Story using Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.503300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.503300Z digest=sha256:850084ecbb520c220815f16898406915d4c2c1331af11b76ea1490a3b03d724c

Observation 81da03ba-3216-4cf8-89db-5451544752cf · outbound

This paper cites Efficient-3DiM: Learning a Generalizable Single-image Novel-view Synthesizer in One Day.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Efficient-3DiM: Learning a Generalizable Single-image Novel-view Synthesizer in One Day

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:16:57.774377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.507235Z digest=sha256:fb3d252ea0774edce315981bfbf6fb36d053e0e6090d65902551592f33494d2a

Observation af4858b5-5442-437b-a4e9-1d97d6311811 · outbound

This paper cites Training generative adver- sarial networks with limited data.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Training generative adver- sarial networks with limited data

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.032619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.511352Z digest=sha256:d9320e2108d7e0a721b46dc3691aa5a5001179d05db02e24e1025c13907fcc00

Observation 7a3699f7-dd5b-4fb8-9a7f-f6c523913fda · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Imagic: Text-based real image editing with diffusion models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.022232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.514743Z digest=sha256:6d7e77b7aa33a300ec47a213876944f769510c31f44b4be8b11874ec043380d2

Observation 5793a165-ea26-4aa6-a172-26b157d1e200 · outbound

This paper cites Segment Anything.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Segment Anything

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.518338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.518338Z digest=sha256:f77cb93d58e6a71cbf956c33231a400bd1d66e8135f7be83ece30423d044d584

Observation 1938e02b-9a48-467f-bd65-3ebc9bdc9a87 · outbound

This paper cites Putting people in their place: Affordance-aware hu- man insertion into scenes.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Putting people in their place: Affordance-aware hu- man insertion into scenes

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:58.010822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.522271Z digest=sha256:ec1b8d246f5f72ce1a7a007e6b2b18a16b5ba5278ee47f4f1510c1717080ec03

Observation cbb16619-855a-40ac-ab30-d06b7c209923 · outbound

This paper cites The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.999176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.525868Z digest=sha256:439b5fd2c4baf9c2fd34cba5cefbddd0916b8232f20877ab6f55796d6c87cfb6

Observation 724e1f8b-aa45-48dc-b1c8-1244646201f8 · outbound

This paper cites MoEController: Instruction-based Arbitrary Image Manipulation with Mixture-of-Expert Controllers.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion MoEController: Instruction-based Arbitrary Image Manipulation with Mixture-of-Expert Controllers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.529289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.529289Z digest=sha256:b5cfda40e02aa344e5c21cd70cb41b9294bb794c6d27508d519c1d2db000bfb0

Observation 1b1e7e11-7ec0-4261-8c64-7c10bfc5b11d · outbound

This paper cites DreamEdit: Subject-driven Image Editing.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion DreamEdit: Subject-driven Image Editing

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.533394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.533394Z digest=sha256:05ed0acce4d666f5afffac63a711fa537cb95709794e27657cf3789199588093

Observation 1b45a55b-5ca3-4822-b35d-860ac47487e8 · outbound

This paper cites Mat: Mask-aware transformer for large hole im- age inpainting.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Mat: Mask-aware transformer for large hole im- age inpainting

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.988849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.537558Z digest=sha256:8383366fa5701e3230c3e3e607b6b7169598cec7c63e5fc047e1433cd05e0d1b

Observation e4597ee9-d80c-4a20-a429-d8985652ee6b · outbound

This paper cites Putting humans in a scene: Learning affordance in 3d indoor environments.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Putting humans in a scene: Learning affordance in 3d indoor environments

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.979005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.541489Z digest=sha256:8946d7d05662d9d83e76ab728ed9c09c4ee9c7e1cda0bceb989c73f6561ff6b7

Observation 27d9d004-8ee1-44a2-be51-7b54e10a89b2 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Gligen: Open-set grounded text-to-image generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.968521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.546594Z digest=sha256:7617a48d44c7653afc88d3e6675eb7c20f92da0e43afb529c4ff85ff9c18014a

Observation 3e5f0a40-b753-4cf6-8918-a10326c91612 · outbound

This paper cites Microsoft coco: Common objects in context.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Microsoft coco: Common objects in context

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.550299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.550299Z digest=sha256:7ac3fc287f0b9582ca9718a2f02a50666c6a9560cf6bab9fb0fd75a8f3380eda

Observation 19f77ef3-2fda-453d-aaaf-40606bec44f8 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.554033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.554033Z digest=sha256:228813f3c7c11ae9c72ad3c117051f03c00ec859946e82a31b90ba3554b7d9f4

Observation 5c78e3ed-eec0-446c-a746-74f51f6ca9d7 · outbound

This paper cites HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.558054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.558054Z digest=sha256:cd3458ff1baf2168a34350b157c46f8a713b41267886ad4f044aa2c17d7b1a93

Observation 9575db9a-e909-45c9-a599-a66f1e09ee18 · outbound

This paper cites DreamCom: Finetuning Text-guided Inpainting Model for Image Composition.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion DreamCom: Finetuning Text-guided Inpainting Model for Image Composition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.561951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.561951Z digest=sha256:f7d32702e62353d63fea29674a14dd2236372b09ee7cd92d23759402d454bf79

Observation 711171c3-57d6-4566-b55f-76d3fcc245ab · outbound

This paper cites Tf-icon: Diffusion-based training-free cross-domain image composi- tion.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Tf-icon: Diffusion-based training-free cross-domain image composi- tion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.951685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.565498Z digest=sha256:2ca4f0981cd96b9dfc92c316da508997078bdffa266480af8cb6f05d1d3688fe

Observation f9f36ab1-a217-466f-a7a5-d138890029bd · outbound

This paper cites Repaint: Inpainting using denoising diffusion probabilistic models.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Repaint: Inpainting using denoising diffusion probabilistic models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.940319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.569133Z digest=sha256:a1f49a8226439b2f2d77871ef3ddaaa1e7e39d89bf7bda0376ac3c8668dc1a99

Observation 1123e9b7-5323-4d63-bc82-75e845c6f7cd · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion DINOv2: Learning Robust Visual Features without Supervision

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.572697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.572697Z digest=sha256:336586f071413614ada7ac86241de617c189864944d776c81b77720e2e8c50d2

Observation 2794e99b-b81b-4641-95d7-2f4696196672 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.576598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.576598Z digest=sha256:ce2ac1036346dafcd00599efeb4a57d95446adb0fb9d2f9c7f2b2e7df5eddfcb

Observation 5437708b-5d8d-46b3-842b-4d2524b852bb · outbound

This paper cites Language models are unsuper- vised multitask learners.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Language models are unsuper- vised multitask learners

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.580419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.580419Z digest=sha256:2b36f4059b70e663dada583a3c3d2e39d125cdc514267e1b96c860afe85e9106

Observation 8de9a654-bfe1-49d6-8fea-37fa1e00300d · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Learning transferable visual models from natural language supervi- sion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.583740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.583740Z digest=sha256:058686b96d80641666009c974f9ccc7293c1fcf139429774a4d884969e5e6e7b

Observation 4261f1ea-a305-4545-a09b-8c0f13f30e37 · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.587455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.587455Z digest=sha256:4ce35f67ad94733141566df423046fde19214cdf64553dda44dea460094e5ab1

Observation e70e19e7-3bec-4ff6-89c9-b52b24b5f9ea · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion High-resolution image synthesis with latent diffusion models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.918207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.591669Z digest=sha256:52280e5c2b928f5898143eea20a30a60ac256e107f9a31c0e56b0f54603640a4

Observation 4a704ad3-4265-460a-acf3-007d443ea6a8 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.595631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.595631Z digest=sha256:5e37faef3fff6dcac040b967e0236086cebb6d3e38f05b2751d622d9954a9558

Observation 4a84b050-dd72-46fc-b707-436477fbfe55 · outbound

This paper cites Object- stitch: Object compositing with diffusion model.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Object- stitch: Object compositing with diffusion model

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.903069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.599224Z digest=sha256:14b8ed75cf95bc9ea0b61240710741d3ef4638fe599bad4dcb98ae66e29c80ef

Observation ca1e956c-a428-41ce-bd6f-43dc89a24c6a · outbound

This paper cites Resolution-robust Large Mask Inpainting with Fourier Convolutions.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Resolution-robust Large Mask Inpainting with Fourier Convolutions

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.602584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.602584Z digest=sha256:a8d51b960a21bf1440105df4922483b4cadf1189e1527818a01c48e75d6685ef

Observation 6390207f-7bd5-48d3-b739-2838958cd737 · outbound

This paper cites Effective Data Augmentation With Diffusion Models.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Effective Data Augmentation With Diffusion Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.606991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.606991Z digest=sha256:e00be6e47388341c43153898f4d5564a41380651f77d74059bf841211265bfd2

Observation 11ae331e-40b8-459f-b86b-5639a6656028 · outbound

This paper cites Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.893103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.610733Z digest=sha256:72365fc2267860861805a36b6f4b1c8f2304c84d8d17801f567ec4ab6ac8b4e7

Observation 3fbe74ec-20aa-436d-989f-0c7d519b9b16 · outbound

This paper cites Binge watching: Scaling affordance learning from sitcoms.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Binge watching: Scaling affordance learning from sitcoms

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.883228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.614170Z digest=sha256:8705b0cde45478a688a293797260cdc7c830ac55ccc56e2c760bb9761f219fa3

Observation 593565af-8529-4905-b12a-ff6a34365a57 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion mod- els.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Paint by example: Exemplar-based image editing with diffusion mod- els

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.617991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.617991Z digest=sha256:2fde0df8eb918a88eb6e4626ac8769f715164cda9fa30252d73f58145c45b04e

Observation 17abc852-6991-4da8-905f-6d5900b0dd1d · outbound

This paper cites Imagebrush: Learning visual in-context instructions for exemplar-based image ma- nipulation.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Imagebrush: Learning visual in-context instructions for exemplar-based image ma- nipulation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.868196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.621768Z digest=sha256:62ad4d77d0da7408e166ecd698fea69c8e19e4063474fed46f5a7622a97bb115

Observation 18bd51b3-0262-4445-bd25-738c324c46bf · outbound

This paper cites Modeling mutual context of object and human pose in human-object interaction activi- ties.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Modeling mutual context of object and human pose in human-object interaction activi- ties

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:16:57.857699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.625305Z digest=sha256:ee322a18e343227acad53da9c23e4b34fb6e05401d93b15d706e1794d4924953

Observation dc86a260-da2a-4a44-813f-f5df13dc58a4 · outbound

This paper cites Generative image inpainting with con- textual attention.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Generative image inpainting with con- textual attention

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.628663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.628663Z digest=sha256:029d41095ef3e7a2bddef2235086ead56a871b1e774eac23326ea5a642df20fb

Observation b9932150-13f1-4ade-9c99-7582157eca52 · outbound

This paper cites Recognize Anything: A Strong Image Tagging Model.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Recognize Anything: A Strong Image Tagging Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:16:57.633472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:16:57.633472Z digest=sha256:b6d71f4b56da90f96b286f4689f7c73875fcd7e906814b9e3eefd46409cbc699

Observation 28cf6b85-4bf1-4abf-b272-e2acd732bc39 · outbound

This paper cites Reasoning about object affordances in a knowledge base representation.

Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion Reasoning about object affordances in a knowledge base representation

Reference 54

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T12:16:57.842011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T12:16:57.637400Z digest=sha256:918a6e68330b0f81dae8352d5d21cf5286e7cd05f0711800d8bb51f673d63adb

Pith citing papers

Observation 8bfc216a-2152-4c65-9926-8e6e7bf65499 · inbound

HOComp: Interaction-Aware Human-Object Composition cites this paper.

HOComp: Interaction-Aware Human-Object Composition Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:09:51.868938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:09:51.868938Z digest=sha256:203cd8d18cbc9313322fa5d4a145d269ab0d2b9298dab338b972ad3881f57dab

Observation 4db632e8-457e-4131-8670-d0650658c1f5 · inbound

HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement cites this paper.

HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:00.659353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T16:10:16.251424Z digest=sha256:e928b123f4cdeccf90273fbb81c4a27f3a5144d5094145730f25312e33bf10f0