Pith. sign in

Paper Citation Record · LEDGER

Rethink Sparse Signals for Pose-guided Text-to-image Generation

As of 8 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2506.20983.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20983 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:40:23.917814Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact3
  • verified fuzzy44
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 833bc21f-dc56-4240-95b8-ce59b7836de8 · outbound

This paper cites Spatext: Spatio-textual representation for con- trollable image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Spatext: Spatio-textual representation for con- trollable image generation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.798550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:18.387547Z digest=sha256:5ef2540d321704ea4a7b5706c523e484a43a83b8b05aefd2ec3737d03126ea12

Observation 45547b5a-3975-438f-8a80-e554d2d9599b · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Rethink Sparse Signals for Pose-guided Text-to-image Generation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:18.447091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:18.447091Z digest=sha256:ab3bfad8df17c5fb9ba8e3e1b44b190189561433edcfe22bd99047d0c1a9b655

Observation 44218b58-17a2-47d5-b3f3-1ea80d3510a3 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

Rethink Sparse Signals for Pose-guided Text-to-image Generation ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:18.534281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:18.534281Z digest=sha256:e1c31c68720cc6e52cd604b62a6af39c66c0b5304ddd5d2d85db426680c932d5

Observation 5628c2ab-30af-4872-888c-e7383a3d72a1 · outbound

This paper cites Person image synthesis via de- noising diffusion model.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Person image synthesis via de- noising diffusion model

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.639476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:18.624771Z digest=sha256:2c10239e96f660f28c49fd83d872c23a21131237386b6ff9bfdf2c2f7d6f22ea

Observation 5ab2491b-4378-4387-9606-075d955b3969 · outbound

This paper cites Openpose: Realtime multi-person 2d pose estimation using part affinity fields.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Openpose: Realtime multi-person 2d pose estimation using part affinity fields

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.444330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:18.685457Z digest=sha256:0faeb4a976e9a888e1d1912ea6517bfc45a894900d9d4ad9a006f31cbefd4fcd

Observation f0ee8c9b-2a7e-4cf6-9751-6724cb5b10a3 · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffu- sion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffu- sion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.248922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:18.778458Z digest=sha256:c83027370fe3ab4bd652cff1107b161e4f47c856544b815d5c9e655204907ada

Observation 607853e3-52f3-44ea-b3c7-8e5ac184cf04 · outbound

This paper cites Up- gpt: Universal diffusion model for person image generation, editing and pose transfer.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Up- gpt: Universal diffusion model for person image generation, editing and pose transfer

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.064430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:18.884292Z digest=sha256:cce12c80d187b242b90e3eef1c3480418999ca14ddc868a9df1bd11be6e9094e

Observation fe2847af-3a21-4c59-a7b5-e3523d7f771d · outbound

This paper cites Openmmlab pose estimation tool- box and benchmark.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Openmmlab pose estimation tool- box and benchmark

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.812432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:18.965482Z digest=sha256:ac09456a03bd1088593c3eb5cb0d65d62fce69c586c9e6b6ee12c654c81aeea3

Observation c22768ff-8cb5-466d-a26d-0bd141b8e2cd · outbound

This paper cites Make-a-scene: Scene- based text-to-image generation with human priors.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Make-a-scene: Scene- based text-to-image generation with human priors

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:19.108485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:19.108485Z digest=sha256:a2224961abd8093211e8cea94601e3384c09774abd74b07d14d7b81baab55a8e

Observation 2c05867f-a099-4f8a-9a47-ee4101208c0a · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:19.223666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:19.223666Z digest=sha256:9bb46f48bf7f5bc15551d0621c126452f07160090fcb0fe3d70154c1c3c997c4

Observation acac9e1e-f3f4-4ff5-b258-00547f502898 · outbound

This paper cites Controllable person image synthesis with pose-constrained latent diffusion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controllable person image synthesis with pose-constrained latent diffusion

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.564562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:19.334652Z digest=sha256:f0069ab56fc03c9c987b4314c3db656d9dedb7ddedcfa77d64d70294de1aea51

Observation 7d530aa7-e0e8-4bda-bc64-6701f675d284 · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Prompt-to-prompt image editing with cross-attention control

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.391334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:19.479717Z digest=sha256:1b9806d6df09ae0bc58f9af474f405f7f73da086e6a019ccfe4e631a1e73d5e3

Observation e5d999f1-1c13-4f02-9527-39893fc807d6 · outbound

This paper cites Clipscore: A reference-free evaluation met- ric for image captioning.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Clipscore: A reference-free evaluation met- ric for image captioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:19.644611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:19.644611Z digest=sha256:e9ec615e31353cc23ef60a03a2e366b03ca6619115fc6f37b5429279c92edc6c

Observation a01b8f56-54a9-47fb-8b03-71f48e7721fc · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Animate anyone: Consistent and controllable image- to-video synthesis for character animation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.130883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:19.795100Z digest=sha256:6c5187ec8ec9fca3ab2604ca442487a5bc1eb4f2bf4263b45f9465c7f374f1e7

Observation 6a1bccc9-3dc9-4d47-93f4-1a8cb391bf44 · outbound

This paper cites Composer: Creative and controllable im- age synthesis with composable conditions.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Composer: Creative and controllable im- age synthesis with composable conditions

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.925992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:19.888999Z digest=sha256:1b19699961eb729857d90bab88e596c20d8b4ee1a4d40631ea1ed7f472a59fd1

Observation f941104b-625f-4ac8-ac21-f92395e651d6 · outbound

This paper cites Stable-pose: Leveraging transformers for pose-guided text-to-image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Stable-pose: Leveraging transformers for pose-guided text-to-image generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.703108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:19.965654Z digest=sha256:b3e26b6211a04985c19255c75f2c0b5e22ba6a950a310bd556238ed6eb944887

Observation 3cf98ae1-f372-48e2-b8d6-6706dcd805a4 · outbound

This paper cites SPAC-Net: Synthetic Pose-aware Animal ControlNet for Enhanced Pose Estimation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation SPAC-Net: Synthetic Pose-aware Animal ControlNet for Enhanced Pose Estimation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:40:24.462808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.049697Z digest=sha256:b9757f29770ec975a8e04a17bbfea94080b4d7d0269da125c8b51704f0bd761d

Observation af2d24ef-0310-4618-a707-be84a04873cf · outbound

This paper cites Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:40:24.274129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.131528Z digest=sha256:4cff750172f6dc62daa60e4a90519cba0a334ced9ebd19ecbe9a763da8f8e27e

Observation e090c3f2-84a6-4ede-b58e-7564d8eada9c · outbound

This paper cites Human-art: A versatile human-centric dataset bridg- ing natural and artificial scenes.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Human-art: A versatile human-centric dataset bridg- ing natural and artificial scenes

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.449937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.219216Z digest=sha256:72f4dde691c0ec6ff5046fa68bb906635ff2753669ca5b3c00c693f8a8d3b969

Observation 2d79e33b-121f-47f2-94b5-b25e00a34248 · outbound

This paper cites Humansd: A native skeleton-guided diffusion model for human image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Humansd: A native skeleton-guided diffusion model for human image generation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.256036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.293266Z digest=sha256:7f15bad492fe62cb549f5d074c37f397c2ebbc0a7d94db9e77ac3cc95432a335

Observation 25cb4b2a-f888-4c65-8186-54b21222ff5c · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Dreampose: Fashion image-to-video synthesis via stable diffusion

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.116455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.396502Z digest=sha256:68622f27cb563074a1e2aeb3c28e55b920cc5eb3f988421ffdfbb1958bbaa763

Observation 7184493b-22ca-44a7-b19d-82dd367fcff4 · outbound

This paper cites Segment anything in high quality.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Segment anything in high quality

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.972257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.512885Z digest=sha256:bc3c77a2157d0509f4223d2191f90d3f9abb1db166337c470ac345f740a47700

Observation 1ec9def2-159a-4df0-891f-0413aa5e4cb8 · outbound

This paper cites Dense text-to-image generation with attention modulation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Dense text-to-image generation with attention modulation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.804186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.595026Z digest=sha256:f68f4d39bd2eecc383cd8f6592d69d6213dd5f8521c90a70b0c7becfe6e0ae34

Observation 607a66ca-c9bd-42e4-9e61-6732cd769cce · outbound

This paper cites DisPose: Disentangling Pose Guidance for Controllable Human Image Animation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation DisPose: Disentangling Pose Guidance for Controllable Human Image Animation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:20.708669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:20.708669Z digest=sha256:7a13549c6d92ba4c8890fbeb301a354dedaeca160a8fd5d9f7dbcdadd357e286

Observation 783e2f72-2b44-4d2e-a2bd-49810a47c9c5 · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controlnet++: Improving conditional controls with efficient consistency feedback

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.666948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:20.793234Z digest=sha256:e6dd91da6db4e78115f82848d65ea42a4c8772f82d71db0f8c9c3ea8b7c7c1c3

Observation f69f1314-6a64-4010-a217-9626108da35b · outbound

This paper cites ECNet: Effective Controllable Text-to-Image Diffusion Models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation ECNet: Effective Controllable Text-to-Image Diffusion Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:20.850877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:20.850877Z digest=sha256:039419cd9ac662b8868dbecf2995f8565b1ccc2bdcd374b72f3de118f5b95552

Observation 52b2b2fb-567d-4835-a095-42c980c5b3f2 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Gligen: Open-set grounded text-to-image generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:20.931244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:20.931244Z digest=sha256:0d27be09d1885eeaba863354e146bd3f93d31d6fd8308970383386eff5e3213f

Observation 119d5b4a-1c81-4081-b23b-e238c38c97ce · outbound

This paper cites Controllable text-to-3d generation via surface-aligned gaus- sian splatting.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controllable text-to-3d generation via surface-aligned gaus- sian splatting

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.491203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.050057Z digest=sha256:9f161f5d868d75ad3bf8184d71665876c89f75185462013e53a17bff7d4d7435

Observation d38beb51-20af-492f-8711-9b700076ecec · outbound

This paper cites Ctrl-x: Controlling structure and appear- ance for text-to-image generation without guidance.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Ctrl-x: Controlling structure and appear- ance for text-to-image generation without guidance

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.354096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.171516Z digest=sha256:3395b7b1cd1039866c994ac3f10e4b3f9c752fb498063e9054a864b62aff9f22

Observation ee0095ed-0ed8-43f5-9242-aec745325c01 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:21.255648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:21.255648Z digest=sha256:f26e2e646f5ba4da3dc434d0ca73c2a690e7dbf71f946bd5334c73f7a46b3051

Observation 13543dd2-06d3-4ca0-92a3-4747a533ad52 · outbound

This paper cites Hyperhuman: Hyper-realistic human generation with latent structural diffusion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Hyperhuman: Hyper-realistic human generation with latent structural diffusion

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.191061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.369782Z digest=sha256:87f2bc814889e8d6d262109bfb8e47fb4f1c45a2d0228d9f9ea5cb01ffcd03a2

Observation f1dd4690-39e6-4a9c-9caa-9cf447c62710 · outbound

This paper cites Smartcontrol: Enhancing controlnet for handling rough visual conditions.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Smartcontrol: Enhancing controlnet for handling rough visual conditions

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.015061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.456151Z digest=sha256:d9c0b505040d5acf4d9ca25820a644db1641f105f1b95f400c7a8458508f679f

Observation 993a21a4-7d32-43e1-a379-0d171d8b410b · outbound

This paper cites Controllable person image synthesis with attribute-decomposed gan.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controllable person image synthesis with attribute-decomposed gan

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.843139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.568663Z digest=sha256:59502810a2ff9fc9fb8f2ea335ca55d616a6e9dcb1a1c5f8c519d62c20c415b2

Observation 91adaafd-c272-4938-8cf1-6dde34bff1de · outbound

This paper cites Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.645243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.644746Z digest=sha256:4a9df1a4e29f9316b8e0010a98563d7be779ff168bed75711579885b54d61060

Observation 808ad8c8-ba86-4a9d-a047-274bd831b626 · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.506372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.732972Z digest=sha256:cc085584867838b0407b398ba828fbc52ac8f15c3aa7e2bd4f3eb9274ca8f010

Observation 1d451bb1-a1f3-4a38-bf7f-618b1ac63fa6 · outbound

This paper cites Consolidating attention features for multi-view image editing.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Consolidating attention features for multi-view image editing

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.334312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:21.818579Z digest=sha256:633b42ffc97d4ae7df234534334fd12f52340b1ab66cbab902cc35dccfc401c6

Observation 9c39c07c-2e35-41cf-86bd-dd14baf73fc2 · outbound

This paper cites ControlNeXt: Powerful and Efficient Control for Image and Video Generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation ControlNeXt: Powerful and Efficient Control for Image and Video Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:21.916192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:21.916192Z digest=sha256:b4af3319197e4ac08267ea2929363c14a16fbcfc889199f11467e332e36bfa3f

Observation 0d88ba47-61e9-41bd-b7e9-1e56f189f0c0 · outbound

This paper cites Learn, imagine and create: Text-to-image generation from prior knowledge.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Learn, imagine and create: Text-to-image generation from prior knowledge

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.132901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.006626Z digest=sha256:4904855d45649ec6b200dcdd1fe0bec5b64460bb80ed1d3c9afac2d8b4765c87

Observation 05a64b06-8e81-48fc-85a5-3ca9b9b271ad · outbound

This paper cites Mirrorgan: Learning text-to-image generation by re- description.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Mirrorgan: Learning text-to-image generation by re- description

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.933999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.076431Z digest=sha256:6d058d030f3948e912429db27ec1e06d5f0902318b096cb9ae937b1d9054061e

Observation 72703871-4aa1-4f6a-8c2f-94d34b17d84a · outbound

This paper cites Unicontrol: A unified diffu- sion model for controllable visual generation in the wild.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Unicontrol: A unified diffu- sion model for controllable visual generation in the wild

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.738273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.160510Z digest=sha256:30c406e40426d6c13001325447c5d1a99cffedabc450f97c3a4e189d83f004ee

Observation 1a0eead5-185a-4469-b772-310b5493210b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation High-resolution image synthesis with latent diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.558803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.267932Z digest=sha256:e6edd837c78d8979575f862e073e52347ce016ccd08e362010875b03eb3f0217

Observation 83d0645b-7bff-4c3e-9d2a-5eccb02dc216 · outbound

This paper cites Benchmarking and error diagnosis in multi-instance pose estimation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Benchmarking and error diagnosis in multi-instance pose estimation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.374856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.354649Z digest=sha256:dbd909fcd8e5c3d37cbaf69e7bf499b51c6fee6bca419e9c9f374a7aa7fce766

Observation c87315b9-434f-461f-8cff-34b5f42a6f15 · outbound

This paper cites Stable Diffusion v1-5, 2022.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Stable Diffusion v1-5, 2022

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.151560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.437441Z digest=sha256:8547e9a61d95cc391f8c622130141d41330892c02679c28c509ac9efe9c4c0fd

Observation 86d5a380-274a-47dd-9cef-f2acde020253 · outbound

This paper cites Advancing pose-guided image synthesis with pro- gressive conditional diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Advancing pose-guided image synthesis with pro- gressive conditional diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.932845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.559546Z digest=sha256:41457f61ec6cb84ffac1db8e27166d6848915ae63018337fa6732f2359b5fb28

Observation 1445cae7-bdb8-4e1e-b56d-87fe324abe08 · outbound

This paper cites Deep high-resolution representation learning for human pose es- timation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Deep high-resolution representation learning for human pose es- timation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.729070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.670807Z digest=sha256:363e8f6fbf4a6ad98d8cdd029a551bbeab27b763c68a51058ad862808a91cb43

Observation 272e4a01-41a1-4d12-b231-ca0ffe1d8095 · outbound

This paper cites What the DAAM: Interpreting stable diffu- sion using cross attention.

Rethink Sparse Signals for Pose-guided Text-to-image Generation What the DAAM: Interpreting stable diffu- sion using cross attention

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.481776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.795614Z digest=sha256:a24a9ff01c64cd701f13b3cdd591bfdcfa32000c1ac3732df79833a175a65665

Observation 3af9e700-3e09-4412-bd5e-0140522cdec1 · outbound

This paper cites Visualizing data using t-sne.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Visualizing data using t-sne

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.260908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.889767Z digest=sha256:cb409d3c22f4f1045b508717bca9467e44cc402a0ee3d6703c750d5def882384

Observation b36ee586-7416-44ad-a2bf-8227b61bd322 · outbound

This paper cites AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment.

Rethink Sparse Signals for Pose-guided Text-to-image Generation AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:40:24.089793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:22.979387Z digest=sha256:8eca1a60b516f4b6c362b1c2f88fb91528ed2f53670fca13659e4043c092543b

Observation 8fcca519-5e3b-4860-9e55-37269ad6398a · outbound

This paper cites Vit- pose++: Vision transformer for generic body pose estima- tion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Vit- pose++: Vision transformer for generic body pose estima- tion

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.127713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.064502Z digest=sha256:149ebaf932ef5869ee2fd62955b774b2b7a0b2fb78d4eabfb64cf3a7438c8d57

Observation 34235a47-1d52-499f-9d9d-7f07ea86d56b · outbound

This paper cites Magicanimate: Temporally consistent human im- age animation using diffusion model.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Magicanimate: Temporally consistent human im- age animation using diffusion model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:23.153655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:23.153655Z digest=sha256:d1dc5fd9e8e4c48ea2fe8866fe9e17af3af052a1941a76b0611680f38c1e1253

Observation 6358fd66-da64-4fd1-876c-ee32e3ac8255 · outbound

This paper cites When controlnet meets inexplicit masks: A case study of controlnet on its contour- following ability.

Rethink Sparse Signals for Pose-guided Text-to-image Generation When controlnet meets inexplicit masks: A case study of controlnet on its contour- following ability

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.958791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.205566Z digest=sha256:b936c9dbc9027cd1b9c4fd7590db1720b2d522328aa04535518cd3e8265eb30f

Observation a06cd6da-07c0-4043-bfaf-f54265992b12 · outbound

This paper cites Grpose: Learning graph relations for human image generation with pose priors.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Grpose: Learning graph relations for human image generation with pose priors

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.776788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.282647Z digest=sha256:9fa5a26f6f82e13547dff9a683984721497c2fdbb6ccd133abb0fdf63504e192

Observation 1ac9b915-db30-4958-8005-df02b151b425 · outbound

This paper cites Ap-10k: A benchmark for animal pose estima- tion in the wild.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Ap-10k: A benchmark for animal pose estima- tion in the wild

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.616665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.382631Z digest=sha256:1028c2a872354979533b954f7ff33c6fd08a5c049503f76543c965d7779e78c3

Observation 9207dcfd-c897-4005-b695-294835838505 · outbound

This paper cites Pise: Person image synthesis and editing with decoupled gan.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Pise: Person image synthesis and editing with decoupled gan

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.458888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.483847Z digest=sha256:32b2b0ae49a8d7eddf36ce5c122109242af0bc9dfb63d6bc5e2d45c4a746f836

Observation 1ccaa88c-f7b5-463c-9e73-d35c2edba244 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Adding conditional control to text-to-image diffusion models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.283427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.564585Z digest=sha256:bbf8b4c87540d71e0dfef47b51b0d6920fe6855da73ad4a4247008357875a6db

Observation 9d89d071-f55c-404f-828f-2ad97bd5aa08 · outbound

This paper cites Uni-controlnet: All-in-one control to text-to-image diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Uni-controlnet: All-in-one control to text-to-image diffusion models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.105082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.662160Z digest=sha256:8980669e4bb71aa2409f35421162010ab31581192f6aef2030998605be9a3d5e

Observation 5208c7f9-7def-4072-8b94-6218f2966c44 · outbound

This paper cites Unipc: A unified predictor-corrector framework for fast sampling of diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Unipc: A unified predictor-corrector framework for fast sampling of diffusion models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:24.924657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.776671Z digest=sha256:929d953d6a8bfa836bd6521bb9ca47a4537e096d4e36f94542dfa4c1af6d3955

Observation 647fed9d-f1d4-4794-ab9e-ee636f9fe1ea · outbound

This paper cites Cross attention based style distribution for controllable person image synthesis.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Cross attention based style distribution for controllable person image synthesis

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:24.769862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.848578Z digest=sha256:b892396f88629b6eb570c6e9017a091f63118e67246446f9057561203ef71d3a

Observation 9475ba99-a852-441d-a906-d2c102ee2fdb · outbound

This paper cites Champ: Controllable and consistent human image an- imation with 3d parametric guidance.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Champ: Controllable and consistent human image an- imation with 3d parametric guidance

Reference 59

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T22:40:24.624338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:40:23.917814Z digest=sha256:b83e2eefc4495e60b0038097edbf4855e9b01b9cca0d68a4e06965982be1aa6f

Pith citing papers

No inbound Pith citation observations are available.