Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T17:30:48.007551Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2501.12173.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T17:30:48.007551Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d490b31e-3d0d-4c43-b68c-25bcb4026d10 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Spatext: Spatio-textual representation for con- trollable image generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation df8c0798-d017-47ee-be16-10ae77b31b18 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba481bdf-bd9b-4bab-8d27-4b19380350ed · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Sutherland, Michael Arbel, and Arthur Gretton
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e15659c6-0fe5-466e-a797-2181377d1e4c · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions InstructPix2Pix: Learning to Follow Image Editing Instructions
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7e13581-a7e8-4a81-a5d6-46a2e4255ddd · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b51779b0-dd86-4f76-8c11-dd7cd6fe8f80 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Training-free layout control with cross-attention guidance
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64544040-90d9-42c6-b024-18658ba570df · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions AnyDoor: Zero-shot Object-level Image Customization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84e5c1ab-d65e-4699-993f-ed758ebd7b9d · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions LayoutDiffuse: Adapting Foundational Diffusion Models for Layout-to-Image Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74322a34-e077-4956-a9bb-5c5d8e95b8ae · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions KPE: Keypoint Pose Encoding for Transformer-based Image Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7463dbd3-605e-4864-920a-c7f8f865d4b9 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Viton-hd: High-resolution virtual try-on via misalignment-aware normalization
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a75b5cc1-77df-4446-a201-24c77098d514 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Measures of the amount of ecologic association between species
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f4ce4cb4-362e-4645-b07b-e210759a69ec · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Stylegan-human: A data-centric odyssey of human genera- tion
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c72b606b-1c4d-4089-bd16-dd1af06003e9 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba458290-f62a-4d56-9c3a-441c41d42976 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions A versatile benchmark for de- tection, pose estimation, segmentation and re-identification of clothing images
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6536aca5-b052-495a-9982-daf7022e9f41 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Context- aware layout to image generation with enhanced object ap- pearance
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 75ae43e7-c555-42b4-af0a-fe35d4fb7642 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Denoising dif- fusion probabilistic models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0be945d-b4ad-488d-954a-f4da9bd4ba9b · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions CogVLM2: Visual Language Models for Image and Video Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66cc599d-d0ce-48dd-bff7-158d4bc5a32d · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions LoRA: Low-Rank Adaptation of Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d177a0fb-e45b-4063-aacc-13e6ea135bf0 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions ´Etude comparative de la distribution florale dans une portion des alpes et des jura
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ee5ccada-a3ec-471c-9e24-3d1f12f62f34 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Text2human: Text-driven controllable human image generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fd99474c-d18a-49c5-8832-fbd29d4872e3 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Dense text-to-image generation with attention modulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e86d74a4-fa59-48f2-adf9-23d908bed295 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Auto-Encoding Variational Bayes
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c50343be-f86d-46d3-9e95-fb61a07c58c5 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Segment Anything
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a33e0e4-b2cd-445b-ad39-14aa23fdc43c · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Multi-concept customization of text-to-image diffusion
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b658aa63-f2b8-4934-8adc-0ad7dbfcf72c · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Self- correction for human parsing
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94595758-8b25-4f86-8660-f8c8f87ae733 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Gligen: Open-set grounded text-to-image generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 71814155-4120-4460-bfe9-62f5873f8a01 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Image synthesis from layout with locality- aware mask adaption
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d292a7d4-5b9c-433c-ab03-2970a6ca3eb8 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1438a259-7aa5-4b42-8dcb-6d800e6bf84b · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Dress code: High- resolution multi-category virtual try-on, 2022
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f07dde4c-a53b-47c4-832f-dc8412757e0f · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions $\lambda$-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c705960-f287-4452-90e0-032760beeb40 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Learning transferable visual models from natural language supervi- sion
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50d1761e-a282-46e0-bc46-e26e004913e6 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions High-resolution image syn- thesis with latent diffusion models, 2021
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f680884f-6307-44e3-80dc-cb74615c614b · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30885056-b095-4de1-8e94-778a40c7062f · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Humangan: A generative model of hu- man images
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f95a8c05-5756-442a-8ad7-55e739b573ce · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions pytorch-fid: FID Score for PyTorch
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f523453-f07d-4eea-bfd7-c5e5ae280c68 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions In- stantbooth: Personalized text-to-image generation without test-time finetuning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cc25dca7-73fc-4be4-82c6-f8d7fcdf2ad4 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Image synthesis from reconfig- urable layout and style
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation df2be0fc-bf8b-4631-9efb-cdf5184e4130 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Object-centric image genera- tion from layouts
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 236f6967-d908-4754-a255-3195ece747eb · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Interactive image synthesis with panoptic layout generation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 99752318-c0cb-4658-902b-360198564b6e · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions InstantID: Zero-shot Identity-Preserving Generation in Seconds
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 485acd33-d9bb-4334-87ce-05158caaa609 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Instancediffusion: Instance-level control for image generation, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a2f20f46-d721-4072-aa23-3776efa3ea22 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Image quality assessment: from error visibility to structural similarity
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 43066a6f-3058-4fa9-8fab-dc8d4c3bc853 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 414dc079-dcff-402a-b380-7eb015485c60 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Fastcomposer: Tuning-free multi- subject image generation with localized attention
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 04342a8f-1d53-4922-b7e2-4683bc520249 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions R&B: Region and Boundary Aware Zero-shot Grounded Text-to-image Generation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f84a73-4883-4025-8cd0-92a06168b2fc · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 78a9c7b3-b863-49bf-9550-172e70599253 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Reco: Region-controlled text-to-image genera- tion
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 038fce44-f971-4687-a31e-755c0091c412 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5f886bb-68c3-462f-b0e4-f26e54b30344 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Customnet: Zero-shot object customization with variable-viewpoints in text-to-image dif- fusion models, 2023
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ba63e5ea-e2e2-4041-96e5-31f3b0664d9b · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions HumanDiffusion: a Coarse-to-Fine Alignment Diffusion Framework for Controllable Text-Driven Person Image Generation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 784dc3f6-544e-41cb-874d-bae8c2f5aef7 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions The unreasonable effectiveness of deep features as a perceptual metric
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 544697bd-b9d8-430b-a76c-8fa90abb9735 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Layoutdiffusion: Controllable diffu- sion model for layout-to-image generation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 70dcb08b-9447-4c06-9930-be3b2f0b5f9b · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions clip-score: CLIP Score for Py- Torch
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f573d113-9743-4574-a9d8-2b343db30785 · outbound
ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions Migc: Multi-instance generation controller for text-to-image synthesis, 2024
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
No inbound Pith citation observations are available.