Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:26:37.951818Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2507.02713.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:26:37.951818Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:58:15.910837Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T08:55:35.719491Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 18a8a7e8-d58f-436a-829a-374dee2eb931 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Layer Normalization
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce664aa-8222-4bee-a01d-41e488039fde · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation YOLO-World: Real-Time Open-Vocabulary Object Detection
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a86b7fb-5a21-456b-b4c0-e8a9e72225e8 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Other works (Esser et al., 2024; Chen et al., 2023; OpenAI, 2024a; Gao et al., 2024; Li et al., 2024b; Xie et al., 2023; Nair et al.,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4459dfd4-ebc3-4b8e-a686-1c3c4491139d · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Orange kittenwith bright blue eyes on rocky ground
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c21c29c2-f788-422c-9333-00e92ba0796d · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Prompt-to-Prompt Image Editing with Cross Attention Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50d7046c-663c-4008-9ad5-29b874c4478b · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation RTMPose: Real-Time Multi-Person Pose Estimation based on MMPose
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01613cf8-1f4b-4777-9245-2199ba8b590b · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7215d8bd-8432-4be5-a886-136d3118ed11 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation FiT: Flexible Vision Transformer for Diffusion Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b242ca2-ef67-4c57-a9a8-d9dda2a33e65 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d997714a-58dc-400b-a68f-f60983debfc6 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Diffscaler: Enhancing the Generative Prowess of Diffusion Transformers
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e46863c8-3d73-4bb5-90f9-27a220622298 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34055887-880b-46fb-bc8a-0a98c08cd7dc · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c00fb91-bada-4d3b-8495-fda0d80c9b88 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Plug- and-play diffusion features for text-driven image-to- image translation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 546da1e9-c8f0-4b9f-8677-79a3345d2600 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation CogVLM: Visual Expert for Pretrained Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a97ff1bb-d775-4d86-b0e4-feeb561fb681 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation InstanceDiffusion: Instance-level Control for Image Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ef55050-4ddb-493d-af15-66ccca64d5a0 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation MagicPose4D: Crafting Articulated Models with Appearance and Motion Control
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 076047ce-e063-4916-ae4b-eb62f3a7ba30 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a3e5616-3fab-43f9-84b3-eaaf67b1dc47 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Dataset A.1
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5744c3c2-3163-4b75-af18-0695f3ca3986 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53db4ca1-4d6f-46f0-8937-f1de40e25cdb · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a793e732-afa4-499e-b52b-adaf3782c812 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation In this work, we use PIXART-α as our backbone, which is a variant of DiT (Peebles & Xie, 2023)
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d81187c-e584-48ed-bdca-80089b5f3f4c · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation During the sampling process, the model can transform Gaussian noise of normal distribution to real samples step-by-step
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8fb12fa2-5987-4aef-a44b-4a71474f2a56 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Latent Diffusion Model & PIXART-α
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a77f4d82-7611-4f2b-b617-79222886e324 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation At the inference stage, we can reconstruct the generated image through the decoder ˆx = D(ˆz)
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be50e2ba-d19a-4114-aa3d-92c320cc81b9 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b1cdd5-10a6-46bd-92b0-d2a793596c66 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation SyDog: A Synthetic Dog Dataset for Improved 2D Pose Estimation
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66b5de23-ce72-49b8-9087-f6d3de5844cc · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c6e5d5b-eb49-4408-9bf2-6a90c31c6534 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e9c2a99-12d5-4a3d-ba4f-936d9f692126 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation DiffiT: Diffusion Vision Transformers for Image Generation
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d69922-5558-45b9-be06-f30d5f5af92e · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 826e5b75-7a8c-4d47-9523-abc80896f559 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 507a892c-920f-4d4f-8c5d-79a184db2557 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff6a9c9-759d-4e83-9fbd-d7349d8a2396 · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Demystifying MMD GANs
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8e0e912-55f9-409c-8f16-dbe75a0f90cf · outbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation CosmicMan: A Text-to-Image Foundation Model for Humans
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19d9ffe3-fda1-4681-b56d-7aa338b5475f · inbound
TerraDiT-$\Omega$: Unified Spatial Control for Satellite Image Synthesis with Any Geospatial Primitive UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.