Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T14:58:37.383749Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 100 inbound Pith citation observations for arXiv:2405.08748.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T14:58:37.383749Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:31.264007Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
41 of 41 outbound references displayed
External citation measurements
1
pith, observed 2026-08-05T02:28:24.338817Z
Observation 0c574b84-5f3a-4d03-a150-467b23204e7d · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding https://www.midjourney.com/home
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 039c4d4d-be1b-4460-9f15-a82fceac5df1 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 40c7ac0e-f72c-4415-a366-1a4b84bb0193 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8593d81b-356a-4f68-bda6-e9926667a283 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding All are worth words: A vit backbone for diffusion models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6f64cbf0-2d0b-42a1-9e50-5cca60d37f1a · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Improving image generation with better captions
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 40350f7a-cdac-4ba3-a06b-6a626f54c1f9 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Muse: Text-to-image generation via masked generative transformers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 581e6b9a-8d09-4f9d-8b70-64f55df30ad4 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Pixart-\alpha: Fast training of diffusion transformer for photorealistic text-to-image synthesis
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d7ee4fbd-c76c-454f-919d-0c74c54748df · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0723abeb-487a-49fd-aebf-02d570f19eed · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding An image is worth 16x16 words: Transformers for image recognition at scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 06ea00d7-df55-4979-8d97-edf34b1c8b75 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 239459fe-7563-48fd-a297-eca4d5de87d5 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Making LLaMA SEE and Draw with SEED Tokenizer
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fc659712-5713-49dc-8a1e-11a9fa8ebe90 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Matryoshka diffusion models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a562c28b-18a9-4b2f-ae70-b6f40eb956c4 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Query-key normalization for transformers
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e738f32f-cb26-4a95-abd1-213b077cb13f · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Gans trained by a two time-scale update rule converge to a local nash equilibrium
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aab99592-907a-43ce-8e5b-57bfea5fa980 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding DialogGen: Multi-modal Interactive Dialogue System for Multi-turn Text-to-Image Generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c828b109-e8ce-4c3f-acb1-6a46c0f85c15 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 906aa91c-7c34-497a-9f21-d74968aeaabe · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ce80a498-3bb0-424a-997e-2e225d8a2be4 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Swinv2-imagen: Hierarchical vision transformer diffusion models for text-to-image generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cf3631ec-65a9-4c21-8fa0-8c3f833a8f41 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Microsoft coco: Common objects in context
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0357996c-8a73-4404-a89a-9b856e3d90b5 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Improved Baselines with Visual Instruction Tuning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 49898400-25a7-4fab-a029-f035f9f43716 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Instaflow: One step is enough for high-quality diffusion-based text-to-image generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f1f8c7b2-f33b-44f0-8ca5-ec49c773b450 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 92525b43-496a-4e6a-b0fb-3ad6866b5bd9 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Scalable diffusion models with transformers
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4b421b8b-cce4-43ef-9a75-aa9814a74e6c · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Sdxl: Improving latent diffusion models for high-resolution image synthesis
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e044e6b6-5b34-4e22-a1ee-91824dfd3ba5 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Learning transferable visual models from natural language supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ff7379f4-2eff-4f58-baf3-6277d9280564 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0817e3f4-34ad-4a5f-874d-fb924a198d12 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Zero: Memory optimizations toward training trillion parameter models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6f26e00a-fc3d-4085-9615-d1c07f0055b0 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding High-resolution image synthesis with latent diffusion models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e2deabf2-87a2-4bd1-bc43-1c41295c135b · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Photorealistic text-to-image diffusion models with deep language understanding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e12772ff-067d-4b1e-b59f-3caea0ed7bf1 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Progressive distillation for fast sampling of diffusion models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 40df8d60-f4ca-468b-b68b-8d4529eddc54 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Adversarial Diffusion Distillation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 496b6f5b-db62-4586-8d14-79b01dd9ae0f · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Roformer: Enhanced transformer with rotary position embedding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cf757cd1-032d-4489-9283-539769e0af20 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Attention is all you need
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9c83a454-109f-4355-af2b-7413fc309d7d · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding PAI-Diffusion: Constructing and Serving a Family of Open Chinese Diffusion Models for Text-to-image Synthesis on the Cloud
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 082d8ffd-e82a-4384-9e7f-0bc07055dbf8 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding NExT-GPT: Any-to-Any Multimodal LLM
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 17dcc32c-0090-4f4f-b188-e71e1e898e35 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Taiyi-Diffusion-XL: Advancing Bilingual Text-to-Image Generation with Large Vision-Language Model Support
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e03bf071-645a-47ad-8781-513f76be546c · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding UFOGen: You Forward Once Large Scale Text-to-Image Generation via Diffusion GANs
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d2c356f3-c195-426d-8230-ba8200344824 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5bfd222e-5971-46f7-b3aa-f0b6e370dbaa · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding Altdiffusion: A multilingual text-to-image diffusion model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b7dc518b-cf5d-429f-8d5c-aa5a690bb261 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding One-step Diffusion with Distribution Matching Distillation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a7276b04-4dc2-40a0-9297-790122fa7487 · outbound
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding CapsFusion: Rethinking Image-Text Data at Scale
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e95e4235-bd61-4cc4-98c2-271aa5e4a753 · inbound
Emu3: Next-Token Prediction is All You Need Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b784b988-ed8e-4877-8bf6-3f1c26089a6a · inbound
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 981345a0-6025-41c6-a7b9-3e4e260114b9 · inbound
High-Resolution Image Synthesis via Next-Token Prediction Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35a26a3b-db9c-4b16-a921-46dc957d7ccf · inbound
Text-to-Image Synthesis: A Decade Survey Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3153ac62-9931-449c-a407-586072d518d3 · inbound
One Diffusion to Generate Them All Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0358a844-7806-4255-9a7c-cc027c1d2bf3 · inbound
Efficient Multi-modal Large Language Models via Visual Token Grouping Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41b9b3fb-433a-4e58-a9b7-1e3ab7cd8412 · inbound
Open-Sora Plan: Open-Source Large Video Generation Model Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f1b82155-c6c3-4532-b83a-08dd778c6193 · inbound
HoloDrive: Holistic 2D-3D Multi-Modal Street Scene Generation for Autonomous Driving Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7765ed2f-9fb4-482c-a343-69def11b1db7 · inbound
Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d696d16-2c28-4696-ba80-897ca68eead4 · inbound
SnapGen: Taming High-Resolution Text-to-Image Models for Mobile Devices with Efficient Architectures and Training Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29af3adc-1aeb-4752-b769-eb686e7c6d92 · inbound
Zigzag Diffusion Sampling: Diffusion Models Can Self-Improve via Self-Reflection Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb023b8d-cfb7-4994-8ec8-83b8fd4eff4d · inbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbb84cfd-0e08-491d-b4e3-b9c6eb9be24d · inbound
F-Bench: Rethinking Human Preference Evaluation Metrics for Benchmarking Face Generation, Customization, and Restoration Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8803063-afa4-44ac-98f4-725ef1f37e45 · inbound
CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9abed91e-9e08-447c-94ad-77d2fa0eafaf · inbound
PromptLA: Towards Integrity Verification of Black-box Text-to-Image Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6decca13-abc0-427c-a26b-d644163b295b · inbound
EvalMuse-40K: A Reliable and Fine-Grained Benchmark with Comprehensive Human Annotations for Text-to-Image Generation Model Evaluation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5a4f1d5-8e79-4e79-bb7a-dcb878fd2cb5 · inbound
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f407df20-bd64-4035-920a-d02445d15108 · inbound
SeedVR: Seeding Infinity in Diffusion Transformer Towards Generic Video Restoration Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 007b7b0d-6cae-4683-9d25-75725bc7c622 · inbound
Enhancing Image Generation Fidelity via Progressive Prompts Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82d4b05c-a5fa-4a76-8362-e9ebb2c09444 · inbound
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4287647a-e996-43af-9b67-db4af3afc96b · inbound
T2ISafety: Benchmark for Assessing Fairness, Toxicity, and Privacy in Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abb4556a-db91-4677-9123-e5c6d201c538 · inbound
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b77b361-f091-4e10-b15f-9c6f63666604 · inbound
Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 613fc111-a701-4619-92a3-5df42363d9a0 · inbound
SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bc4c100-3750-4c4e-a460-f2075022732a · inbound
UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea13ae8c-cd82-4109-969a-4b4c3f4fa44f · inbound
VFX Creator: Animated Visual Effect Generation with Controllable Diffusion Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f728abfe-1ac2-47eb-b4a0-b36b74f59ed7 · inbound
Matrix3D: Large Photogrammetry Model All-in-One Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93a71d0d-a75f-49d1-ac8b-be50586c74c0 · inbound
Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3ea81b16-5537-4f3f-8adb-e6216c3153de · inbound
RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 76dcd2c9-854b-48f0-8874-9b2212109e86 · inbound
VACE: All-in-One Video Creation and Editing Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 619bf7e2-d1fb-4938-8c63-eea9a23e6b4e · inbound
Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7ac6b9e6-7cf5-4f49-9e6d-cb4c928b463f · inbound
Instruction-augmented Multimodal Alignment for Image-Text and Element Matching Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc9f41c5-f5f1-47ab-b673-6f68df039d55 · inbound
Understanding Attention Mechanism in Video Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f56009-acaa-4905-9dd4-7fb6034a54fe · inbound
SUDO: Enhancing Text-to-Image Diffusion Models with Self-Supervised Direct Preference Optimization Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dea6047-9081-4053-ae64-92d6faf54234 · inbound
From Reflection to Perfection: Scaling Inference-Time Optimization for Text-to-Image Diffusion Models via Reflection Tuning Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ccc9090-6d8b-4dcb-bc33-982fb70a9314 · inbound
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e11c36e8-7e6e-4681-8621-271b32ffcb65 · inbound
LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cd9816c-ceae-437d-a5f7-7a33db577638 · inbound
Hunyuan-Game: Industrial-grade Intelligent Game Creation Model Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bbb5788-a59c-4612-a250-14e6e2bce774 · inbound
UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25088bc4-a1f9-4f6b-90a1-acd6d0d13a56 · inbound
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a1a8f78-555b-489e-b3db-a88b273631c5 · inbound
OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7464fcef-16b5-4120-bebc-5a2d6fff5501 · inbound
MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c39549-f37a-4bdc-b515-ceb6be0a7050 · inbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 586e70c8-b0c4-4609-b585-1b9f4cbb586b · inbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 130
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfd03ea0-5c34-4716-b360-59d3162780ea · inbound
AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2077f18a-effd-48eb-be04-092b0ee7208c · inbound
OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ba3e372-9816-449e-8acf-d4a831eb2ab6 · inbound
GenSpace: Benchmarking Spatially-Aware Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90fc9349-56ef-4074-ae6e-9e5335767cba · inbound
NoiseAR: AutoRegressing Initial Noise Prior for Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33884d4d-d15d-40be-8aad-eb28daf44fed · inbound
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce85b1c3-0fda-4b7a-b908-926aba5f9c5f · inbound
Marrying Autoregressive Transformer and Diffusion with Multi-Reference Autoregression Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4158dac0-cec8-4cd3-bf99-534b47953d49 · inbound
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b75c20ac-9734-4f78-8c5f-350876d716c1 · inbound
COME: Adding Scene-Centric Forecasting Control to Occupancy World Model Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba124ae-8510-4cb4-890d-62a2b53796e0 · inbound
ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9305dfa-6c32-4c91-ab1e-c02c83693a8d · inbound
OmniGen2: Towards Instruction-Aligned Multimodal Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ca526121-fa56-4e51-8f05-80d22f967955 · inbound
Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a179ed4e-2cdc-4e92-a040-54e4580ed319 · inbound
Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83f93069-964f-43f1-bc48-b538e56401bb · inbound
Med-Art: Diffusion Transformer for 2D Medical Text-to-Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 887d158b-933f-4a91-9be9-639b82c323dd · inbound
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1011fda0-9765-4d35-9bc8-af0f014b05e4 · inbound
On the Feasibility of Poisoning Text-to-Image AI Models via Adversarial Mislabeling Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01613cf8-1f4b-4777-9245-2199ba8b590b · inbound
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e175b62-cf9c-4935-acea-0f58edf05a1d · inbound
CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 131fe098-44f5-4301-a768-3fb169eca97f · inbound
CharaConsist: Fine-Grained Consistent Character Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc134c95-a660-44d6-be19-d5fa8236b61e · inbound
AnimeColor: Reference-based Animation Colorization with Diffusion Transformers Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f39972e5-3f49-4445-ae4a-8fa12728a4a4 · inbound
T2I-Copilot: A Training-Free Multi-Agent Text-to-Image System for Enhanced Prompt Interpretation and Interactive Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a12d2be-2083-4d49-a18a-e96e4b528ae4 · inbound
ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a96f96-3294-407b-ac8c-dd8960487d9b · inbound
A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7854012-5878-4145-a2f3-f0b127f40fdd · inbound
Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf68c54a-d28b-4b4e-887a-decbab4b14fc · inbound
Transition Models: Rethinking the Generative Learning Objective Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ca1c93a-5adc-46ee-935f-5a453af8e7e3 · inbound
Home-made Diffusion Model from Scratch to Hatch Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aea1fa6-9b79-49d5-af67-7cdfd5a622c4 · inbound
Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preference Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0356551c-13d7-480c-8ac5-f129a797c09f · inbound
FLUX-Reason-6M & PRISM-Bench: A Million-Scale Text-to-Image Reasoning Dataset and Comprehensive Benchmark Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f70f34df-e7cb-4993-9ec6-9e54491e60ec · inbound
Compute Only 16 Tokens in One Timestep: Accelerating Diffusion Transformers with Cluster-Driven Feature Caching Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4560414f-06f5-4d1a-aada-954181a4b0ea · inbound
HunyuanImage 3.0 Technical Report Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d4482470-c6fe-4ea1-985e-0cb3fa272087 · inbound
HunyuanImage 3.0 Technical Report Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37ccd311-6343-46c1-921a-b5e96caee0ab · inbound
Few-Shot Synthetic Image Attribution: Identifying Unseen Generators with Limited Samples Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59025a7d-7553-453e-b225-0d57ba994af1 · inbound
Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 001e6a19-564b-40af-aef8-eb00406fd66b · inbound
Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8596c85-a782-40fc-af9d-409f28cb433c · inbound
PixelDiT: Pixel Diffusion Transformers for Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5a541b64-c887-40fd-9858-c92028ea12a4 · inbound
Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 31537b14-4724-42bc-9266-d28b5e4755c2 · inbound
Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0bbf535-9216-4c07-8dea-6120b2872110 · inbound
Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86bdbf9d-1baa-4439-b1af-bf37f8db312f · inbound
Guiding Token-Sparse Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caba5c44-c215-4525-95ee-5dcf5f2a0fab · inbound
TexTailor: Inference-Time Textual Guidance Tailoring for Multimodal Diffusion Transformers Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21edd1ef-12eb-47a5-9ea5-5c5afc398638 · inbound
The Algorithmic Gaze of Image Quality Assessment: An Audit and Trace Ethnography of the LAION-Aesthetics Predictor Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 59d9d0a1-507c-4d17-a774-c773e8bfe673 · inbound
Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2345c93a-837e-4460-9f5c-7cc8a2a04a41 · inbound
OmniFysics: Towards Physical Intelligence Evolution via Omni-Modal Signal Processing and Network Optimization Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c127b862-4de1-4325-8945-a6b1fdd850c9 · inbound
When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e1e45613-ed44-4f95-b3ec-b929758b557b · inbound
CineAGI: Character-Consistent Movie Creation through LLM-Orchestrated Multi-Modal Generation and Cross-Scene Integration Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6ff9371d-2cf4-4076-a9de-44dc210e2828 · inbound
Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 25a3dff3-6d9e-4f6a-9200-69e4c2d47ff8 · inbound
ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ae66cecb-b76a-418f-85d1-e2da5783fb5d · inbound
Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9717243b-4342-4726-a454-7c5b71be061e · inbound
Leveraging Verifier-Based Reinforcement Learning in Image Editing Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c6b71185-2754-4f18-9e0b-4ec7af0773b4 · inbound
Leveraging Verifier-Based Reinforcement Learning in Image Editing Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6e4765f9-5092-4f08-b420-40ba1ec391cd · inbound
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 48430f21-1021-4096-b0ec-2ca5c20ff89a · inbound
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b4877599-743e-4b84-b4c6-fa145bfe051b · inbound
Qwen-Image-2.0 Technical Report Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9d5a4e95-03ff-4132-8a32-06f8e89223cb · inbound
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 732da0c5-a6ad-416c-b3d1-10d99fddbe1c · inbound
Asymmetric Flow Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f164c029-7467-4a9b-bf47-e53501ccf378 · inbound
Asymmetric Flow Models Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1a4b0297-079b-4eb6-b530-1bcf3220cd70 · inbound
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.