Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:10:27.446268Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 100 of 109 outbound references and 23 inbound Pith citation observations for arXiv:2411.17440.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:10:27.446268Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:29:36.698279Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:49:42.422443Z
100 of 109 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c2d0e0ec-1a23-492c-8683-b8ff4ccc6bc8 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition BoT-SORT: Robust Associations Multi-Pedestrian Tracking
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b01816a-a9b4-48ed-ae0f-b1d93d77f301 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Vatt: Transformers for multimodal self-supervised learning from raw video, audio and text
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11e9c834-ac41-47e1-8ccf-46e76d5e3859 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9304edf-a19f-4be8-820f-f7c9781cd1de · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Improving vision transformers by revis- iting high-frequency components
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 721cd4b0-9dc0-4a75-b00f-88288aec8286 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition All are worth words: A vit backbone for diffusion models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 723be1c0-4869-42cf-b611-e11687c59717 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a8391a1-cc4b-442d-8d99-b9af9bfdf7a0 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11854595-7c91-4e08-ad11-48ff98a0570d · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Video generation models as world simulators
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 214b5091-34aa-4cc0-ba00-20fbd4b754ff · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Vggface2: A dataset for recognising faces across pose and age
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 368d66d3-70ec-4b0e-94dc-8116dbc690f9 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Still-Moving: Customized Video Generation without Customized Video Data
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab41c5f0-1efa-4138-bcda-1b37fcc5db25 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7e0707c-a7ca-47d7-b2a9-4e3f20cf528f · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Boosting Camera Motion Control for Video Diffusion Transformers
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 460b06dc-c729-45d4-8c39-1d27f9fda29c · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Arcface: Additive angular margin loss for deep face recognition
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03475cd6-06a5-4412-9d8d-b84c305963f5 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition DH-FaceVid-1K: A Large-Scale High-Quality Dataset for Face Video Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cc5c8f9-4f50-4a02-8604-fe8a5b1b0761 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1125e7cb-1f32-4d58-93b7-8edfa8158e43 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Animatediff: Animate your personalized text-to- image diffusion models without specific tuning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59868d19-d236-4225-8ac2-125bb57ce12f · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition PuLID: Pure and Lightning ID Customization via Contrastive Alignment
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f6a9c2f-3607-4bbb-884b-6f68dddb49f8 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f99e9b-69c8-4810-add0-a2c7cb8feec5 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be4cbf3f-6460-4e4a-adb1-66951930f25c · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Imagine yourself: Tuning-Free Personalized Image Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a70c6db-af35-4834-8628-962d20f2c769 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition CLIPScore: A Reference-free Evaluation Metric for Image Captioning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2783bbce-ae3c-435c-984a-91b68ef028c0 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a468ab-0a64-4879-87c5-31f6fa45c1b4 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Classifier-Free Diffusion Guidance
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25908296-956f-4e5c-8630-52d1eb7c6ca6 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Denoising diffu- sion probabilistic models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6792fc7b-cdc7-4f79-9c4b-185c7a95bd47 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition LoRA: Low-Rank Adaptation of Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd01a0ac-281e-49f5-8a20-4de8d2b3a8df · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Deep networks with stochastic depth
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc4ae9da-bfed-4094-9cc6-5c987530d06a · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Curricularface: adaptive curriculum learning loss for deep face recognition
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a846910f-7ff8-462a-81c9-8e89fcf3e053 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Real-time intermediate flow estimation for video frame interpolation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5eca017-db05-4e46-b140-46f5b1eb1afd · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Vbench: Comprehensive bench- mark suite for video generative models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 945f2c14-a021-4d75-87f5-534a92c910b4 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition ultralytics/yolov5: v7
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05eb1106-bb5d-4eef-afac-c7e0e1578847 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Progressive Growing of GANs for Improved Quality, Stability, and Variation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c833d34-2396-422a-867b-bdc66102ef7d · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition VideoPoet: A Large Language Model for Zero-Shot Video Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc08e78-ac1f-4a44-80aa-51f7a4b8e724 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Multi-concept customiza- tion of text-to-image diffusion
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e3dea8-633a-45ec-b833-4c73096f853f · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 890166e4-fbda-49d3-9fbd-dd41597cbecf · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 800e02b5-ad1a-43f1-a255-64ef9138bbc5 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Photomaker: Customizing re- alistic human photos via stacked id embedding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c4de337-d5d3-42da-84c1-f8de135de424 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Open-Sora Plan: Open-Source Large Video Generation Model
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bb877a6-12e6-4fb8-adc0-6608a723ee47 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf7357a8-f1aa-40ab-a956-1856d518fad4 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition EvalCrafter: Benchmarking and Evaluating Large Video Generation Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b366be98-ae72-4d63-a80d-79ee0e587352 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Fetv: A bench- mark for fine-grained evaluation of open-domain text-to- video generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75dafaa2-adbb-40e2-99f4-60bb54e30105 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Latte: Latent Diffusion Transformer for Video Generation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2e3827-9be0-4234-8c31-dafce71c4619 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition MagicStick: Controllable Video Editing via Control Handle Transformations
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3afb0d6b-c836-4e11-8c7c-2aaffa160b5c · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Follow your pose: Pose- guided text-to-video generation using pose-free videos
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1922c170-7ae4-43c6-8d6e-b242ca9438cd · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b647d22-0fa6-42ca-a3ec-4c40d4d21e84 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Follow-your-emoji: Fine-controllable and expressive freestyle portrait animation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5629f967-a2e3-4960-a258-fb87152d2e3f · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Magic-Me: Identity-Specific Video Customized Diffusion
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b7768e7-5bfd-4036-b0ee-afd2f5b75a39 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ae029cf-1607-405d-8b4f-4fdc1f92bf6b · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition V oxceleb: Large-scale speaker verification in the wild
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 718e76bb-b221-449d-8500-17b13c27d0c5 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Toward verifiable and reproducible human evaluation for text-to-image generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8c1dfa9-d079-4cef-a2ae-9a467e3b30fd · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Movie Gen: A Cast of Media Foundation Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a193d71-9f60-47f1-b3e3-c1d72e60e2f4 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Learn- ing transferable visual models from natural language super- vision
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9894d1b5-045d-45ee-891e-cdf0d97035ce · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Zero-shot text-to-image generation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcc49a71-8317-4efa-ab22-343620b348e4 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Rethinking Video Deblurring with Wavelet-Aware Dynamic Transformer and Diffusion Model
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6725d220-8b9a-4408-9c47-801ea0ba63bc · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition SAM 2: Segment Anything in Images and Videos
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc71d83e-778c-4940-8d7a-4b12d41690e5 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8be134f-7a51-49aa-a531-c8802c49aba9 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition High-resolution image syn- thesis with latent diffusion models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d01fc48-4e47-4083-bcb7-f49b72fd6535 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b6a115-2731-4bb7-bb90-2e93616ca0bc · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34962758-460c-4aca-9b3a-85791136ef97 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25b6ff77-ff17-42ac-9c63-8e89a8484903 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition LightningDrag: Lightning Fast and Accurate Drag-based Image Editing Emerging from Videos
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bce8180-8e55-4f97-939d-eb0df03d68b0 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Freeu: Free lunch in diffusion u-net
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afaac3b4-2a48-4ae1-a679-e05d2ca7f8d5 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Deep unsupervised learning using nonequilibrium thermodynamics
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 355ca527-c795-4339-b96b-d2ba52084856 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Denois- ing diffusion implicit models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2f75b500-4eea-4385-9a84-891b45ab226a · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Rethinking the inception ar- chitecture for computer vision
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8e132efe-df6d-4f35-bdbf-155665106af7 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Cycle3D: High-quality and Consistent Image-to-3D Generation via Generation-Reconstruction Cycle
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b53f5f2-5606-4b45-a4f0-ed911523ecbe · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93931778-6e20-42af-836a-81b311793404 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3da74bf-4d3b-4a39-a803-3c5307d73654 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition InstantID: Zero-shot Identity-Preserving Generation in Seconds
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0fc3ef0-fe18-46c4-8089-0723dfd9201e · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition One-shot free-view neural talking-head synthesis for video conferenc- ing
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 86c2a531-d16f-4958-a020-f60add71ee80 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Customvideo: Customizing text- to-video generation with multiple subjects
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ed1ce0c-614c-48c8-a84c-66a17cd16a0c · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Elite: Encoding visual con- cepts into textual embeddings for customized text-to-image generation
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cf792dd2-35c9-4846-aa7a-9f6720b453fe · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Dreamvideo: Composing your dream videos with customized subject and motion
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 07aa1095-693d-4394-87ec-0dc9f0d5d9a8 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition MotionBooth: Motion-Aware Customized Text-to-Video Generation
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff9151ad-01c8-4f3c-b5d4-055ac1b25727 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Dynamicrafter: Animating open-domain images with video diffusion priors
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3b0035df-556a-42d0-a748-2c8e81ab871e · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Autoregressive Models in Vision: A Survey
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee086857-eb1b-4928-8bc0-849e57061636 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Easyanimate: A high-performance long video generation method based on transformer architecture
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b29d11b-317b-45ad-ae07-3e8fc19ab70e · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition xgen-mm (blip-3): A family of open large multimodal models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7815a96d-65db-4063-870a-45442c7e6602 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Ucf: Uncovering common features for generalizable deep- fake detection
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d2d24c38-733d-4568-8187-8cde5a5a31a9 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Transcending forgery specificity with latent space augmentation for generalizable deepfake detection
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 77912a41-63ef-4572-902f-85246471f1d4 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition DF40: Toward Next-Generation Deepfake Detection
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd4b482a-eed3-44f6-8e4d-fa20f8210918 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d4ac17-7192-40df-8b76-e794b3aa3036 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Is Parameter Collision Hindering Continual Learning in LLMs?
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 292c531f-d710-4a65-8307-20a55105d023 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45ab19fa-3548-4a39-963c-788094c1d185 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e9dd844-0b8f-45ec-a99b-a3eee5efe5f7 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Celebv-text: A large-scale facial text-video dataset
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c96b841a-4d5b-4f3c-996e-6d86321146f1 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition EvaGaussians: Event Stream Assisted Gaussian Splatting from Blurry Images
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 138b37f9-405a-4511-953b-4bf4be403990 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bde9494-b3ab-4ea8-8af4-e0215d332ee6 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition PromptFix: You Prompt and We Fix the Photo
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c999eb4-805d-497a-a8bd-d6268108168b · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f0041b0-430c-4dc4-86f2-86272fb14e4d · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Chronomagic-bench: A benchmark for metamorphic evaluation of text-to-time-lapse video gen- eration
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1815ef45-f2c1-48f7-a3de-dafa8659d9e8 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Adding conditional control to text-to-image diffusion models
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24417da9-9ee4-43bd-b3ea-f64ef3ee9f8c · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Bytetrack: Multi-object tracking by associating every detection box
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation de62994f-ac24-4243-a4a4-89378a38a8ff · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Flow-guided one-shot talking face generation with a high- resolution audio-visual dataset
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3e25e70c-5698-4d2d-98e1-d3f937668bb6 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Tora: Trajectory-oriented Diffusion Transformer for Video Generation
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 309f281a-43bf-4e63-9485-20a473ebb7cd · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Videogen-of-thought: A collab- orative framework for multi-shot video generation
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44332149-83a2-4357-a81b-d6ffbd579a35 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Open-sora: Democratizing efficient video production for all
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f647344a-a72b-4487-9aaa-b17a7f8bb391 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Allegro: Open the Black Box of Commercial-Level Video Generation Model
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7c612bd-a6f2-4b27-858b-33b58e9784c8 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Celebv- hq: A large-scale video facial attributes dataset
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f07f4994-ab72-4c30-9e73-6f2532861887 · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Comparison with Closed-source Method
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c1cbe94f-536e-4e77-8ece-ca07c01e635e · outbound
Identity-Preserving Text-to-Video Generation by Frequency Decomposition Visualization of Different Injection Methods 6 3.2
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d379e5df-9720-4c55-a0bf-0781afdd4dbf · inbound
PersonalVideo: High ID-Fidelity Video Customization without Dynamic and Semantic Degradation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ce45a66-28a6-4abc-9a71-afef92ef4e3d · inbound
WF-VAE: Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Model Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94c1d87a-70f9-4ab6-ae4d-14ec73a6f39a · inbound
Ingredients: Blending Custom Photos with Video Diffusion Transformers Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 654c9a00-8f6d-4039-9593-bbe9cf27e3e1 · inbound
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 157cdf74-9c98-48e7-9520-9f6312c91447 · inbound
Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 29436255-9733-4d60-a3ad-87f0001dae02 · inbound
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0f306ac-9e16-44a7-9521-3256679dafe1 · inbound
ImgEdit: A Unified Image Editing Dataset and Benchmark Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ab3fa46f-dd86-49fc-afe4-7e1a6d1947b2 · inbound
OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 113
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46a1ad32-721b-48e2-a3d0-d42d1fac1549 · inbound
UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aaadd1b3-ad36-4399-88df-ec3aae29cbb2 · inbound
UNIC: Unified In-Context Video Editing Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fef7acec-22d0-4247-bda7-769ee6ae4dbd · inbound
Follow-Your-Creation: Empowering 4D Creation through Video Inpainting Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3761732-b039-47c0-a26f-1e59f3b149d6 · inbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 241d36e5-0297-4afd-97b0-cd103d00fd43 · inbound
DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75be7455-bf0e-4288-ad58-2f692377874b · inbound
SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9a45fe12-994c-4630-b9ce-94e3317f18ec · inbound
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 368fc191-8da6-4ce5-9ed0-67e9b7be3445 · inbound
A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4434e8-b090-4277-b226-97d0e06fddcc · inbound
Evolution of Video Generative Foundations Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 196
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0975307d-bb4c-460d-ae3a-0e738b08c3fc · inbound
Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 609aa3f1-5b4b-42f1-a784-2f903f394879 · inbound
MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation beaaa26e-a4a8-4af6-88e3-1f09932a983f · inbound
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 22a7e0d9-e635-464f-ba6d-a50f53319213 · inbound
Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5a9005f6-bf64-40aa-a5a7-0e02b26ee519 · inbound
A Comprehensive Ecosystem for Open-Domain Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2819a73a-0c3c-428b-8e6b-a5cfa6e6ec2d · inbound
Customizing Video Portraits via Identity-ActionDecoupling Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.