Pith. sign in

Paper Citation Record · LEDGER

TransPixeler: Advancing Text-to-Video Generation with Transparency

As of 11 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2501.03006.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.03006 v2

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:03:13.651954Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 802c79ed-11f4-4261-89c1-f9c40e8215a3 · outbound

This paper cites One transformer fits all distributions in multi-modal diffu- sion at scale.

TransPixeler: Advancing Text-to-Video Generation with Transparency One transformer fits all distributions in multi-modal diffu- sion at scale

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.388236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.424433Z digest=sha256:ba95e6ab319b525cefebbcf4469c7ce1dd45c81b75c61688b953c17858297d48

Observation b7601d13-fbb2-4539-8cc5-fd1a613a5774 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

TransPixeler: Advancing Text-to-Video Generation with Transparency Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.428470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.428470Z digest=sha256:20a807b6c279bd920efcfaeb877061125a681f52636fe4ab9584496e5ec16635

Observation 629939d6-0a1d-4c3e-8059-093ec09859ee · outbound

This paper cites Video generation models as world simulators.

TransPixeler: Advancing Text-to-Video Generation with Transparency Video generation models as world simulators

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.432127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.432127Z digest=sha256:bfa602cf333ed01f7b5beb537c867ef7bdec0e48b7a1f2406b81b43fc6fa6fa2

Observation 79175ade-57a9-4ab9-8336-9af812f98ec1 · outbound

This paper cites Magick: A large-scale captioned dataset from matting generated images using chroma keying.

TransPixeler: Advancing Text-to-Video Generation with Transparency Magick: A large-scale captioned dataset from matting generated images using chroma keying

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.371380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.435878Z digest=sha256:474a0e1fbc5963437578ac204d676ecfd87f36e3dad47a22c50efb45514bfd92

Observation 44460e55-b00b-4dc2-9814-7c50ce0f1c7e · outbound

This paper cites zeroscope v2.

TransPixeler: Advancing Text-to-Video Generation with Transparency zeroscope v2

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.361772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.439450Z digest=sha256:12dacf91c578513a419c428b3eccd60fa88677898dc072add05253009f2d319e

Observation 53addafb-f405-455f-9d7c-2be7e37aeacd · outbound

This paper cites PP-Matting: High-Accuracy Natural Image Matting.

TransPixeler: Advancing Text-to-Video Generation with Transparency PP-Matting: High-Accuracy Natural Image Matting

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.443451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.443451Z digest=sha256:1fcad44121d93d6ff490d074c26e805300912637b3af0ecc68c6cbe1129adea9

Observation 034e862b-e019-4f0a-88d4-7a587affd2c1 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.447836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.447836Z digest=sha256:c674d359c91eb214dfcdf8de891d099541f24e5877aedd50a814687a8b75cde0

Observation 7dd7c79c-d318-43fa-941a-28b93962ab1d · outbound

This paper cites VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models.

TransPixeler: Advancing Text-to-Video Generation with Transparency VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.452815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.452815Z digest=sha256:a2fa5700f0232d22f5689e47c336da4161f3c21470ca7c829cff18fd731047c4

Observation 58a51d4e-9345-4b65-b35a-6255df38eaac · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

TransPixeler: Advancing Text-to-Video Generation with Transparency PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.457302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.457302Z digest=sha256:2daa422731230067e696be9338eb5942838a36f4fb70efa8b2f17fa451ece955

Observation a30c5492-528c-4b5d-aaac-d095a551f953 · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

TransPixeler: Advancing Text-to-Video Generation with Transparency ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.462730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.462730Z digest=sha256:9bf0599ceefcb70a3fcbf2757ae5ff81c75ff5f38a7de38fb21a87bb255000ce

Observation 6fceeca7-5cc1-4d81-8ac3-98cba82c8e3d · outbound

This paper cites Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning.

TransPixeler: Advancing Text-to-Video Generation with Transparency Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.466003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.466003Z digest=sha256:bd564851946b0b54d1c461b8ee8b7af841c0ff5cf993b2a52f6aea4d5a61b2f9

Observation 167a6211-dec9-4903-9357-17bd340a63dc · outbound

This paper cites Longnet: Scaling transformers to 1,000,000,000 tokens.

TransPixeler: Advancing Text-to-Video Generation with Transparency Longnet: Scaling transformers to 1,000,000,000 tokens

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.351711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.468903Z digest=sha256:7090eabd3591798753f3083ff8fd4b371b59da1581d9003a60c853cd832c3ac3

Observation 6e77cf48-e743-40b7-9195-4595c0c5f3c0 · outbound

This paper cites Two-frame motion estimation based on polynomial expansion.

TransPixeler: Advancing Text-to-Video Generation with Transparency Two-frame motion estimation based on polynomial expansion

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.341940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.475167Z digest=sha256:a7bc1ce6208608aa2601067f58016fa0c3d55580a59fed04181f3c332527ff9b

Observation 628f1105-f5b1-436c-81ff-6567d904b1cc · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

TransPixeler: Advancing Text-to-Video Generation with Transparency TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.477931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.477931Z digest=sha256:95d0ba332a75239d798ad4311a215cc4a46edca66e7db172243e7807ef5165db

Observation fd197d53-291d-47a6-82d0-6fe9a6de6704 · outbound

This paper cites LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control.

TransPixeler: Advancing Text-to-Video Generation with Transparency LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.481255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.481255Z digest=sha256:d4fbcdbd217626409d6241de48332896f7ea00b0d6880295e7c74cb4756a7f0b

Observation 14579c03-51cc-479a-80ad-a5641f5ab87d · outbound

This paper cites I2V-Adapter: A General Image-to-Video Adapter for Diffusion Models.

TransPixeler: Advancing Text-to-Video Generation with Transparency I2V-Adapter: A General Image-to-Video Adapter for Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.484792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.484792Z digest=sha256:4a78fc2f08a07558bc600c8224551ae79aa26ec891197abd11a2520c778618ba

Observation 88b1f8a7-6b39-4b6c-aed8-6c7887d3e3ff · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

TransPixeler: Advancing Text-to-Video Generation with Transparency AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.488150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.488150Z digest=sha256:4fe0b31b89be9b4073e292e35f767630fe13be337bb9770052b2e846dcd5cf88

Observation c3e8f7b8-b204-4140-832b-f58a943943ca · outbound

This paper cites Lu- cidfusion: Generating 3d gaussians with arbitrary unposed images, 2024.

TransPixeler: Advancing Text-to-Video Generation with Transparency Lu- cidfusion: Generating 3d gaussians with arbitrary unposed images, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.332283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.491488Z digest=sha256:a0fd83b7cf74fb602684675bde0b1fa18de4601c528b750d8a22a10f56e31d28

Observation 7da24728-5fa5-4ac4-bb6c-4ed251ddbe54 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.494418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.494418Z digest=sha256:e0e9d9077e6bec272b357899a888f4ec595b525acbf3b2c5b4588148c01d2167

Observation 3b0a4cdd-7e3a-4945-852c-cf785ee4056f · outbound

This paper cites Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction.

TransPixeler: Advancing Text-to-Video Generation with Transparency Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.497600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.497600Z digest=sha256:c35393932bf55897f76530db942ed3c2b001bd88497024904e1103f548040fea

Observation 36563d41-0e49-4420-b8b2-1753fc1c840b · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.500825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.500825Z digest=sha256:6c24682a008330735203bfb1216ea4650923629461a678bb361a674cc62b7ef2

Observation e818f96d-3bcd-4db9-95e1-3fe1ecaaf7f7 · outbound

This paper cites Denoising dif- fusion probabilistic models.

TransPixeler: Advancing Text-to-Video Generation with Transparency Denoising dif- fusion probabilistic models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.503727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.503727Z digest=sha256:db2a4e045924b37fd1047ce24f8d297300b67de7cb0ff886a2f0358aea517ad3

Observation a02183b4-8321-4586-9b97-7f256fc0ae9f · outbound

This paper cites Determining opti- cal flow.

TransPixeler: Advancing Text-to-Video Generation with Transparency Determining opti- cal flow

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.315439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.506571Z digest=sha256:e682775ea036434d819a3e6b00e590bbd4d0ee2de1aa39b616bc60e62fe9346b

Observation 452ce017-7d37-4721-be4e-4459bfacf3a1 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

TransPixeler: Advancing Text-to-Video Generation with Transparency LoRA: Low-Rank Adaptation of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.509331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.509331Z digest=sha256:9f224aa557ad6a0c8ef652decf00fc92de15ebe0041ea73de2f75cfb31f458dc

Observation 38dd361e-257e-4160-adad-5bf8905f4452 · outbound

This paper cites DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing.

TransPixeler: Advancing Text-to-Video Generation with Transparency DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.512595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.512595Z digest=sha256:214c5fbc49c85367445f45f604b7d13f4c9bf375b5ebc3b723a3c4d8bfde0fa9

Observation 038d4fee-4f6c-4919-8394-f41e2f340969 · outbound

This paper cites Repurpos- ing diffusion-based image generators for monocular depth estimation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Repurpos- ing diffusion-based image generators for monocular depth estimation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.515644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.515644Z digest=sha256:70a16ec2bec9c8276404b4a32fc10d2e8b17d68d80d02dfca5024a8ac1191a17

Observation f188966a-a10f-48d0-b405-4de32e1ee50a · outbound

This paper cites Open-sora-plan, 2024.

TransPixeler: Advancing Text-to-Video Generation with Transparency Open-sora-plan, 2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.298555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.518388Z digest=sha256:435f117927883a54507a25754f7f777cd3a11f1af91b2d0436eba247c8937988

Observation 5327516c-8149-4c04-9c1a-a909ba28ef4b · outbound

This paper cites Matting anything.

TransPixeler: Advancing Text-to-Video Generation with Transparency Matting anything

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.287705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.521329Z digest=sha256:c97f1bc7a91386f83d4aa4f16f54126596761a23bb04141db68c545c7444c29d

Observation 45d6a804-9eaf-44b9-9a22-2a894a107d0a · outbound

This paper cites Omnimat- terf: Robust omnimatte with 3d background modeling.

TransPixeler: Advancing Text-to-Video Generation with Transparency Omnimat- terf: Robust omnimatte with 3d background modeling

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.274367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.524446Z digest=sha256:76b10fc2f1ee0e65fb31345cecdd67374ed52ea681c5062c9ddbde4aae67b54f

Observation 69ad9895-e05f-4f2c-91a8-11653903f275 · outbound

This paper cites Real-time high-resolution background matting.

TransPixeler: Advancing Text-to-Video Generation with Transparency Real-time high-resolution background matting

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.262900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.527761Z digest=sha256:b312ef12d8e0117bb84c0c28492f67a33b33688d4c886e61c5157c873c16e7d8

Observation edb28094-cde5-4fbe-98ba-ef51ca91fd3b · outbound

This paper cites Robust high-resolution video matting with tempo- ral guidance.

TransPixeler: Advancing Text-to-Video Generation with Transparency Robust high-resolution video matting with tempo- ral guidance

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.251951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.530996Z digest=sha256:c90d4fd57e36e5793a750a812d56332d86172e481797df622083c7d8f212b230

Observation 5c92ac8d-3490-4e9f-8c4d-5a4a67a64fdc · outbound

This paper cites MotionClone: Training-Free Motion Cloning for Controllable Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency MotionClone: Training-Free Motion Cloning for Controllable Video Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.534345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.534345Z digest=sha256:a384ea27d32c9f6420f06afdc5aa14454e0a105bf4408d3e8e4286bad95b2e64

Observation e77631c0-2b1c-41d1-8c84-1fe7be37657e · outbound

This paper cites Video-p2p: Video editing with cross-attention control.

TransPixeler: Advancing Text-to-Video Generation with Transparency Video-p2p: Video editing with cross-attention control

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.537985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.537985Z digest=sha256:0d50a5b67ce17069470a4c403420e1ea2d7c00dfdea64c89443eb6f55e41d18e

Observation 73b3bc10-07f2-49b6-b63d-0d20087e970a · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

TransPixeler: Advancing Text-to-Video Generation with Transparency Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.540995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.540995Z digest=sha256:477dfbaf9e2f808cc8b8543e4f24c7725dc76f4557fccfd2607dbf02f155422c

Observation 1f20a71a-6e2e-49a3-aa33-8673cf478c84 · outbound

This paper cites Wonder3d: Sin- gle image to 3d using cross-domain diffusion.

TransPixeler: Advancing Text-to-Video Generation with Transparency Wonder3d: Sin- gle image to 3d using cross-domain diffusion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.234487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.544546Z digest=sha256:d9b60846f8b0160d292568b1b902dc646f663e1975b323f8938446fd4de0f8fd

Observation 6a6090dc-e376-45f9-936f-03ad39ddaaab · outbound

This paper cites Intrinsicdiffusion: Joint in- trinsic layers from latent diffusion models.

TransPixeler: Advancing Text-to-Video Generation with Transparency Intrinsicdiffusion: Joint in- trinsic layers from latent diffusion models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.223573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.548015Z digest=sha256:1eb49fafb044b89c741555594f968936f1d5547bd0f0fafdca3e70d45cdbedc4

Observation 4b1f40ac-4964-4c72-a75e-24dde494b1ef · outbound

This paper cites TrailBlazer: Trajectory Control for Diffusion-Based Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency TrailBlazer: Trajectory Control for Diffusion-Based Video Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.551286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.551286Z digest=sha256:8303aabe6959799d66093cb9127ea0f01724862d4e90d4f88cf319b3006fafe7

Observation 3db71a60-8289-4f83-91d0-680777d1e82d · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Latte: Latent Diffusion Transformer for Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.554678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.554678Z digest=sha256:8d564fbe5a1c625e14c8a8aaf7c8e6d17a2ada7c2c6d45e80238f7a582987923

Observation 0eaac291-299e-412f-9c83-c30b48190e0c · outbound

This paper cites MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model.

TransPixeler: Advancing Text-to-Video Generation with Transparency MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.557527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.557527Z digest=sha256:1d77f662a5d6ee40b9f23510c8da2eb75cb92b9967f74c68923cbf6f176e54c3

Observation 72746aa4-5145-4be5-907d-f8a76d3a14b6 · outbound

This paper cites Fatezero: Fus- ing attentions for zero-shot text-based video editing.

TransPixeler: Advancing Text-to-Video Generation with Transparency Fatezero: Fus- ing attentions for zero-shot text-based video editing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.560797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.560797Z digest=sha256:2e2c6255ac17a1e2df5e0ff30d958a1215ac5c98fc2e167b3af163107604d6b1

Observation db742f99-9932-407b-8807-fe12c9932609 · outbound

This paper cites Bimatting: Efficient video matting via binarization.

TransPixeler: Advancing Text-to-Video Generation with Transparency Bimatting: Efficient video matting via binarization

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.206501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.564688Z digest=sha256:6a958e716696536e5c874ea0e05f79bd78dead50dd217e945f6273e0993b7b9e

Observation 64eed75c-bf0c-4969-83bb-40aaea962e3b · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

TransPixeler: Advancing Text-to-Video Generation with Transparency SAM 2: Segment Anything in Images and Videos

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.567677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.567677Z digest=sha256:495b8e35981d05a40d63c60c782b4d34af9ae985ffe69693f9fef158f3d53980

Observation b109daf2-2980-4743-958d-42f79c300e2f · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

TransPixeler: Advancing Text-to-Video Generation with Transparency High-resolution image synthesis with latent diffusion models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.571367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.571367Z digest=sha256:7d33fa75867ba680caca483fc32f5a8ab691b554f61649045a812bf4b5f9851c

Observation 6b771ff2-5635-44b5-8a17-01c9229c72bc · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

TransPixeler: Advancing Text-to-Video Generation with Transparency Roformer: Enhanced transformer with rotary position embedding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.574514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.574514Z digest=sha256:8a81e95a48ccb0bed03713806761b0f976a5e05c178ccaac1dd799c74e3dc994

Observation 4b887705-b724-4b4f-9706-5f674216886e · outbound

This paper cites Mochi, 2024.

TransPixeler: Advancing Text-to-Video Generation with Transparency Mochi, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.182265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.578026Z digest=sha256:5b3623740764bc887c4de1a0f70dc2673b32b885a21b5a6e51bc1c25f579d537

Observation 2f6809d9-415c-4d7f-b01e-54c197622cf9 · outbound

This paper cites Fvd: A new metric for video generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Fvd: A new metric for video generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.171286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.581084Z digest=sha256:f83499790eb2b276a9cb839bf1efab3f4ca74827602dbe4d674ca6f3ce92b34d

Observation c2980faf-133d-4ffa-b348-550d351a2b2b · outbound

This paper cites Collaborative Control for Geometry-Conditioned PBR Image Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Collaborative Control for Geometry-Conditioned PBR Image Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.584320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.584320Z digest=sha256:1e20c4935673a7f697f932d25ca072c6265e7c6323c9a872e57b2e99b5a6aa45

Observation 6a872c3a-1c30-4301-8639-4cb92218b276 · outbound

This paper cites ModelScope Text-to-Video Technical Report.

TransPixeler: Advancing Text-to-Video Generation with Transparency ModelScope Text-to-Video Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.587944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.587944Z digest=sha256:3cdd8214ed60db99aa811b9d65c13eb38a2c916fbb27462b8fcb45b7483e077a

Observation 6b3dcc44-6f3f-4e3c-8ee4-888028d8e35b · outbound

This paper cites Motion Inversion for Video Customization.

TransPixeler: Advancing Text-to-Video Generation with Transparency Motion Inversion for Video Customization

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.591248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.591248Z digest=sha256:6b899e0b680c802102bbccaba24943dc874c5b84474fa799f543a330b47f650f

Observation 512d08bd-f631-41c7-9d90-58751db2b9c7 · outbound

This paper cites Linformer: Self-Attention with Linear Complexity.

TransPixeler: Advancing Text-to-Video Generation with Transparency Linformer: Self-Attention with Linear Complexity

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.594637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.594637Z digest=sha256:424e4ba976c98a77398724d90fda79c694a903f9e1071020a7de4830c864ab34

Observation 63c5cdd1-33ff-4cf2-8e4e-c9a19f0aeeb0 · outbound

This paper cites Videocomposer: Compositional video synthesis with motion controllability.

TransPixeler: Advancing Text-to-Video Generation with Transparency Videocomposer: Compositional video synthesis with motion controllability

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.159682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.598510Z digest=sha256:c853a771bc2039e49f88a8b671b629f0436f5823db927d80c1864aef19075b5a

Observation b6b97726-9a36-4c7b-a5d5-2d35ee45692c · outbound

This paper cites Matting by gen- eration.

TransPixeler: Advancing Text-to-Video Generation with Transparency Matting by gen- eration

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.146550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.601790Z digest=sha256:b247b5d549e998a5b6e013adc47f4fabcbf16eee3be09f30e589c75d01746b19

Observation 6ad6d65a-c755-4999-8fd4-84eb1afd773d · outbound

This paper cites Motionctrl: A unified and flexible motion controller for video generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Motionctrl: A unified and flexible motion controller for video generation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.133403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.604754Z digest=sha256:76d6f832cc105e9a4f496208d23edda535f2bde02c22ead1360f26a5a590eea5

Observation a2292f6f-e30b-407e-8475-aac7795b01d5 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.607961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.607961Z digest=sha256:e7fd63eaf9960ccbfad01a6a873154b4f095454c7f2dfc585f34fad2515553a0

Observation cb7266e3-3d5a-4b73-be7b-5e68cfddd3a5 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

TransPixeler: Advancing Text-to-Video Generation with Transparency Depth anything: Unleashing the power of large-scale unlabeled data

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.611264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.611264Z digest=sha256:4cf701dffb6c46825fa028962a2b432bf75e97adbd2685f643fc5f528dfa7b4a

Observation 4d5867ce-f8bf-4556-886f-e54222dece11 · outbound

This paper cites Defect spectrum: A granular look of large-scale defect datasets with rich semantics, 2023.

TransPixeler: Advancing Text-to-Video Generation with Transparency Defect spectrum: A granular look of large-scale defect datasets with rich semantics, 2023

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.097866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.614284Z digest=sha256:e3fc971d439ebc1dae7e34c2efd5338c14a26bea2e240506cfbee65413b1b37b

Observation 17a399d5-c3b3-4574-90ae-9e87fe0a46b0 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Rerender a video: Zero-shot text-guided video-to-video translation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.082993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.617986Z digest=sha256:0d6b6d71c4472eee64541847b8fc5a666052545231c03cc43b42a7658edf0a73

Observation 076a4e8c-9e4f-4c45-a651-b7bb78e3a649 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

TransPixeler: Advancing Text-to-Video Generation with Transparency CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.621400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.621400Z digest=sha256:1fb5dc018da7f7e81af1862eb857b1b46c46adcf2b1821d566e9ae0f6aed73be

Observation fa0a8843-2384-4d79-8045-fada524902e8 · outbound

This paper cites Vitmatte: Boosting image matting with pre- trained plain vision transformers.

TransPixeler: Advancing Text-to-Video Generation with Transparency Vitmatte: Boosting image matting with pre- trained plain vision transformers

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.068558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.625224Z digest=sha256:0c537089d32cc0818e7d13c6176e4a0fbfa508f78bc7ced291d26a245556a0f9

Observation d579d974-74f6-4ba5-bc0f-7b745600b61c · outbound

This paper cites Matte anything: Interactive natural image matting with seg- ment anything model.

TransPixeler: Advancing Text-to-Video Generation with Transparency Matte anything: Interactive natural image matting with seg- ment anything model

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.054262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.628305Z digest=sha256:7e9a9ae8f459421af2c73f4039bc7427f68a0753be69b5f9d02754e802f8d703

Observation 85d6428e-d914-41d8-81b8-e38b830e588e · outbound

This paper cites DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory.

TransPixeler: Advancing Text-to-Video Generation with Transparency DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.631543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.631543Z digest=sha256:175d56e4f89c60c67c076ad3267c0c40a6099e8aa2ab5a0a116013ea58ddc9e8

Observation 1372ece5-a478-4521-9186-b3cd2980fe5c · outbound

This paper cites Rgb ↔x: Image decomposition and synthesis using material-and lighting-aware diffusion models.

TransPixeler: Advancing Text-to-Video Generation with Transparency Rgb ↔x: Image decomposition and synthesis using material-and lighting-aware diffusion models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.038519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.635056Z digest=sha256:23874ef842bd400afc56889dd0a86be4aaad0f8247dbca6897a42dc9a9dbc1a9

Observation 0a911b25-bec6-4f3e-ae88-c69ad6e900f2 · outbound

This paper cites Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation.

TransPixeler: Advancing Text-to-Video Generation with Transparency Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.638542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.638542Z digest=sha256:13016898558192de68d59fc64dd8915199bcb1a8529756f7e4804a915e43bf79

Observation 51f06818-c4bc-4b66-a96d-abe6d9a07655 · outbound

This paper cites Moonshot: Towards Controllable Video Generation and Editing with Multimodal Conditions.

TransPixeler: Advancing Text-to-Video Generation with Transparency Moonshot: Towards Controllable Video Generation and Editing with Multimodal Conditions

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.641920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.641920Z digest=sha256:6770431f69a2cccef2e2cc9bfdbd289c82854b764488fad201f39b988d1f10d0

Observation 4ea32c1b-ee0a-4ca0-94af-761e78d1a6d2 · outbound

This paper cites Transparent Image Layer Diffusion using Latent Transparency.

TransPixeler: Advancing Text-to-Video Generation with Transparency Transparent Image Layer Diffusion using Latent Transparency

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.645613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.645613Z digest=sha256:4441cf069da568acd3e781108306cc4223e4ad913546ec657a1f4ae485bc9424

Observation d7f1cd08-dae3-42d8-b294-952a66ebaa54 · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

TransPixeler: Advancing Text-to-Video Generation with Transparency Open-sora: Democratizing efficient video production for all, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.024905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.648867Z digest=sha256:0c532ebf1112c5559b6103b077aac76af8fe106ee8e88e9825caec968b0008ad

Observation 101c528c-8c6d-47fb-8c0a-7296197812fe · outbound

This paper cites Long-short transformer: Efficient transformers for language and vision.

TransPixeler: Advancing Text-to-Video Generation with Transparency Long-short transformer: Efficient transformers for language and vision

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:14.007505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T22:03:13.651954Z digest=sha256:d7ed9583fc2d7f1831e8d2fcf19426b9e6ebe1a59cbafef6c78a578e3ae4c8dc

Observation 15c9ba3e-9780-4234-80b1-162c66cc75c6 · outbound

This paper cites LongNet: Scaling Transformers to 1,000,000,000 Tokens.

TransPixeler: Advancing Text-to-Video Generation with Transparency LongNet: Scaling Transformers to 1,000,000,000 Tokens

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.472136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.472136Z digest=sha256:f753ccb5cc3977523950045b4470a60b117c8851421e6a2a0583a3792315dc07

Pith citing papers

No inbound Pith citation observations are available.