Pith. sign in

Paper Citation Record · LEDGER

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2508.00289.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.00289 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:19:02.457827Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved14
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 61f32914-3bf4-4ef5-99a5-9ed11e27bdd0 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:56.826240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:56.826240Z digest=sha256:7facb25989ed037554b580f35688c00c423d363da987993a4a734b89018792b0

Observation 97ff0b24-2483-4afe-b4c9-ff89109da48c · outbound

This paper cites Universal guidance for diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Universal guidance for diffusion models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:56.952634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:56.952634Z digest=sha256:f44a8e2388e4d6af8d1af14cf381e13d2d74251208db57b6bf63e5787d6f753f

Observation 65cb2da8-1ade-4ef7-9343-8835e291098a · outbound

This paper cites Gradients without Backpropagation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Gradients without Backpropagation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:57.082183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:57.082183Z digest=sha256:4763fe15376d8b83838e088c70a32c4e4d2647c6c44c04afc640db90371cf376

Observation 5e741762-9efd-4d2e-9896-cd1944370230 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Vggsound: A large-scale audio-visual dataset

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:11.702942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:57.192695Z digest=sha256:5aeef76c41b02783454d4bb20cf4fb2477d8a74523d30c23d862e613471df4b7

Observation 85c9296b-5c14-4974-a0a9-539afb891a64 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:57.269811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:57.269811Z digest=sha256:0f59a576ea058cfc76f66ad84cdfcfc998f8b976993cbd6d8f498f976e4849fd

Observation 6613ac4e-a0ee-4d69-888d-5f7d6afcd8a9 · outbound

This paper cites FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:57.439470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:57.439470Z digest=sha256:f0a8f6504f60c4f02b721544a60b1602e86fa880cf8a49c7e9cf943b9fac35b5

Observation 4794412c-6459-41b2-8565-d8ec71da7764 · outbound

This paper cites Tweedie’s formula and selection bias.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Tweedie’s formula and selection bias

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:11.424128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:57.613325Z digest=sha256:ccfa0d23e73179c63d3a0f0a4c7076df0e1853477b5dc2798d268714c01ffc15

Observation e045067b-375a-4c81-9083-815f84451c25 · outbound

This paper cites Douceur, Jon Howell, and Jared Saul.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Douceur, Jon Howell, and Jared Saul

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:11.166560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:57.724040Z digest=sha256:6cc29c4945a43454c876744e3da54ab7f5f6e8ac9603c348b8bc41719bab0725

Observation e1f4623e-a616-43d3-b1b2-ef71f27d9dbc · outbound

This paper cites Can forward gra- dient match backpropagation? In International Conference on Machine Learning, pages 10249–10264.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Can forward gra- dient match backpropagation? In International Conference on Machine Learning, pages 10249–10264

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:10.878163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:57.821630Z digest=sha256:8b3d923e815387b195e8bec3b1ec4f2b5242eb3ca84dcfa5eaa6671ace896387

Observation 780e7832-8161-4df9-8f09-8a3602faf2b3 · outbound

This paper cites Imagebind: One embedding space to bind them all.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Imagebind: One embedding space to bind them all

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:10.534267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:57.972339Z digest=sha256:c4e715b35709b89faa1f01ccc3c124ab18a88762aea89c59cbdd56a78fc069c7

Observation 353fe7c7-d430-439c-a0c6-2b15c9063727 · outbound

This paper cites Animatediff: Animate your personalized text-to- image diffusion models without specific tuning.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Animatediff: Animate your personalized text-to- image diffusion models without specific tuning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:10.212799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:58.085809Z digest=sha256:e4e626013fecb9f901e4142fc863141a3f18d0714b5222edba66ec7a5fc7849c

Observation 0c7c502e-e2ee-4244-aa30-ef2f52fa8d91 · outbound

This paper cites Manifold preserving guided diffusion.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Manifold preserving guided diffusion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:09.868357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:58.205603Z digest=sha256:80bc75a69fd06756d4ef99f52652dbb08594a77f59feb656d23bc1da19abe6a1

Observation e5f97f8a-a3fc-42d4-af10-67abb484eccc · outbound

This paper cites Denoising dif- fusion probabilistic models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Denoising dif- fusion probabilistic models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:09.502025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:58.394471Z digest=sha256:4a8f9259d921a860c874a77d7211e94a44c4293347eaab83949c5ac7ac91427b

Observation ca448dae-a797-4ecf-91a5-667456f374ad · outbound

This paper cites Deep learning face attributes in the wild.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Deep learning face attributes in the wild

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:09.273484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:58.564074Z digest=sha256:8359eae9e31cd85be4a14ce188c53c99d62547511824b2e1240119a78704b1cf

Observation 109e8f0a-ab82-41a2-97a5-9426a5050ddd · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:58.687765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:58.687765Z digest=sha256:93bac358c98bb196c50e671e466144ac21b95ee8dc3a08d8eb5421594f3aa7ec

Observation bc306bb7-b91a-4b83-871f-8e77b75dd7e2 · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:58.787547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:58.787547Z digest=sha256:b16ccb09385d27c37b97e7f7f7eafa3a87622fad27aac6d903e9bfffcbce4052

Observation f701c130-ba86-49e1-8dc5-720392997abb · outbound

This paper cites Patel, and Tim K.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Patel, and Tim K

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:08.928742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:58.894322Z digest=sha256:ad1ff5b470e0e9b8fac0ef38c6ce832c9aaf2ea526800b71068d5e4af89bbe34

Observation b664519c-6427-4068-9362-0873ec69e3ca · outbound

This paper cites Steered diffusion: A generalized framework for plug- and-play conditional image synthesis.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Steered diffusion: A generalized framework for plug- and-play conditional image synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:08.636474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:59.035988Z digest=sha256:18b51913c11c184c5343818f044e868d469fa3aa8b4598f094d42042cb6dcaaa

Observation 84481e46-f891-4932-a068-bdd3f63dc99d · outbound

This paper cites Ditto: Diffusion inference-time t- optimization for music generation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Ditto: Diffusion inference-time t- optimization for music generation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:08.310674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:59.138737Z digest=sha256:524b2da9e96810cbd3728c37681a210e9dee4a467c42a76b9f70c0ea032d641d

Observation e8293c18-bbc8-4b46-83b0-03320f0856f9 · outbound

This paper cites Automatic differentiation in pytorch.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Automatic differentiation in pytorch

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.937331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:59.271633Z digest=sha256:68aed46ad43d9ed140500fc718346a7a916af8a05d2fa74087436c7e5c634848

Observation 6f0f0cf5-2519-4903-bfe6-a7ba4f508d54 · outbound

This paper cites Styleclip: Text-driven manipulation of stylegan imagery.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Styleclip: Text-driven manipulation of stylegan imagery

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.673876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:59.412916Z digest=sha256:1a4c2f3263fcdeb70510b60c16a07918d7bb843345c300272757242cd324b301

Observation f431413d-636f-4661-9c85-40a0ccca9d2d · outbound

This paper cites Scaling Forward Gradient With Local Losses.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Scaling Forward Gradient With Local Losses

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:59.548616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:59.548616Z digest=sha256:153dbb1ff70e747826865aa3777a57f07980b1672c9cb7615b1f7b0510b7d1e7

Observation 05ce7fb0-539f-4024-a778-1b7e3702cfcc · outbound

This paper cites Denoising Diffusion Implicit Models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Denoising Diffusion Implicit Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:59.650703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:59.650703Z digest=sha256:5c3edd4345b8eb25a0b8c949508ec14f6dfd4d62caae813b6c6853b99d65a64c

Observation 5b7dfdf3-f96a-4646-af6a-c4f7a011a005 · outbound

This paper cites Loss-guided diffusion models for plug-and-play controllable generation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Loss-guided diffusion models for plug-and-play controllable generation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.430315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:18:59.799031Z digest=sha256:2967eee86aab81f2bf5619cfed8b42088dd0d223000c0a21100bd0706e6c2185

Observation d0929762-7452-4243-8153-6d1db2256879 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Score-Based Generative Modeling through Stochastic Differential Equations

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:59.912795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:59.912795Z digest=sha256:b801f576aaf1b679077f40adcf6f2dcbe564112890d221dad42e4cf1cf34043f

Observation ed4047d1-92d2-4b7a-bcc7-8cc58542da2f · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T10:19:00.091556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:19:00.091556Z digest=sha256:2c33375f7c2a80a23fed5109bea3d191de75af29cecab63553b13bafde37daf6

Observation 3cc2f208-d7cc-4330-aaea-86349019cc39 · outbound

This paper cites EDICT: Exact Diffusion Inversion via Coupled Transformations.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models EDICT: Exact Diffusion Inversion via Coupled Transformations

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T10:19:00.213508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:19:00.213508Z digest=sha256:b0e93f2e5b481b945d47931c1dac2e410e62a47c34020bf08050ef145f81e8b9

Observation 78aa970c-3d51-492f-99c1-2bbb2c9ab251 · outbound

This paper cites End-to-end diffusion latent optimization improves classifier guidance.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models End-to-end diffusion latent optimization improves classifier guidance

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.137812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:00.336983Z digest=sha256:4bc2ecaa0b9c8fe658c957ad9ebb6438d123a6c2bc70231622dedf20f737511b

Observation 99ed74fa-e27b-4e31-a287-b25c27dc2ee7 · outbound

This paper cites Exploring video quality assessment on user generated contents from aesthetic and technical perspectives.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Exploring video quality assessment on user generated contents from aesthetic and technical perspectives

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:06.744134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:00.614727Z digest=sha256:2fa834e8051964e8fad3337f957095812964c761ae549325671afbd5d1ef150f

Observation ea998301-69ef-477d-a762-6a86dd3169c6 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:06.458827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:00.710597Z digest=sha256:ede81119fae29de23b86e814204b06bc8049cee13f9bc4c1064b85a6a36dbb83

Observation 8bec9df8-22b1-4dad-bf9e-1b9172dcf155 · outbound

This paper cites Cvpr 2023 text guided video editing competition, 2023.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Cvpr 2023 text guided video editing competition, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:06.190487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:00.838968Z digest=sha256:828b3aedf4ea6fddfca1e36ef8e9dc68d780254ee534fc79edb111a8e244bd1b

Observation 044e5235-c78a-4483-8093-2f2414993387 · outbound

This paper cites Seeing and hearing: Open-domain visual- audio generation with diffusion latent aligners.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Seeing and hearing: Open-domain visual- audio generation with diffusion latent aligners

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:05.917568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:00.953443Z digest=sha256:1fd846377747ee93d64b30f02cf0dfc04b6e1ff41781bae8490229e9cd6d30c2

Observation 25503311-d8e4-4b04-bb5f-b1719b6d39e0 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:19:01.037098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:19:01.037098Z digest=sha256:65fe3531409c79a56da3e821af101d594830e991311ff4ce7cc2cb4d377110ec

Observation 6167f9db-803e-41b5-8ee3-9628d404505e · outbound

This paper cites Diverse and aligned audio-to-video genera- tion via text-to-video model adaptation, 2023.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Diverse and aligned audio-to-video genera- tion via text-to-video model adaptation, 2023

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:05.634407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:01.163520Z digest=sha256:72294614541483b4bcc098693b43ef6bb0b1ff4e7f42c58b85ebe3f8575dded1

Observation 7cfc12ab-ee6c-498c-83d7-65717fb6f311 · outbound

This paper cites Tfg: Unified training-free guidance for diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Tfg: Unified training-free guidance for diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:05.321040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:01.311418Z digest=sha256:f5d70ab649831789df12a7ffab68d8fda28e7da79aad6ebf2a41faa78074fd95

Observation f0c7fe81-8da2-4468-b3eb-380fd05c7921 · outbound

This paper cites Freedom: Training-free energy-guided condi- tional diffusion model.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Freedom: Training-free energy-guided condi- tional diffusion model

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:04.670407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:01.596884Z digest=sha256:6647b5fa1d9b3fd9fb6a9ffe2f0a900126992d0eb92204ce55834e52b7d269b5

Observation 536b11ec-c3a9-494b-b212-70f0c23e8c2f · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Adding conditional control to text-to-image diffusion models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:04.314085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:01.748117Z digest=sha256:0eb6003110b9d0b48f6213348b3c044772e4d09561719fc0b783daf6c6e6b68f

Observation 712f5a20-7764-4359-a4a4-f09137e3f6f0 · outbound

This paper cites Let f : Rm − →Rn.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Let f : Rm − →Rn

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:04.092191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:01.905666Z digest=sha256:02424f4811b1cd339e35a389462cff553083b8406503e0602dba4be78fc59a88

Observation 497cee0c-5b7c-429f-bc98-83ce86c2c8f6 · outbound

This paper cites In all experiments, we em- ploy AnimateDiff [11] with epiCRealism 2 as the base text- to-image model to generate 16 frames 8fps.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models In all experiments, we em- ploy AnimateDiff [11] with epiCRealism 2 as the base text- to-image model to generate 16 frames 8fps

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:03.806570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:02.041513Z digest=sha256:9b2c759d3dd4382ff82b3ce75e38a0c958eb483b5c6ae442ea96cf04c437985e

Observation ad2b222c-fb6a-40cc-b1c5-8c00e1f73c1a · outbound

This paper cites A bird in a forest.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models A bird in a forest

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:03.496573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:02.169544Z digest=sha256:a8d08bd0747ce5d0a1205016ed0f3695c7a1dfb21895cd65a536da3923945c94

Observation 554474bf-cd2d-4e37-87b8-6eb1402f756a · outbound

This paper cites However, since our primary focus is text- to-video tasks, we do not explore this aspect in depth but provide evidence of its applicability.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models However, since our primary focus is text- to-video tasks, we do not explore this aspect in depth but provide evidence of its applicability

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:03.243478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:02.326749Z digest=sha256:86098f8d4bc6c8bd9d490ce0fd2b004e5af53e9efc947d893d4538c870c999fa

Observation 9cc8f007-c6f9-46bb-bc68-7ac05ac747c9 · outbound

This paper cites Following [35], for super resolution, and deblurring tasks, we use the CAT-DDPM diffusion model trained on the CAT dataset [8].

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Following [35], for super resolution, and deblurring tasks, we use the CAT-DDPM diffusion model trained on the CAT dataset [8]

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:02.964526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:02.457827Z digest=sha256:38801026ec7d0d221ec5fc439d4b5c3833540a8c830d409c469188716e2b6ddf

Observation 1fa22e79-3b2e-423f-934b-7d712ac44f24 · outbound

This paper cites an unresolved cited work.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Unresolved cited work

Reference 2023

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T10:19:06.976010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:00.464586Z digest=sha256:473dc755181e8e86f5cf6eb951ff5bcb63c88721682fbc02b69c4833b1a041b6

Observation f2631690-0718-47f1-babf-55c4a7fb9daf · outbound

This paper cites an unresolved cited work.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:19:04.970152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:19:01.475513Z digest=sha256:9bcaf4eda7a54d0050529c59a924d6073d0816b56dfc25574e87d2e13dea1e88

Pith citing papers

No inbound Pith citation observations are available.