Pith. sign in

Paper Citation Record · LEDGER

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2508.00289.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.00289 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:19:02.457827Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved14
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 61f32914-3bf4-4ef5-99a5-9ed11e27bdd0 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:56.826240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:56.826240Z digest=sha256:7facb25989ed037554b580f35688c00c423d363da987993a4a734b89018792b0

Observation 97ff0b24-2483-4afe-b4c9-ff89109da48c · outbound

This paper cites Universal guidance for diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Universal guidance for diffusion models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:56.952634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:56.952634Z digest=sha256:f44a8e2388e4d6af8d1af14cf381e13d2d74251208db57b6bf63e5787d6f753f

Observation 65cb2da8-1ade-4ef7-9343-8835e291098a · outbound

This paper cites Gradients without Backpropagation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Gradients without Backpropagation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:57.082183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:57.082183Z digest=sha256:4763fe15376d8b83838e088c70a32c4e4d2647c6c44c04afc640db90371cf376

Observation 5e741762-9efd-4d2e-9896-cd1944370230 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Vggsound: A large-scale audio-visual dataset

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:11.702942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:57.192695Z digest=sha256:5c11c7d80e72f069bd8ba5e2e863ef24783b3069ec85f936201ca0b23a03e912

Observation 85c9296b-5c14-4974-a0a9-539afb891a64 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:57.269811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:57.269811Z digest=sha256:0f59a576ea058cfc76f66ad84cdfcfc998f8b976993cbd6d8f498f976e4849fd

Observation 6613ac4e-a0ee-4d69-888d-5f7d6afcd8a9 · outbound

This paper cites FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:57.439470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:57.439470Z digest=sha256:f0a8f6504f60c4f02b721544a60b1602e86fa880cf8a49c7e9cf943b9fac35b5

Observation 4794412c-6459-41b2-8565-d8ec71da7764 · outbound

This paper cites Tweedie’s formula and selection bias.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Tweedie’s formula and selection bias

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:11.424128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:57.613325Z digest=sha256:dda89b49c4b9bf4e4d88702f9723d243998c4612a8710450eec5cbb737efc26d

Observation e045067b-375a-4c81-9083-815f84451c25 · outbound

This paper cites Douceur, Jon Howell, and Jared Saul.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Douceur, Jon Howell, and Jared Saul

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:11.166560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:57.724040Z digest=sha256:5f83256afaa7da2f1676b5092dd81bece372a0895dd5f10db9eb6884c4f3f824

Observation e1f4623e-a616-43d3-b1b2-ef71f27d9dbc · outbound

This paper cites Can forward gra- dient match backpropagation? In International Conference on Machine Learning, pages 10249–10264.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Can forward gra- dient match backpropagation? In International Conference on Machine Learning, pages 10249–10264

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:10.878163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:57.821630Z digest=sha256:579abd3910f35b2bc63e2cca87e4fc4a1979a1479ff53ad1478695c7e3cc8391

Observation 780e7832-8161-4df9-8f09-8a3602faf2b3 · outbound

This paper cites Imagebind: One embedding space to bind them all.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Imagebind: One embedding space to bind them all

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:10.534267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:57.972339Z digest=sha256:329bf3a3df0d8ca48b111170a4971ac8afbf9b7dc3cf9172aa562209bbffd87f

Observation 353fe7c7-d430-439c-a0c6-2b15c9063727 · outbound

This paper cites Animatediff: Animate your personalized text-to- image diffusion models without specific tuning.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Animatediff: Animate your personalized text-to- image diffusion models without specific tuning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:10.212799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:58.085809Z digest=sha256:fd6a35130752794ec9b1f7a953df1fe1e985e13b72e433c1411e919d8d2f8e50

Observation 0c7c502e-e2ee-4244-aa30-ef2f52fa8d91 · outbound

This paper cites Manifold preserving guided diffusion.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Manifold preserving guided diffusion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:09.868357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:58.205603Z digest=sha256:15cc8e3b292db769ba95a083a73af33ace382bd8da3c489ce1fbf15b6f5b0721

Observation e5f97f8a-a3fc-42d4-af10-67abb484eccc · outbound

This paper cites Denoising dif- fusion probabilistic models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Denoising dif- fusion probabilistic models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:09.502025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:58.394471Z digest=sha256:06006411cda59b392d8f0150b42158fae1e1abaffe996e8215dc8e0664290fc6

Observation ca448dae-a797-4ecf-91a5-667456f374ad · outbound

This paper cites Deep learning face attributes in the wild.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Deep learning face attributes in the wild

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:09.273484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:58.564074Z digest=sha256:a08a1c3abe11a56da056bda23d0f122f1c0eb8f9cd263dbcc8d9e12dcd98758d

Observation 109e8f0a-ab82-41a2-97a5-9426a5050ddd · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:58.687765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:58.687765Z digest=sha256:77e8b728a59559dd5d5eb1e229f450a2f05c5b2275f7b8ddcd31845e972d1a92

Observation bc306bb7-b91a-4b83-871f-8e77b75dd7e2 · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:58.787547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:58.787547Z digest=sha256:b16ccb09385d27c37b97e7f7f7eafa3a87622fad27aac6d903e9bfffcbce4052

Observation f701c130-ba86-49e1-8dc5-720392997abb · outbound

This paper cites Patel, and Tim K.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Patel, and Tim K

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:08.928742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:58.894322Z digest=sha256:6bb90bf65d40355f238b79127c232580f1b2886418a14a80029b4b39256d1d81

Observation b664519c-6427-4068-9362-0873ec69e3ca · outbound

This paper cites Steered diffusion: A generalized framework for plug- and-play conditional image synthesis.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Steered diffusion: A generalized framework for plug- and-play conditional image synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:08.636474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:59.035988Z digest=sha256:2cfb780a75a3ae536867fb582d3692f4be544e433ae948955258a26dac3c06aa

Observation 84481e46-f891-4932-a068-bdd3f63dc99d · outbound

This paper cites Ditto: Diffusion inference-time t- optimization for music generation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Ditto: Diffusion inference-time t- optimization for music generation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:08.310674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:59.138737Z digest=sha256:7e6e6b5958325c9f0bb1822efb9069e6fb1107c2c31c66e33d53cefbdccee6a3

Observation e8293c18-bbc8-4b46-83b0-03320f0856f9 · outbound

This paper cites Automatic differentiation in pytorch.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Automatic differentiation in pytorch

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.937331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:59.271633Z digest=sha256:7e20301fdbc878c69fb53eb2f8f0c7ed9bdd14481a77c969182cf7c1bce95f0e

Observation 6f0f0cf5-2519-4903-bfe6-a7ba4f508d54 · outbound

This paper cites Styleclip: Text-driven manipulation of stylegan imagery.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Styleclip: Text-driven manipulation of stylegan imagery

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.673876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:59.412916Z digest=sha256:4eb083312ca99fae880e640cdb32ba12a92d971545fc4a2c7ea3a0f8beb2afbd

Observation f431413d-636f-4661-9c85-40a0ccca9d2d · outbound

This paper cites Scaling Forward Gradient With Local Losses.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Scaling Forward Gradient With Local Losses

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:59.548616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:59.548616Z digest=sha256:153dbb1ff70e747826865aa3777a57f07980b1672c9cb7615b1f7b0510b7d1e7

Observation 05ce7fb0-539f-4024-a778-1b7e3702cfcc · outbound

This paper cites Denoising Diffusion Implicit Models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Denoising Diffusion Implicit Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:59.650703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:59.650703Z digest=sha256:5c3edd4345b8eb25a0b8c949508ec14f6dfd4d62caae813b6c6853b99d65a64c

Observation 5b7dfdf3-f96a-4646-af6a-c4f7a011a005 · outbound

This paper cites Loss-guided diffusion models for plug-and-play controllable generation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Loss-guided diffusion models for plug-and-play controllable generation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.430315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:18:59.799031Z digest=sha256:5cbf13808a8e46040485880014cf2409775d873d0764f77df7cd2e35cc48f253

Observation d0929762-7452-4243-8153-6d1db2256879 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Score-Based Generative Modeling through Stochastic Differential Equations

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:59.912795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:18:59.912795Z digest=sha256:b801f576aaf1b679077f40adcf6f2dcbe564112890d221dad42e4cf1cf34043f

Observation ed4047d1-92d2-4b7a-bcc7-8cc58542da2f · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T10:19:00.091556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:19:00.091556Z digest=sha256:2c33375f7c2a80a23fed5109bea3d191de75af29cecab63553b13bafde37daf6

Observation 3cc2f208-d7cc-4330-aaea-86349019cc39 · outbound

This paper cites EDICT: Exact Diffusion Inversion via Coupled Transformations.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models EDICT: Exact Diffusion Inversion via Coupled Transformations

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T10:19:00.213508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:19:00.213508Z digest=sha256:b0e93f2e5b481b945d47931c1dac2e410e62a47c34020bf08050ef145f81e8b9

Observation 78aa970c-3d51-492f-99c1-2bbb2c9ab251 · outbound

This paper cites End-to-end diffusion latent optimization improves classifier guidance.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models End-to-end diffusion latent optimization improves classifier guidance

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:07.137812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:00.336983Z digest=sha256:f66236bf6b83046f7019769a19995396fa462d84c71cf89e941a9317441ae07e

Observation 99ed74fa-e27b-4e31-a287-b25c27dc2ee7 · outbound

This paper cites Exploring video quality assessment on user generated contents from aesthetic and technical perspectives.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Exploring video quality assessment on user generated contents from aesthetic and technical perspectives

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:06.744134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:00.614727Z digest=sha256:84aaf0928ab462565e05dd7cd6d12fb07042df27327863b5b67e42142d16a714

Observation ea998301-69ef-477d-a762-6a86dd3169c6 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:06.458827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:00.710597Z digest=sha256:79bdcc29444feedac06b98bece23ede826d4fc5ee07b094189bbf690961aa66c

Observation 8bec9df8-22b1-4dad-bf9e-1b9172dcf155 · outbound

This paper cites Cvpr 2023 text guided video editing competition, 2023.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Cvpr 2023 text guided video editing competition, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:06.190487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:00.838968Z digest=sha256:f982dff23fa88e95f35238d3ccd5b0b891c76a990d401f79e82ddf07746ed37b

Observation 044e5235-c78a-4483-8093-2f2414993387 · outbound

This paper cites Seeing and hearing: Open-domain visual- audio generation with diffusion latent aligners.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Seeing and hearing: Open-domain visual- audio generation with diffusion latent aligners

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:05.917568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:00.953443Z digest=sha256:8c278e9132bbcf72970302241006af37e5be90b6705595346966e33b30c6774b

Observation 25503311-d8e4-4b04-bb5f-b1719b6d39e0 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:19:01.037098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:19:01.037098Z digest=sha256:65fe3531409c79a56da3e821af101d594830e991311ff4ce7cc2cb4d377110ec

Observation 6167f9db-803e-41b5-8ee3-9628d404505e · outbound

This paper cites Diverse and aligned audio-to-video genera- tion via text-to-video model adaptation, 2023.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Diverse and aligned audio-to-video genera- tion via text-to-video model adaptation, 2023

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:05.634407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:01.163520Z digest=sha256:8f7ac8e57064e7a94633cca591c03d714738384dd2c71ff2e5efff63ea5818ee

Observation 7cfc12ab-ee6c-498c-83d7-65717fb6f311 · outbound

This paper cites Tfg: Unified training-free guidance for diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Tfg: Unified training-free guidance for diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:05.321040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:01.311418Z digest=sha256:08dc34017c452dc9b5dcc05da590a08fe693e6a54badd0a0ab4a04add5e9418b

Observation f0c7fe81-8da2-4468-b3eb-380fd05c7921 · outbound

This paper cites Freedom: Training-free energy-guided condi- tional diffusion model.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Freedom: Training-free energy-guided condi- tional diffusion model

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:04.670407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:01.596884Z digest=sha256:a93c07804d06b2286c55b6f2798d68f90eb08c8fb109e324a41ca06f5acc889a

Observation 536b11ec-c3a9-494b-b212-70f0c23e8c2f · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Adding conditional control to text-to-image diffusion models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:04.314085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:01.748117Z digest=sha256:c311ce4794e3371f297cc6d3f60b7026edd60a54e96afe6efb2d257a17213549

Observation 712f5a20-7764-4359-a4a4-f09137e3f6f0 · outbound

This paper cites Let f : Rm − →Rn.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Let f : Rm − →Rn

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:04.092191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:01.905666Z digest=sha256:7ebb0a1d926a524f4857ec646ed7a48ea31ab90115308fdbba90960781dcb6b2

Observation 497cee0c-5b7c-429f-bc98-83ce86c2c8f6 · outbound

This paper cites In all experiments, we em- ploy AnimateDiff [11] with epiCRealism 2 as the base text- to-image model to generate 16 frames 8fps.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models In all experiments, we em- ploy AnimateDiff [11] with epiCRealism 2 as the base text- to-image model to generate 16 frames 8fps

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:03.806570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:02.041513Z digest=sha256:f3e7ba2c26bc0df64a411f84c752c047b52f74d5aaf2be7bbf8c414ab65835f4

Observation ad2b222c-fb6a-40cc-b1c5-8c00e1f73c1a · outbound

This paper cites A bird in a forest.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models A bird in a forest

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:03.496573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:02.169544Z digest=sha256:2b858af2d0e8ae2e535859caa3842d50ccc7c061bb694834f8451570da2a877b

Observation 554474bf-cd2d-4e37-87b8-6eb1402f756a · outbound

This paper cites However, since our primary focus is text- to-video tasks, we do not explore this aspect in depth but provide evidence of its applicability.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models However, since our primary focus is text- to-video tasks, we do not explore this aspect in depth but provide evidence of its applicability

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:03.243478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:02.326749Z digest=sha256:0a95bfeaead860e04d24854fc1bba3d20fa1cb73492dc1c06d521a75b5884b0a

Observation 9cc8f007-c6f9-46bb-bc68-7ac05ac747c9 · outbound

This paper cites Following [35], for super resolution, and deblurring tasks, we use the CAT-DDPM diffusion model trained on the CAT dataset [8].

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Following [35], for super resolution, and deblurring tasks, we use the CAT-DDPM diffusion model trained on the CAT dataset [8]

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:19:02.964526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:02.457827Z digest=sha256:46da1f0e3ff127683c47a6f8943d56e4393c71d4e7eb3ec3f0393d141267e579

Observation 1fa22e79-3b2e-423f-934b-7d712ac44f24 · outbound

This paper cites an unresolved cited work.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Unresolved cited work

Reference 2023

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T10:19:06.976010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:00.464586Z digest=sha256:3d9364d54014d43be2068a87f764f6fd9c1cf17d0444740d84f19c62609cec2d

Observation f2631690-0718-47f1-babf-55c4a7fb9daf · outbound

This paper cites an unresolved cited work.

TITAN-Guide: Taming Inference-Time AligNment for Guided Text-to-Video Diffusion Models Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:19:04.970152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:19:01.475513Z digest=sha256:01ddd3cd2fbf51fe9331860d514e973e679e2838b86f7d19bda16d7792821ad2

Pith citing papers

No inbound Pith citation observations are available.