Pith. sign in

Paper Citation Record · LEDGER

LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2412.04814.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.04814 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T03:04:44.499036Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:09:29.415956Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 26a2e067-1bc8-4ef8-b328-94d6cc290264 · inbound

Improving Video Generation with Human Feedback cites this paper.

Improving Video Generation with Human Feedback LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:30:02.820925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T15:30:02.578430Z digest=sha256:0df363217d78e535f1b80a27922a2b5653ac38887518b8fec00542a744bfacfc

Observation fb551883-4bf8-404d-84e4-118056d81d74 · inbound

Unified Reward Model for Multimodal Understanding and Generation cites this paper.

Unified Reward Model for Multimodal Understanding and Generation LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:44:30.907784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T00:44:30.558048Z digest=sha256:b14cf790e6c87a70865bba0ab379bf83344f2f512af0469ecef9f6476c379f5b

Observation 7223992c-bdde-4e4c-92d6-60682d147d0b · inbound

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling cites this paper.

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:15:54.437321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T05:13:42.934115Z digest=sha256:b25ccd085e24a6c19d045ab6a59ef0148267210edeb7b5ef89c6bf257b7ddd24

Observation 3f4938c0-35dd-41a5-9dac-d47cbd1d8b98 · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:00:27.260501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:86889570a2c682253acddf5485d38625a3c2ee7b419764fb800a8efbe5668f7a

Observation dc25a1a6-9e7f-4971-a8d6-4e92b8c3e659 · inbound

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models cites this paper.

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:28:05.557128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T16:25:03.743594Z digest=sha256:76d47d4b0dd4be430f67c97db87bbb820dc4a69c99c60bacc1415741b7a2022e

Observation 30d9a368-faee-43d0-a259-42251d648c52 · inbound

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models cites this paper.

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-21T16:04:14.695189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T16:01:52.150950Z digest=sha256:22492f4a4e852b043da97bdca2f1e96a3059b246f4dd83755fdabcf8048f152b

Observation b0eead6d-ecf0-4ee5-91ca-cdc3b2f7d1b6 · inbound

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation cites this paper.

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:44.499036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:44.499036Z digest=sha256:40fe1be0c1f9e8ebcf81562a017427b09698102d14c2cb4837e510b7dbef99e5

Observation 40da0efd-d045-4131-9956-2a02ba3bc2e7 · inbound

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling cites this paper.

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.755627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T07:37:52.346280Z digest=sha256:8dec1719b5062324ec23c11693eb84a71d2f65895ce83bdc645b691679cecad3

Observation c7b756ed-008d-4e65-9680-f1c02f06a5ca · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:22.083007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T06:00:31.582714Z digest=sha256:b1693594cef7070a082826bc1478366ad315f0900a97443ef3ddc94a878bdc65

Observation 4a59fe8f-2958-4793-829f-2ccdbf7ac32a · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:55:45.202918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T22:38:43.102769Z digest=sha256:42339ddcadc778ac9948e5bc0462e7ba4ec5dac11d3dbfa06ade11aeb32f3439

Observation e72140b6-6d2d-4569-9803-5c41a7bb1ea0 · inbound

DRM: Diffusion-based Reward Model With Step-wise Guidance cites this paper.

DRM: Diffusion-based Reward Model With Step-wise Guidance LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:13:59.429079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T22:12:39.225557Z digest=sha256:31d682cf8eb1c19a759ba8566c116e7a07941915c9ebb6f2f908a02f60b06c25

Observation 1b0fc3d7-38f8-46a3-97a6-ea35b1e814cd · inbound

Through the PRISM: Preference Representation in Intermediate States of Video Diffusion Models cites this paper.

Through the PRISM: Preference Representation in Intermediate States of Video Diffusion Models LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:09:29.417783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T18:25:49.041639Z digest=sha256:1eb1c5d1dcc50d7b928979cc9a676b3b918f1c60d18716a3ca8a0cceffe500b9

Observation e06c7cb8-e6e0-4222-94cc-022811bc940e · inbound

Reward Lightning: Fast Video Generation via Homologous Preference Distillation cites this paper.

Reward Lightning: Fast Video Generation via Homologous Preference Distillation LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-11T22:44:24.796514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:44:24.796514Z digest=sha256:2269c280c744c71a6f3cd25dd306bb05766980ff8041c99c065e27d86d4e9e71

Observation 547c4f4d-b0eb-40c6-8583-dadb4b70a32b · inbound

Temporal Concentration from Rollout Errors: Implicit Preference Optimization for Text-to-Video Diffusion cites this paper.

Temporal Concentration from Rollout Errors: Implicit Preference Optimization for Text-to-Video Diffusion LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T19:09:45.675824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T19:09:45.675824Z digest=sha256:aceeb14a60922392a59c3229654bc54987a52fadb4fdbf138f94cd915f500677