Pith. sign in

Paper Citation Record · LEDGER

Diffusion Model Alignment Using Direct Preference Optimization

As of 22 July 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2311.12908.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.12908 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-22T06:31:00.163083+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-09T02:17:20.589485Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T02:25:55.896149Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d0cce014-8d2e-41ce-ad5e-ba55dec556c8 · inbound

Scaling Rectified Flow Transformers for High-Resolution Image Synthesis cites this paper.

Scaling Rectified Flow Transformers for High-Resolution Image Synthesis Diffusion Model Alignment Using Direct Preference Optimization

Reference 191

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:27:53.624782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-12T08:27:53.446686Z digest=sha256:14480f0645fc39992c1770438b9a10b12ce84745c423f598e8a80e9e4efa5158

Observation 985d286e-bda1-4c57-81d3-b0dcdb611313 · inbound

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models cites this paper.

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Diffusion Model Alignment Using Direct Preference Optimization

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:26:37.380158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-15T12:26:37.300599Z digest=sha256:ad9795d74a85ab8f91db79905b7cb7d3a3423a049f92879f775cbeffe38ac7ac

Observation 985cc89e-da0d-48ab-9eea-ac61e42cc091 · inbound

VideoPhy: Evaluating Physical Commonsense for Video Generation cites this paper.

VideoPhy: Evaluating Physical Commonsense for Video Generation Diffusion Model Alignment Using Direct Preference Optimization

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:34:37.694145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-20T11:34:37.599691Z digest=sha256:7b221e39fdd8fca56661b52d942ab86909967b8e9e4b9b98e08e0f962967476f

Observation 47df5f88-d8fa-4abf-bd93-f6f80394c194 · inbound

Diffusion Policy Policy Optimization cites this paper.

Diffusion Policy Policy Optimization Diffusion Model Alignment Using Direct Preference Optimization

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:48:15.024198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-16T08:48:14.776754Z digest=sha256:8fde29425fa8db360453fad6a50208afedf18b2f3734259543dfa687b0068398

Observation 509484c6-bea9-4a1f-8ed3-df809414bbd7 · inbound

Alignment and Safety of Diffusion Models via Reinforcement Learning and Reward Modeling: A Survey cites this paper.

Alignment and Safety of Diffusion Models via Reinforcement Learning and Reward Modeling: A Survey Diffusion Model Alignment Using Direct Preference Optimization

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T02:30:56.323210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-22T02:28:23.754561Z digest=sha256:5b18aff20fc5d0baeff4a50b1b6eb4680a5778cf275ebd12307870430caab4b0

Observation 92607226-6b1e-4b0f-81f4-da6dd73d7892 · inbound

Listener-Rewarded Thinking in VLMs for Image Preferences cites this paper.

Listener-Rewarded Thinking in VLMs for Image Preferences Diffusion Model Alignment Using Direct Preference Optimization

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:42:09.214747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-19T07:38:52.273903Z digest=sha256:dcf3d959f193e0c3019f0485bcc917c989cd3579d8ad488c74abaafeecc6e185

Observation 5057d666-5317-4f57-b2f9-32a4ddf03735 · inbound

Collective Recourse for Generative Urban Visualizations cites this paper.

Collective Recourse for Generative Urban Visualizations Diffusion Model Alignment Using Direct Preference Optimization

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:11:40.169243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-18T17:10:54.720256Z digest=sha256:ea6256d7edd28b1749bff5ca9d5b0cd8a96255b84e4372c04339d6c405f3152e

Observation bd822f63-01ca-41be-8b19-3e9c34300254 · inbound

D2 Actor Critic: Diffusion Actor Meets Distributional Critic cites this paper.

D2 Actor Critic: Diffusion Actor Meets Distributional Critic Diffusion Model Alignment Using Direct Preference Optimization

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:35:29.484004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-25T07:31:27.330951Z digest=sha256:3fa90220d438390f0c89c94ae4fd670192f3319edd50b1db2d4a9ea95d02983a

Observation 89ee6b99-e330-4c66-ab0c-ad610206dc4c · inbound

IdGlow: Dynamic Identity Modulation for Multi-Subject Generation cites this paper.

IdGlow: Dynamic Identity Modulation for Multi-Subject Generation Diffusion Model Alignment Using Direct Preference Optimization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:00:10.047290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-21T12:56:27.928008Z digest=sha256:8a0ce018db75df729a11c2524a38406a3dfc84a454e828ea6f38014061c58089

Observation c8d802c4-28ad-47bb-99ac-4ffb238413d5 · inbound

VASR: Variance-Aware Systematic Resampling for Reward-Guided Diffusion cites this paper.

VASR: Variance-Aware Systematic Resampling for Reward-Guided Diffusion Diffusion Model Alignment Using Direct Preference Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:26:00.391432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T18:08:03.771557Z digest=sha256:a33b0bc685585e13630b8144e374921a47d6c50d7d8d983485dc49a8bbab235d

Observation e8e5510d-108e-4ebd-9910-b9e5c5cb900d · inbound

VASR: Variance-Aware Systematic Resampling for Reward-Guided Diffusion cites this paper.

VASR: Variance-Aware Systematic Resampling for Reward-Guided Diffusion Diffusion Model Alignment Using Direct Preference Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:52:06.521142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-13T01:08:01.321456Z digest=sha256:4704afdc4a1e84242a7e8f00a076cfea1a3f725bb6e29e0b58e8d12f0cfbfecc

Observation 79dea1e9-fe3d-45f9-99f4-b5f17dc32eb4 · inbound

$Z^2$-Sampling: Zero-Cost Zigzag Trajectories for Semantic Alignment in Diffusion Models cites this paper.

$Z^2$-Sampling: Zero-Cost Zigzag Trajectories for Semantic Alignment in Diffusion Models Diffusion Model Alignment Using Direct Preference Optimization

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:15.138064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-08T06:41:04.597012Z digest=sha256:8736325e9b3086a8f8a8099639b4f26c848224d9250c03cf2a53017085afe04a

Observation a525aeb9-c600-45cf-9e28-829ec3301625 · inbound

Diffusion Domain Expansion: Learning to Coordinate Pre-trained Diffusion Models cites this paper.

Diffusion Domain Expansion: Learning to Coordinate Pre-trained Diffusion Models Diffusion Model Alignment Using Direct Preference Optimization

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:45:20.283617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=arxiv_source observed=2026-05-25T04:42:42.564679Z digest=sha256:9d3e9cf803306712338da107583482c6f3cd01defbe2be6ef36047e91de21791

Observation cf53192e-4ab4-44f5-b574-9a1e2a7ccffc · inbound

Pluralistic-Alignment Urbanism: Operationalizing a Right to AI for Inclusive Public Space cites this paper.

Pluralistic-Alignment Urbanism: Operationalizing a Right to AI for Inclusive Public Space Diffusion Model Alignment Using Direct Preference Optimization

Reference 88

Resolution
malformed identifier
arxiv_id, observed 2026-06-30T18:55:00.095674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-30T18:53:58.717920Z digest=sha256:1a3dae7131984750f1b64393208b27150b1da592d3c3b12576ec3b5ba3e84e9c

Observation a9974c30-882d-4af3-bb61-0646710db165 · inbound

Judging to Improve: A De-biased VLM-as-3D-Judge Protocol for Single-Image 3D Generation cites this paper.

Judging to Improve: A De-biased VLM-as-3D-Judge Protocol for Single-Image 3D Generation Diffusion Model Alignment Using Direct Preference Optimization

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:09:29.495922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-26T18:25:45.000030Z digest=sha256:ffc0c7e25d4bed149a01af7cf2bacf2b86635629b2447bfebffa5850761a899f

Observation b72a5761-4e94-47c6-9d38-7e19fd9ae561 · inbound

Curvature-Adaptive Consistency Flow Matching: Autonomous Trajectory Optimization via Reinforcement Learning cites this paper.

Curvature-Adaptive Consistency Flow Matching: Autonomous Trajectory Optimization via Reinforcement Learning Diffusion Model Alignment Using Direct Preference Optimization

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:59:43.050731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-26T10:40:22.767129Z digest=sha256:3696d8df0e0600821d58971a6d66b368ca7d68fdef929ef52b2583810d4b07e1

Observation 409f87a4-9564-4c12-a330-5f5f178b9045 · inbound

Flow Reasoning Models: Scaling Reasoning Through Iterative Self-Refinement cites this paper.

Flow Reasoning Models: Scaling Reasoning Through Iterative Self-Refinement Diffusion Model Alignment Using Direct Preference Optimization

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T08:04:28.420734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-30T07:55:28.254309Z digest=sha256:24d6c3a851e0cb052b653af763e547e98e69e8ebbed9d2402aee0ba4454a0e16

Observation 050d4b37-ae3a-4772-964a-9aaec3a18bbc · inbound

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF cites this paper.

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF Diffusion Model Alignment Using Direct Preference Optimization

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:25:55.897454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-07-09T02:17:20.589485Z digest=sha256:015477deaf380bae6e0269efa7bc3996054c3207163b1bce938b659263d66373