Pith. sign in

Paper Citation Record · LEDGER

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation

As of 12 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 2 inbound Pith citation observations for arXiv:2411.14871.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14871 v3

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:54:02.564230Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T18:18:46.456767Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T18:23:50.834907Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 22725af1-acb7-4d2d-a6fd-c89201d470d2 · outbound

This paper cites Direct preference optimization with an offset.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Direct preference optimization with an offset

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.361116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.328719Z digest=sha256:93c0107db01acc5a3902057c0fff547cd691e358c3b5d965d416c1d0f3714f92

Observation 758c0570-25e7-4605-b00b-594ec78088a3 · outbound

This paper cites A general theoretical paradigm to understand learning from human prefer- ences.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation A general theoretical paradigm to understand learning from human prefer- ences

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.349920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.333011Z digest=sha256:b2f9568f8d0012d5355f80529b91f7207185639ad75522a7512db49a5a755127

Observation 782e039f-5c30-418e-8e6a-d57dbd4ec1ab · outbound

This paper cites Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.336960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.336960Z digest=sha256:c8d9303a621b61aaad980afd4e554a37747e93a8818924aa854bf300baf780a3

Observation 2b999349-b618-4dcf-aa84-00c42f6e70ad · outbound

This paper cites Training diffusion models with reinforcement learning.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Training diffusion models with reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.337692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.341510Z digest=sha256:b5e30cd84530ea776bf6438185a8cfa8a29d10fee7d1747c7e809f0b1dc2c42f

Observation a2d26ae7-6a78-43e0-9af1-c9c96ac649ba · outbound

This paper cites Align your latents: High-resolution video syn- thesis with latent diffusion models.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Align your latents: High-resolution video syn- thesis with latent diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.327456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.345351Z digest=sha256:83a76cb538bc3b14af7cc94ad23b263d985f5a42dd27cbbd8f4d68025257ee80

Observation b024a3bf-2de2-4a2f-a352-6998ee14252e · outbound

This paper cites Kwok, Ping Luo, Huchuan Lu, and Zhenguo Li.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Kwok, Ping Luo, Huchuan Lu, and Zhenguo Li

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.316856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.349389Z digest=sha256:3577cd0dd597137ed958356aaa0500d09b1143262fa76f5643179375ad4905dd

Observation 977027d9-7eb1-4016-957b-b518703b7e01 · outbound

This paper cites Self-play fine-tuning converts weak language models to strong language models.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Self-play fine-tuning converts weak language models to strong language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.306807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.353167Z digest=sha256:a794d4c30c4248e78716a9f31c03669e962272d321e66f790f753baf08a58c78

Observation bac95065-43eb-497d-8324-21ca538c3fcd · outbound

This paper cites Christiano, Jan Leike, Tom B.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Christiano, Jan Leike, Tom B

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.296275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.357155Z digest=sha256:5895056c3a76373539419439b3b3fd432b5cb890a3ac692fb18e91a8ad2684d0

Observation 0ec493cb-2a78-4ba7-b681-a34471d88dec · outbound

This paper cites Diffu- sion models beat gans on image synthesis.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Diffu- sion models beat gans on image synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.284796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.360853Z digest=sha256:cc20c584a883ca165bd530ced195ac91a0d1b79ce585157b98fbedfc7e017677

Observation f928534c-1c43-4ec2-864a-e6c50512e18e · outbound

This paper cites Scaling rectified flow transformers for high- resolution image synthesis.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Scaling rectified flow transformers for high- resolution image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.273670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.364610Z digest=sha256:d7f8fb9f1b0165a3c0257ebce6b8a0e2371153f904fa3ae6161145a3d54f73db

Observation 36ffea07-be43-4871-96e4-e8a918ea1218 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation KTO: Model Alignment as Prospect Theoretic Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.368175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.368175Z digest=sha256:6d25ee08833a0527d1d6425e77c30c1f0e38c710b42fabbaa5ae270a30b14c55

Observation fbf4ed89-55b6-4bb7-ae79-6042745a5556 · outbound

This paper cites DPOK: reinforcement learning for fine-tuning text-to-image diffusion models.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation DPOK: reinforcement learning for fine-tuning text-to-image diffusion models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.262275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.372209Z digest=sha256:42a1e660c554273d1141bada63cdc8062bdffc5401dbd46ba32cad0c58c2eb95

Observation 4dd6b9a1-9c3f-448c-b2d1-a494f4c22d5f · outbound

This paper cites Scaling laws for reward model overoptimization.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Scaling laws for reward model overoptimization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.250381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.375831Z digest=sha256:f26ea4ce11a5d57b48a7a2becf0492774afd9d7dad7a82172d81cb326b86d816

Observation f405a09f-a8b8-4f0c-b4f1-24aedc49ec4e · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.239054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.379462Z digest=sha256:ccd1c44faae010b6ebcb65a296d200acf826695bfe8e62314a1b3975597f7b2b

Observation f5386393-9f77-45d0-a062-e91a442a7eaa · outbound

This paper cites Brandt, and Tomer Michaeli.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Brandt, and Tomer Michaeli

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.225438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.382998Z digest=sha256:f09f6f4c139efe713948846f89f59389386ab474c56b5ed6b0011009d1bc519f

Observation 0fd283bb-8947-4ec0-a147-71c10288706d · outbound

This paper cites Denoising diffusion probabilistic models.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Denoising diffusion probabilistic models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.212698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.386341Z digest=sha256:c2023e2b1244f566aecf14092cc53b8f1292c62dac4b663f2b74c4ca3f54666d

Observation 7ed99fb3-f4b9-4cb6-b824-6b90756f9817 · outbound

This paper cites Estimation of non-normalized statis- tical models by score matching.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Estimation of non-normalized statis- tical models by score matching

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.201515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.389563Z digest=sha256:f91fb2042293260d04e4466ed86f8f230ada57ecf4a200069f5fd50224cc4d35

Observation 056ea50f-2def-42a5-b416-ba63338c3fbf · outbound

This paper cites Reward learning from human preferences and demonstrations in atari.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Reward learning from human preferences and demonstrations in atari

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.190321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.393106Z digest=sha256:9049d67a0ad1b84f06144ee6a20c08da06823a24f4a5dc0594a7da1306098582

Observation 7cde6093-e2e3-4bba-999e-833d799616a2 · outbound

This paper cites Ryzhakov, Andrei Chertkov, and Ivan V.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Ryzhakov, Andrei Chertkov, and Ivan V

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.178944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.396691Z digest=sha256:2dc778df23369950eab2842d15532b09385d92686fdfe7b3eb07fbb22d66d31d

Observation aecd68b4-bbf9-4011-8345-bf2be980306c · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Pick-a-pic: An open dataset of user preferences for text-to-image generation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.167644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.400628Z digest=sha256:b4b3515005f504473d5cadbd43222e4b92176e497da6e6b655bd19cde8428499

Observation 3832647f-04d5-48aa-aa2f-e45e5d773dd8 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Pick-a-pic: An open dataset of user preferences for text-to-image generation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.155479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.404196Z digest=sha256:da41c595f679e222f28fb8b02e78f3044ca6d7f196b9914ba961b2b93c646192

Observation b884d6d8-9432-4400-bca9-162639bd102b · outbound

This paper cites Bradley Knox and Peter Stone.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Bradley Knox and Peter Stone

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.144654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.408019Z digest=sha256:7bfd9e26f64ba7df6d01b7ca82184246f5d936e2152bbafa3a3348e81b785055

Observation 590c7101-782e-441d-a9f4-6deb5e9fec30 · outbound

This paper cites Dif- fusion models already have A semantic latent space.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Dif- fusion models already have A semantic latent space

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.133590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.411697Z digest=sha256:37b1c77cf1b9b1503ef31ea7a52daf35ecae40f9cbc5fff55abe1d9e3337a52d

Observation b436b6c6-7288-4fef-ba06-b6de3f4005e6 · outbound

This paper cites Aligning Diffusion Models by Optimizing Human Utility.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Aligning Diffusion Models by Optimizing Human Utility

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.415254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.415254Z digest=sha256:79f4c7b4518e0ad8ef94b0afd0b7e71d03731129377ffc92083a00b0fdae033f

Observation c77e5e8a-10a2-4e8d-ad7e-01794de62aae · outbound

This paper cites Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.419272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.419272Z digest=sha256:2827930f2869cdcdb46588085e785569572543bd9391a1f24a39dd39ea190797

Observation 7c60fc4c-8778-483e-bb79-a872fd7fb5c3 · outbound

This paper cites Let’s verify step by step.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Let’s verify step by step

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.123512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.423053Z digest=sha256:6c2ff6566fa5d487f94d2860dd8e86c21de68646d85f75f434d73a4d326b12a6

Observation 78f53924-dca5-4739-b167-d2751c617350 · outbound

This paper cites Lillicrap, Jonathan J.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Lillicrap, Jonathan J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.113164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.426479Z digest=sha256:59c5a24f3f1785ffe7733ec35d00b098d3ca6056e446874bac76ee9e3556c0e2

Observation 77126f30-7f1f-424c-8a20-51262ab6e257 · outbound

This paper cites Alignment of diffusion models: Fun- damentals, challenges, and future.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Alignment of diffusion models: Fun- damentals, challenges, and future

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.430101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.430101Z digest=sha256:876a6bad03101b952929bc6c955aa8df02a864d44b7ac23abe6a9f59b4b01c55

Observation bb3b29a4-363d-41e5-a3be-f53fd5a16684 · outbound

This paper cites Interpretation and generalization of score matching.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Interpretation and generalization of score matching

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.101748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.433620Z digest=sha256:840da523177310752d9293c20a98b68fac5ad72eff558b87ef2d62887f5720ca

Observation e39b0d5b-9dfe-46f8-9c06-104af849cfbf · outbound

This paper cites Ho, Robert Tyler Loftin, Bei Peng, Guan Wang, David L.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Ho, Robert Tyler Loftin, Bei Peng, Guan Wang, David L

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.090375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.437239Z digest=sha256:48c27603361032902fa76a6c56fdb2cf12d56b37c90ed79c6d12b47531b2562d

Observation e95140c4-4e3a-4518-8ae4-c28102384ea7 · outbound

This paper cites Simpo: Simple preference optimization with a reference-free reward.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Simpo: Simple preference optimization with a reference-free reward

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.078511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.441028Z digest=sha256:f463a6e92808bd12c2bb072274895bd5edd669544db2d75fa77b57e0fcdd37fd

Observation 200028e2-5bd6-4e26-bcfc-ac7eedcff526 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Playing Atari with Deep Reinforcement Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.444437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.444437Z digest=sha256:eba883b3d94485da1cf6a72fe41b4f5aa0030c41a4c3192692bd3e1301daaa42

Observation ca13e3af-9670-479c-9189-6df280cc26ba · outbound

This paper cites Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.066456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.448722Z digest=sha256:8bedf4f7df3caa85d32e81eeec69bd5a8e501e748456a3997b305af50432a09b

Observation 45738fa3-cb80-498a-9e4b-6e39974ca1fc · outbound

This paper cites an unresolved cited work.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:54:03.054846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.452354Z digest=sha256:1692835963a4ca47e4f12daf2bc64253fd487150c13550c03d51464f36ec2b01

Observation 26eb8468-259e-4d7e-9703-d63950bece0a · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.456141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.456141Z digest=sha256:00f7563c41c0d100abb1ed0ce6535b35620e6f364bf40b1bbe61781b841d8814

Observation cd834621-b559-4394-b718-9ad409ca307c · outbound

This paper cites Learning transferable visual models from natural language supervision.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Learning transferable visual models from natural language supervision

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.042107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.460432Z digest=sha256:e9d8a2fe35e674a224d2e4da2eedb85edcec9464124733157d132cf0fd7b8fe6

Observation c30638c4-523f-41f1-b562-b0ab06fb12a4 · outbound

This paper cites Manning, Stefano Ermon, and Chelsea Finn.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Manning, Stefano Ermon, and Chelsea Finn

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.030862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.464075Z digest=sha256:2b53e2474f24bee15e718b1ef84b15d02964c917dd2b4639ba9419b49321bac6

Observation 9087a192-90d2-47f3-8c34-8827eff92f0f · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation High-resolution image synthesis with latent diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.019811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.467621Z digest=sha256:7a9eeab78cd4d4ca848e58776d43bb086ce4825a92df961e98e849c833b33827

Observation 62b9beb5-4704-4174-b168-8ce82ef453e4 · outbound

This paper cites Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.471071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.471071Z digest=sha256:fe5623c187df2b24ab7c176c014fca87b654c9914f653d14be9df425b6d7540f

Observation e5f3d85d-3c8b-4930-b489-719eb29f2f50 · outbound

This paper cites Jordan, and Philipp Moritz.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Jordan, and Philipp Moritz

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:03.007942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.475762Z digest=sha256:b599f9fd7f01dec61959315205affcfce56db9a1c0b27c180b933edde5e3f2a3

Observation ea3662ea-bf7e-4387-9349-092eefed81cc · outbound

This paper cites Proximal Policy Optimization Algorithms.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Proximal Policy Optimization Algorithms

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.479171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.479171Z digest=sha256:a83e3b0c070385d8e755e8dc37ff54bf1e923e0e6c08ca234196a18dbc40123e

Observation 65a1d783-c4c5-4063-be56-8316d973431f · outbound

This paper cites Denoising diffusion models on model-based latent space.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Denoising diffusion models on model-based latent space

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.997501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.483369Z digest=sha256:f1f92859d24d05dc66ec108d3dbb1c90e87af7e17485f2932e860d0bb9862943

Observation 65bb36dc-fe03-4a92-b4c8-cc40f26a37e0 · outbound

This paper cites Riedmiller.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Riedmiller

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.985577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.486970Z digest=sha256:8d007b7ae7b19e5abc44cdf4c28bb78cedc237f1169d7f5a1e0f432c2e086f33

Observation f3d2bf19-de5e-4fd2-bc05-c103e2f9413d · outbound

This paper cites Weiss, Niru Mah- eswaranathan, and Surya Ganguli.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Weiss, Niru Mah- eswaranathan, and Surya Ganguli

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.974017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.490776Z digest=sha256:e0251d2e04a24e2b5fa8d205014c126f40e1ada2210cb82098f2e541c0a9728e

Observation dbf00f3e-f92a-4071-b3ad-45c6f05b7845 · outbound

This paper cites Prefer- ence ranking optimization for human alignment.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Prefer- ence ranking optimization for human alignment

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.960541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.494331Z digest=sha256:685701cead4a5c33a476f1540ecb5b0470fa960bb7c1c5f2d979e15637c82e71

Observation 74007324-9572-4c4a-a723-9014230ea04b · outbound

This paper cites De- noising diffusion implicit models.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation De- noising diffusion implicit models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.949774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.497907Z digest=sha256:87085065939ca520d064c25c07b1673b21357d3135b496427700e4835b6c9466

Observation 699b4337-a8ff-4baf-a400-5005f27dd5da · outbound

This paper cites Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.939942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.501599Z digest=sha256:5bf96d78d82bccad819c81c4002af8278409581e94de5108bc3d0c002ae79537

Observation e71875a0-0ec2-4b2c-b1c2-4c31cca034d3 · outbound

This paper cites Ziegler, Ryan Lowe, Chelsea V oss, Alec Radford, Dario Amodei, and Paul F.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Ziegler, Ryan Lowe, Chelsea V oss, Alec Radford, Dario Amodei, and Paul F

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.929645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.505121Z digest=sha256:bb65e72d25614117c25d1c03ade2e1b99625c57c693c18069d9d134e88139a14

Observation 9f894b4a-0b23-46ff-80dc-3688956261a3 · outbound

This paper cites Sutton and Andrew G.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Sutton and Andrew G

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.918825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.508614Z digest=sha256:5fe68c3f835a8d62b7adbb203a3389d89c1f6ef509d4f16e0a40ab8af39ce2d5

Observation b0ffe050-9b4a-4c6f-8e4d-36a5a2812329 · outbound

This paper cites A connection between score matching and denoising autoencoders.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation A connection between score matching and denoising autoencoders

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.907950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.512300Z digest=sha256:341ee2753f4ea2a53109ba4b9a3d3bc0aae87c64684e02c36c3d99c934e2ebbe

Observation 00ec18d6-3d41-43a9-ba98-3860f67799ff · outbound

This paper cites Diffusion model alignment using direct preference op- timization.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Diffusion model alignment using direct preference op- timization

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.897621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.516015Z digest=sha256:7b48cda28705c1098ead973544d267c3ce9d6f16d5846d3f8ed228a9caa76feb

Observation 43d4b163-c4ba-45ed-beb8-06a5d4369a6a · outbound

This paper cites Aligning Large Language Models with Human: A Survey.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Aligning Large Language Models with Human: A Survey

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.519982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.519982Z digest=sha256:242b7c5e5521f15cd3c5c3c6a80aec45c403cc73aba6f8542bd8691f365f2b44

Observation a8a6af0c-b7eb-4f3c-9282-bd5bd90e023e · outbound

This paper cites Dueling network architectures for deep reinforcement learning.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Dueling network architectures for deep reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.886959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.523781Z digest=sha256:48bf448d1819bf5b2449d8bd4e10c6d5aee812ba5e55628068c45a017a85808d

Observation fae401ae-e706-462a-83c9-3e53b52aba3a · outbound

This paper cites Waytowich, Vernon Lawh- ern, and Peter Stone.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Waytowich, Vernon Lawh- ern, and Peter Stone

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.875924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.527351Z digest=sha256:adbd35bf3b6992025dd2399de4013cf1e8ce818e6c5b75007525491f51614662

Observation 17af7ddc-68cc-48ca-a6e1-43804bed4ab4 · outbound

This paper cites β-dpo: Direct preference optimization with dynamic β.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation β-dpo: Direct preference optimization with dynamic β

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.864885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.531043Z digest=sha256:d544797759e514160f536451d62008d61e76c64fc87c3b5cd24f02fd792bf399

Observation ed101848-77b0-44db-a3fb-01f7ef36eb0c · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.534776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.534776Z digest=sha256:111bf9cbc6a9020f67a88ae256cdc8af9d6a8f21650a64f5daab4dd3ef04935d

Observation 5cef9c92-4497-4018-9253-8d9bfb2eb68d · outbound

This paper cites Using human feedback to fine-tune diffusion models without any reward model.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Using human feedback to fine-tune diffusion models without any reward model

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.853808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.538484Z digest=sha256:be900a421bd12661ecec80bd18c74b5f539d87f8a707a29cc8cc9da1a794c3de

Observation 53cef9b3-aff9-4609-86b0-0d88016a0b12 · outbound

This paper cites Diffusion models: A comprehen- sive survey of methods and applications.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Diffusion models: A comprehen- sive survey of methods and applications

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.841657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.541992Z digest=sha256:3a269a8f228c9d92aaca8fa5b3726bc01e62bc157a50add31e61ef5084e2f166

Observation 8b9fbbf3-1f98-450b-9fd6-2178a3f15a72 · outbound

This paper cites A dense reward view on aligning text-to-image diffusion with preference.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation A dense reward view on aligning text-to-image diffusion with preference

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.830317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.545504Z digest=sha256:a308f043be70ccc83bcdcdd4fb5cc7fb399f2084ed9ede5bee7812d29d0b4aad

Observation c2b5b6be-1713-44b7-9968-af7c7e415bc3 · outbound

This paper cites RRHF: Rank Responses to Align Language Models with Human Feedback without tears.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation RRHF: Rank Responses to Align Language Models with Human Feedback without tears

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.549137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.549137Z digest=sha256:52f409ce4cb0e66595037cfc0f14c8db74415f50aaed25d08aad509ddcc468ed

Observation 38561138-6703-4579-9d18-eb5170f7a1ef · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Fine-Tuning Language Models from Human Preferences

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.552889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.552889Z digest=sha256:e6f18594ba9f3d8fa43a233fc74c86fa3ac89ecc7368dfdc41b1beb87f1f6ac9

Observation 22bceab4-0537-47e5-bb51-3cae2fefbeb1 · outbound

This paper cites Derivation of the Loss Function Defined in Eq.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Derivation of the Loss Function Defined in Eq

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.819504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.556868Z digest=sha256:3676b0a88d3895050082a06eb036ad1fda5113a5010840383093b731527f3b53

Observation 584f6e90-b1a3-4f74-a811-2ac38b14fd7e · outbound

This paper cites an unresolved cited work.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:54:02.807475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.560578Z digest=sha256:5a46a4dac41f4eec5eca757d3d28d0446384963e99f59c49963331165f70ce8f

Observation edc19149-b908-4958-8d5d-9ec5f162aea6 · outbound

This paper cites Implementation Details We employ a constant learning rate with a warm-up sched- ule, finalizing at 2.05 × 10−5.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Implementation Details We employ a constant learning rate with a warm-up sched- ule, finalizing at 2.05 × 10−5

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:54:02.795707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:54:02.564230Z digest=sha256:6c1fd997d18e3ea14d46a16c172c3fa208166640863f659f58af3dfc9991bf7c

Pith citing papers

Observation 129b5c4a-d693-4153-bdd2-f3a363f0aa61 · inbound

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs cites this paper.

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:15.491627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-12T02:10:27.595446Z digest=sha256:5e7db0bf267a94b5803c4b1923ba6325f8ff2558dd2ffaf5b741736909f15565

Observation 794a93ca-94ab-4678-b4fc-017d4e8955e6 · inbound

Explicit Critic Guidance for Aligning Diffusion Models cites this paper.

Explicit Critic Guidance for Aligning Diffusion Models Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.836500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-29T18:18:46.456767Z digest=sha256:3979c2933724525038ebe81acb25a37ae749704b3743138faebb61f128e0935e