Pith. sign in

Paper Citation Record · LEDGER

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models

As of 4 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2605.06070.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.06070 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T14:16:55.945455Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact12
  • verified fuzzy31
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f5b3b81-1e4c-4364-b325-10d6827dc961 · outbound

This paper cites A general theoretical paradigm to understand learning from human preferences.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models A general theoretical paradigm to understand learning from human preferences

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.412722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:e3379039df48e1fbb6d6b01a4ecf7c972bbcc55a34a93b0058c62334ae68683c

Observation ec1b893c-abd4-4653-afb9-bb44f553707b · outbound

This paper cites Improving image generation with better captions.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Improving image generation with better captions

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.388935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:2dbec5940a8b8577744141b9b7c1db2ecf0f5b2581916a99c758fb7699bd99d2

Observation 5521c809-cbdb-445d-b464-69a977195e03 · outbound

This paper cites Flux.https://github.com/black-forest-labs/flux.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Flux.https://github.com/black-forest-labs/flux

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.404342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:4e76d08294bb3ea912827a9bea1ad4524398a62c963f0650bee4cd71cbb6c137

Observation 68e4d14d-be81-419d-bead-e515744f43c5 · outbound

This paper cites Rank analysis of incomplete block designs: I.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Rank analysis of incomplete block designs: I

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.397283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:d88367b8bc2b5380b65f301e409b5c69d796c0a2aeaf3deffde35414b265eb42

Observation d04aa348-0a63-4502-83fd-894967313d06 · outbound

This paper cites Getting it right: Improving spatial consistency in text-to-image models.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Getting it right: Improving spatial consistency in text-to-image models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.384728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:b93631ec3dfc4cf5ff4aa0a3c5571c057b494c037d3b8bbe8a37f545f3411192

Observation dc56eb4a-9228-4224-9c1e-971487087908 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:38:53.511834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:03f7127f720f3147e319cc37d4719d7fa7ac0b6e0d2d56b12946a9c22bc0e0a5

Observation db8749f2-ac0a-4512-807f-ed314c8112d1 · outbound

This paper cites Deep reinforcement learning from human preferences.Advances in neural information processing systems, 30.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Deep reinforcement learning from human preferences.Advances in neural information processing systems, 30

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.393260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:69a7e50a32eaac881f64dbed2e0af4d508fcf711365b8c33f6b5aa5f50de873e

Observation ebb6822d-9445-4050-9fdd-4ba2ef133b2e · outbound

This paper cites Diffusion models in vision: A survey.IEEE transactions on pattern analysis and machine intelligence, 45 (9):10850–10869.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Diffusion models in vision: A survey.IEEE transactions on pattern analysis and machine intelligence, 45 (9):10850–10869

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.400942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:deef9d25882f30764b2e7965c14a071c44c0a519edbe6a22560a1f15654d27d8

Observation 4a33e5ab-48d6-4ad1-bb9e-295e63f2edaf · outbound

This paper cites Diffusion-sdpo: Safeguarded direct preference optimization for diffusion models.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Diffusion-sdpo: Safeguarded direct preference optimization for diffusion models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:12.215672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:944489a2c23758b551bdcff5062e4d004e5cb65b0de5e35bd5fc2fe5894daab5

Observation d3f91091-249c-4662-8fd7-2456a092b47e · outbound

This paper cites Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.409142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:b7d784f7647c24225edf28a46e485b01054afc1fef5f4fe7b2640ff4a69a5d5f

Observation f72747d1-10f4-4996-ba31-33669d44ee40 · outbound

This paper cites Margin- aware preference optimization for aligning diffusion models without reference.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Margin- aware preference optimization for aligning diffusion models without reference

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.455564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:5b9fcbbc1168d578d47a85d6f917f918d39f0b3a1fa3f7818fbd920ff84f2d86

Observation 6bc1ae61-9448-40f4-8d10-7239fdf29003 · outbound

This paper cites John wiley & sons.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models John wiley & sons

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.451822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:fdcec18c4ed976f87163024230d5146582f55b0ae05194011bee2a0b07e527fe

Observation 0be4d1ea-602b-409a-8ebe-8040a316057c · outbound

This paper cites Elucidating the design space of diffusion-based generative models.Advances in neural information processing systems, 35: 26565–26577.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Elucidating the design space of diffusion-based generative models.Advances in neural information processing systems, 35: 26565–26577

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.433742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:cdd947c75dcc305381e9b860c6ea3b27b256e276a3f53f1b0297ddd0936c3a72

Observation 4658ae27-2910-4f83-9f5e-2b30728b10c7 · outbound

This paper cites Scalable ranked preference optimization for text-to-image generation.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Scalable ranked preference optimization for text-to-image generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.444833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:228e181d8f2b519c3ebc505c1ba347e3ecf5a9e4b4c09a35a3db79906e410f0a

Observation 39b3ad83-2438-436b-a699-e9e4574be719 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advances in neural information processing systems, 36:36652–36663.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advances in neural information processing systems, 36:36652–36663

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.469390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:a04fb99f937ea14bd34da824b6640093c2ec57f22b26456b1deacfa86d610de1

Observation 8a743779-e94e-4605-96c6-ffd7498d8458 · outbound

This paper cites arXiv preprint arXiv:2507.07510 , year=.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models arXiv preprint arXiv:2507.07510 , year=

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:12.151668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:f4d0d7695868434ff158c4ac084615b0e73cd9dc53bac92ae48d8d14ebec951b

Observation 74d5a422-4e92-474a-9d2a-3b536359e33b · outbound

This paper cites Align- ing diffusion models by optimizing human utility.Advances in Neural Information Processing Systems, 37:24897–24925.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Align- ing diffusion models by optimizing human utility.Advances in Neural Information Processing Systems, 37:24897–24925

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.462985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:77e700616f529d6cbd91bb88b633333090e8460d10bd2d144eb70f09d5e8cca2

Observation cc9bc20c-2e28-4861-a575-c6b2e3f79480 · outbound

This paper cites K-sort eval: Efficient preference evaluation for visual genera- tion via corrected vlm-as-a-judge.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models K-sort eval: Efficient preference evaluation for visual genera- tion via corrected vlm-as-a-judge

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.466264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:e9dcc83da3a71b987b7bdcf228172fe9a3331e10269f956d18e3678b84e1870c

Observation c5726001-85b7-48c0-ae9c-d61204f8121d · outbound

This paper cites K-sort arena: Efficient and reliable benchmarking for generative models via k-wise human preferences.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models K-sort arena: Efficient and reliable benchmarking for generative models via k-wise human preferences

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.440740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:d3cf46206180e4c0d84f3983a81cc365cc5b5cdac58d30e58f726dd132b3b00d

Observation 42836f56-de22-4aab-a351-6116e3089ba9 · outbound

This paper cites Flow Matching for Generative Modeling.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Flow Matching for Generative Modeling

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:41:12.145120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:1f7057f0ceaef1b8ed7d56ce9fcfaaf70d83563a977c1f42877043d1e8463c95

Observation 6054ba67-3f2f-4386-b8aa-7c3393b94121 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:41:12.158173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:699d972216374904e390f2895386dd35b9faaa04c23df338786972150e17bfa2

Observation dfb3a2a3-8f6f-478a-8baa-fc43ffc0d84f · outbound

This paper cites Hpsv3: Towards wide-spectrum hu- man preference score.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Hpsv3: Towards wide-spectrum hu- man preference score

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.437527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:61dc5eadc46b5f49ecca46cd51c9ef8008f8a3487824b8e9043c1091ba856fd2

Observation 3a5b050b-1d52-4000-a131-5b81e3e749ee · outbound

This paper cites Scalable diffusion models with transformers.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Scalable diffusion models with transformers

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.459345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:adb3e564b8c9aaf5e1a181786bf1cd8263df2ff52fc52f420e0eaf829ff7d77c

Observation fd4e28b6-614d-4349-8cb6-605d89cdb507 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:41:12.191002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:fd9a646d0eac32427aa8641d646328e795bdd9c006e53c9dc04748a0be0c4aed

Observation 45aa5059-00e1-4323-896d-c3d1d789cbb4 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Learning transferable visual models from natural language supervision

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.430478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:6ac4c87504766981316e9876fa8594f647dee6061610e55c9dc9426253412f8a

Observation d98c7217-949e-41a8-bce8-9e51dcca37a5 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Direct preference optimization: Your language model is secretly a reward model

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.419907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:33ea434e70223da61c2629e732ce8c3d4a6c8591c94bb7495feda370ab0c2fda

Observation a05ffff9-9e02-43a7-9423-35f9a95a0df3 · outbound

This paper cites High- resolution image synthesis with latent diffusion models.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models High- resolution image synthesis with latent diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.423678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:128ded426254327076ded7caa67541ac007da43bfded53edfdb90dc9434bc40f

Observation d6c7e423-e72c-48c7-8f23-cc9bd4ec895d · outbound

This paper cites LAION-Aesthetics: Predicting the aesthetic quality of images.https://laion.ai/blog/laion-aesthetics/.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models LAION-Aesthetics: Predicting the aesthetic quality of images.https://laion.ai/blog/laion-aesthetics/

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.426900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:fbd44cc47da996176c27a52f9c43674be518cbb21f5df15a7c2e8b93a3a17e74

Observation 94d33e79-16fb-4d24-bfa0-aa38d920690c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Proximal Policy Optimization Algorithms

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:41:12.196535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:bf1a95bd1490042b572325639bddbbf6a70f49bebb37bc596f2ebd20cf31860b

Observation 3410b3e5-293f-4f09-a933-a54a2c7cf25c · outbound

This paper cites Freeu: Free lunch in diffusion u-net.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Freeu: Free lunch in diffusion u-net

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.472677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:934d242703ea2d891c809a28579f75c92174d18f08c1e1e02b15331655bfc535

Observation 659d5408-c724-4187-8d8c-83cf3ac29f6a · outbound

This paper cites Generative modeling by estimating gradients of the data distribution.Advances in neural information processing systems, 32.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Generative modeling by estimating gradients of the data distribution.Advances in neural information processing systems, 32

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.476268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:97150050a067ff3c05a83823cc263cd67c3bf76fa085cdf11a1b9c27bdac4bd8

Observation 25e42aa1-2447-4e5c-875c-58a9fa9b1a3f · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Score-Based Generative Modeling through Stochastic Differential Equations

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:41:12.207170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:2408ddefb7356977fe18fd27d10c79e2342ecb838ad5b680fa652ee376b971e5

Observation 9f65322f-d660-4a04-91f1-483c959e1e97 · outbound

This paper cites Learning to summarize with human feedback.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Learning to summarize with human feedback

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.481687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:a39eb1c6dd850abb0ad4e9e455bdfa10612657749bd9230716200f36657ef3d1

Observation e4309012-79f7-4679-90f5-52099cccacdd · outbound

This paper cites A law of comparative judgment.Psychological Review, 34(4):273–286.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models A law of comparative judgment.Psychological Review, 34(4):273–286

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.484837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:c188de794bfc52942b2f37ac197e1d4fceaef0cbb1de15dcc4445bb34283c68c

Observation 4d361d38-9bcc-4b4a-8c77-e76f6b6034e8 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Diffusion model alignment using direct preference optimization

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.491791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:4787a09fb48ec252602d88cb7991f12b5e88ce4725a10678532c69e1a859efa9

Observation 1ed71ad3-e810-46f1-be42-3f64e689b482 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:41:12.201925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:72630461ba24b7530b1af9c0533cc32d1d061acda1070e812d07771be9409012

Observation 308c885c-0c13-4f08-9fd2-76e43774f0ad · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.495686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:480afa3bea1859c08ab57d839ee26995b1fd1150cad46745d6d6caaf9a56cc22

Observation 619d2178-fad2-404c-a8d8-26456627a8a4 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models DanceGRPO: Unleashing GRPO on Visual Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:28:29.121023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:6220c4ead38ddf83d012d01f75ec2281af5ffd94346f1aac8bf0cba7c15d80f5

Observation 7c1b180c-1c1b-4319-b3a8-56410e42991d · outbound

This paper cites Diffusion models: A comprehensive survey of methods and applications.ACM computing surveys, 56(4):1–39.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Diffusion models: A comprehensive survey of methods and applications.ACM computing surveys, 56(4):1–39

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.448740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:dc5fb06ee54e9fca870a80da116418e9cdfebb6885aae4ce7461f971c634ccba

Observation ad1800b5-4116-404a-88ee-ffe54c02184d · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:49:31.747550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:0110954b12003a023bf441b70e11afb46f26ccc4d5ace620938ea086d9660fe8

Observation a5238165-3f2e-4b53-8f5b-47f134f229a0 · outbound

This paper cites Self-rewarding language models.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Self-rewarding language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.416137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:6e63c7f8302916c951a7288f2fc25d103ebeb190213bed9a1d7898dda53af465

Observation ce1a2567-d07c-4980-970c-b64a3f6ad9ef · outbound

This paper cites Dspo: Direct score preference optimization for diffusion model alignment.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models Dspo: Direct score preference optimization for diffusion model alignment

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:42:15.487958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:e6f19eaf32cda6a5d6265ba4a9a880782101437d05caacb8361ba3f7295ff511

Observation 5e79b816-d7ec-4ae6-bdf2-2448f1997217 · outbound

This paper cites we did not observe stable improvement when training from public implementations on Pick-a-Pic.

Arena as Offline Reward: Efficient Fine-Grained Preference Optimization for Diffusion Models we did not observe stable improvement when training from public implementations on Pick-a-Pic

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:12.167678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T14:16:55.945455Z digest=sha256:33e4a7ec0cc42a68560a05a11c811974711c51fef0fdd783adfbf8dfc2e72e05

Pith citing papers

No inbound Pith citation observations are available.