Pith. sign in

Paper Citation Record · LEDGER

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training

As of 8 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2608.06125.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06125 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:30:34.676779Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact4
  • verified fuzzy18
  • unresolved31
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fcc56792-8cfa-43b7-8710-d72054bb261b · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:39.639842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.052946Z digest=sha256:466241d4e774efe4e72fe3fc044ab928f41e9573ef983c8a938f2455700102c2

Observation 34a208ff-42db-4735-a997-9224ed60c65b · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training High-Resolution Image Synthesis with Latent Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.131605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.131605Z digest=sha256:85dd33cb8fc769566e24745cd93204880def78636c304d562c793a3203c3cefc

Observation bdce378e-1106-410e-bce5-1769c3633294 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.226025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.226025Z digest=sha256:74cd8f9664748624aab7c29082aedc8c2d51e3d3f3f29ee332749b121230ac99

Observation d41e3a64-c0f8-4ab0-86fa-e6d840d3bda2 · outbound

This paper cites Psychological Review , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Psychological Review , volume =

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:39.417762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.329344Z digest=sha256:1aeeceeeb7d342d79106e10418f777217caaa931e7476b1373e6dba9e3683c3d

Observation 605c86c4-a1ce-481d-8e43-b24c47a30026 · outbound

This paper cites Proceedings of the 22nd International Conference on Machine Learning , pages =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 22nd International Conference on Machine Learning , pages =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:39.189641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.396935Z digest=sha256:b5ff14b86f272ce7e5ee2b11776f0ca4e2a916937b1ef30773d57a7e38e8956c

Observation 65f546c9-842e-4788-bb9b-85b27b0205dc · outbound

This paper cites ImageReward: Learning and Evaluating Human Preferences for Text-to-Image Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training ImageReward: Learning and Evaluating Human Preferences for Text-to-Image Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.456120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.456120Z digest=sha256:c40e249a30e9a1096f0d7cb9a58c3e8bcbd2bb7d61f4f8a5626ce456d33e226b

Observation 74b09334-34d3-4b02-a55c-0b6e2d1945d1 · outbound

This paper cites Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.533095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.533095Z digest=sha256:b132a3cb8e3f517949421ebd58045f531e634cc563087106a5de589131a8beb0

Observation 7ab8cc1f-5550-4ab6-88cc-446b586e8a55 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.637742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.637742Z digest=sha256:3b2876578019e169d1072aecf382f48fb1e8b2fcee77b3596ececcb3cf4cd6cb

Observation 75627252-bbce-456b-971a-37e71d79e2ad · outbound

This paper cites Learning Multi-dimensional Human Preference for Text-to-Image Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Learning Multi-dimensional Human Preference for Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.721052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.721052Z digest=sha256:9631e69f2c65c91b3e7a83708405f1177ce448f4146a1ad3667f883f0066792f

Observation 4d2c6121-0afd-460a-8d30-d1167ea97c65 · outbound

This paper cites Unified Reward Model for Multimodal Understanding and Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unified Reward Model for Multimodal Understanding and Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.776673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.776673Z digest=sha256:0c96e172268982df99227d0b467bf93ad2c58cc8895feed11c1282090d8fbea7

Observation dac5f406-1bb1-4f57-9e54-ee20d772b344 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.974589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.861670Z digest=sha256:5e31c30bd5e85261a104f35f511bcca366d09b37fc57c46a3326e56aa7957007

Observation 1742c8d1-d3df-4583-9cfd-3f3a5b6ff456 · outbound

This paper cites 2025 , month = oct, eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2025 , month = oct, eprint =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.788619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.979283Z digest=sha256:bb537a58ade8c4757109bebd2feb520c37072d0e91a997d266cfab361b367eac

Observation d36fdbc5-a0e2-46a0-9064-217064855e4b · outbound

This paper cites an unresolved cited work.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:30:38.662763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.109480Z digest=sha256:55a96a6d936beee55e335d1b39f472c751034084c7b6069c77ca72377faab643

Observation 525228fb-75d6-436c-8588-38ab1b9ead87 · outbound

This paper cites Aligning Text-to-Image Diffusion Models with Reward Backpropagation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.267180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.267180Z digest=sha256:152dc9a9201eec791df85b5b9a51b9df1eb649f255ef2ad901b78a01b5b0c96e

Observation b1864535-b1c2-4e25-9940-c31e04e0e869 · outbound

This paper cites Directly Fine-Tuning Diffusion Models on Differentiable Rewards.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Directly Fine-Tuning Diffusion Models on Differentiable Rewards

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.365332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.365332Z digest=sha256:1b6148bd97711b85c2db772837b09fd0cb9a5ab00831b3bd33b75aa0ec927339

Observation 46249ff1-5af3-4bb4-9808-753c5a8ee607 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Training Diffusion Models with Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.467681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.467681Z digest=sha256:e1e93f7605fb1f6c6a84b66f966582a1d0dd196eaa4e6117b8bf76bbc6138a8f

Observation 60c9e828-4996-433f-b5c0-f9e7e1e2bcc6 · outbound

This paper cites Diffusion Model Alignment Using Direct Preference Optimization.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Diffusion Model Alignment Using Direct Preference Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.554644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.554644Z digest=sha256:122d4aa8b9b3eb2b2af758b9dd7c3809248f2059f4da4be35f79ab179620baa0

Observation 62b913ba-aaf8-4016-bd43-820db10b7e29 · outbound

This paper cites Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.625215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.625215Z digest=sha256:1dcbb140c86d7f14d5975b8aad927f4d659f84b8c20fdd85a56db6ecaa2c75ae

Observation e2e58a6f-45e1-4da3-b9c4-e60e6811de2f · outbound

This paper cites an unresolved cited work.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:30:38.493341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.710740Z digest=sha256:dc176b4e74e86bdc1fcae95161d98dd263a85c0f60c6ea4c551d5f0ad772a4fb

Observation 6410a23a-f23d-4a72-9636-091a066eebb0 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.198546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.803901Z digest=sha256:f41218e9189c5d636d886ce04ac3ac10c061bdf1ef6702ba161da6fbe1a59292

Observation 6bec4435-493a-4f72-acc9-ed739587798a · outbound

This paper cites 2026 , eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2026 , eprint =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.040152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.908798Z digest=sha256:5653493cf5f23045760e2f9d6a2241c5d9748f45ac13485e29461a3eaff20f95

Observation 164682d0-2cef-4822-b803-1943d511c074 · outbound

This paper cites Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:30:35.757355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.955883Z digest=sha256:0e0e619b3bf40139409cea74e4e246a91a4a81ae31fb4e1210df1fb228af2285

Observation a8b9e136-6052-4aac-be8b-b7d227e24c9d · outbound

This paper cites doi:10.48550/arXiv.2509.15110 , url =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training doi:10.48550/arXiv.2509.15110 , url =

Reference 23

Resolution
verified exact
doi, observed 2026-08-07T14:30:35.495333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.048481Z digest=sha256:17efb40c70ae62476e307bc78032e17282ad59b2e3061f118292c95be6e38db8

Observation 6ffd188c-a7c9-4a84-9f06-c8573c91bab9 · outbound

This paper cites Stable Consistency Tuning: Understanding and Improving Consistency Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Stable Consistency Tuning: Understanding and Improving Consistency Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.111942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.111942Z digest=sha256:5679a454659880cffdf8fe069fa2ee6c780822393dbd25768463edf4f9647f0b

Observation 3d365ce6-7510-49f1-a357-fee585c599b3 · outbound

This paper cites Proceedings of the 43rd International Conference on Machine Learning , series =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 43rd International Conference on Machine Learning , series =

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.930730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.185688Z digest=sha256:693e528828a365620585d99c87aaaaa4ac54b68fae46195f6b00cc527a809c38

Observation f7ab9048-6927-4354-97c3-af51283e1d62 · outbound

This paper cites Probabilistic Uncertain Reward Model.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Probabilistic Uncertain Reward Model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.243421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.243421Z digest=sha256:3847f5d05b6303153ea3dd7612faaf12b88118d07a14a674b7de6a92be5660b9

Observation 0687ee4e-06d9-49bd-992a-aaf42d1bc6a1 · outbound

This paper cites Confidence-aware Reward Optimization for Fine-tuning Text-to-Image Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Confidence-aware Reward Optimization for Fine-tuning Text-to-Image Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.316738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.316738Z digest=sha256:e3b18b4913ec053f690244bbbedc3a44698ea5ec598efc443018ef880ec4ba50

Observation ac29a5b2-5e24-40bc-ae1d-9c5d6f6e57d4 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training DanceGRPO: Unleashing GRPO on Visual Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.398918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.398918Z digest=sha256:2bc18c7ce3ce5e7d173d850a324b3ddb96fba3b9d26a79a03057969ee308ac4f

Observation 668a4155-59a7-4016-9132-93354886f5f3 · outbound

This paper cites 2025 , eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2025 , eprint =

Reference 29

Resolution
verified exact
doi, observed 2026-08-07T14:30:35.247598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.488971Z digest=sha256:c8744d7b1d7214bbba2f14152b8820bf15717eeb36ce506f6dd73d5d79c83b55

Observation 39a5409c-14fe-41e6-8918-0e60cecb4cf1 · outbound

This paper cites Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.559115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.559115Z digest=sha256:00f74c8aa8862c4c5065008e440fa7943f064c4c878e3134d1c596ad3ba4b9cd

Observation d3d8effb-6dd2-4636-b25c-1492b28b7121 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.793405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.647386Z digest=sha256:d584d00218f7d380ade1dcbea10799fb56d80972d41c2d96dd79d82357e04ba4

Observation 879988d0-c2b3-4e51-a2e7-b7f5f7bee481 · outbound

This paper cites 2024 , month = jun, eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2024 , month = jun, eprint =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.631205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.723384Z digest=sha256:22c92f049f5453afa05ec1ef9c6e1167a4e091a0a28b3facffe231f56b47a2ea

Observation 093f1db7-26fc-477a-aa27-995367aefbe0 · outbound

This paper cites GenAI Arena: An Open Evaluation Platform for Generative Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training GenAI Arena: An Open Evaluation Platform for Generative Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.815096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.815096Z digest=sha256:9148e5aaf87f0970e25a430d77a29e823ded285bb917c1487d3409d67ff3d035

Observation 3d31178d-865e-48ab-8a57-c1e5802578d9 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , editor =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 41st International Conference on Machine Learning , editor =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.491424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.910082Z digest=sha256:b4cd958888f69d173e0e5ef9f9e336ddf9e0b67f184e700dacf1020ee85b62fe

Observation e9de2a5a-0541-4765-a660-9ea436d782d1 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Wan: Open and Advanced Large-Scale Video Generative Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.029242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.029242Z digest=sha256:36673863c14a4748e8ac94581602c75f2b36f37ae3fa1bcb5d1d7c6e0c508012

Observation de7feba1-e5f7-4bc1-bb14-d18b76d519d8 · outbound

This paper cites an unresolved cited work.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unresolved cited work

Reference 36

Resolution
parse uncertain
no resolver link, observed 2026-08-07T14:30:33.104439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.104439Z digest=sha256:5bbbbf8d8ac01ffb96c5143f204b5fe3f30ba50906478d93ffbf756c48269005

Observation 4b0d3a3d-131d-4f21-9407-be19600054ce · outbound

This paper cites Proceedings of the 38th International Conference on Machine Learning , editor =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 38th International Conference on Machine Learning , editor =

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.292248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:33.189617Z digest=sha256:1efe0f248a999e7d33d847eabfbf247af224c39b32d9dc5646c17c65f07a0657

Observation e1938bfe-41c2-4d62-a309-026b58834a93 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training LLaVA-OneVision: Easy Visual Task Transfer

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.254424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.254424Z digest=sha256:0bf7ca891855c88e9825b5eebd4c98bc3b1a9bf8bd16e746f1390a051aecd939

Observation bd90d109-e981-4f3d-b3c5-0e5e8a8e24da · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.330944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.330944Z digest=sha256:097bc06520374c02101d67ec4603eb2cdd9b3c80a5edd025659e4bf202bbdf2a

Observation 16b93105-1d21-4f4d-a487-57ec79cf5379 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.424109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.424109Z digest=sha256:f36842c0a11737dc9013a2403bec107aae7f170fd9f55b225f7b1309430e1061

Observation 8e1d3deb-dc79-47db-b63d-7fb3620eb0c2 · outbound

This paper cites and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle =

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.512340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.512340Z digest=sha256:e13b501259a64573c8543fb7bb3c21ad1d1de65106d1729e88cb07ff40faaa72

Observation 068eb91d-9b25-4f9e-9b12-57e25b14f66d · outbound

This paper cites Decoupled Weight Decay Regularization.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Decoupled Weight Decay Regularization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.597383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.597383Z digest=sha256:9940d3489147189c9d366d92cc1911a78ed26fb409da2cb3f85495958dc630f0

Observation 75d0641a-7f8e-4805-a35a-12536cf8ec25 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.685852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.685852Z digest=sha256:7feae2c07e8a3aa121eea6375b3ad2e7396c4f11ea0aa9b826c1030bb731ff42

Observation 19329a70-0b4e-44a8-9fb4-add83fb4bc2f · outbound

This paper cites Classifier-Free Diffusion Guidance.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Classifier-Free Diffusion Guidance

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.764819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.764819Z digest=sha256:b9ac20b14ae071805ab8ffb29f5bafdb6d7703457ae8cc7b4c4b4eb8ff42e456

Observation 67a2b587-6591-4ba9-87f3-1543910d9ced · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.872994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.872994Z digest=sha256:e5b1201efebb37f1d9753da47f31971493ada65968dee350d2364d52a86b0f8a

Observation 97c28192-3507-490b-8c5c-5d92bdfed3cd · outbound

This paper cites Proceedings of the European Conference on Computer Vision (ECCV) , pages =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the European Conference on Computer Vision (ECCV) , pages =

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.017611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:33.963745Z digest=sha256:b9d52278a3f0b2a1837a3975e65fec0097a3cabd8cb2ed20b7604fc0578bd401

Observation 1d3ff254-6f41-487b-b6f9-79c4c30317bb · outbound

This paper cites Cross-Iteration Batch Normalization.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Cross-Iteration Batch Normalization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:34.020078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:34.020078Z digest=sha256:a660dc5847e57abc6264eac719ea98c626d7b048d059a38cea561e9058960083

Observation 482463b6-3ee9-47ec-8276-9de7770b4b05 · outbound

This paper cites 2026 , doi =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2026 , doi =

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.810756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.100141Z digest=sha256:187eba2de6e1b6372e3891588f47d23161476565b32883199e0e3fb96122ce9e

Observation 1a703aef-d306-4991-acaa-4beaa279c038 · outbound

This paper cites doi:10.48550/arXiv.2509.22799 , url =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training doi:10.48550/arXiv.2509.22799 , url =

Reference 49

Resolution
verified exact
doi, observed 2026-08-07T14:30:34.947868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.176014Z digest=sha256:7735a4394d86eb37ac1a7d2939ec844032ee880352e9b065749c78c072ed1850

Observation d6a755f5-5984-4e76-a9da-6fc1902bb211 · outbound

This paper cites Advances in Neural Information Processing Systems , editor =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , editor =

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.537984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.271925Z digest=sha256:a403f1f13f3522c5d6014ff3ddcdb85e0a617d156a9ea81bf7cbf00288315b73

Observation 99dfbf8f-e87b-47d1-b8b3-6c79d4f4a4ec · outbound

This paper cites 2024 , address =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2024 , address =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:34.379568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:34.379568Z digest=sha256:db9f350a0bc6ad5c6e46ec589b93e21e3ae01c935561cad18c9eef2c4a20791a

Observation 1e314f45-3b28-498f-b935-6b4d56a929b9 · outbound

This paper cites 2025 , url =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2025 , url =

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.309374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.475798Z digest=sha256:58363ccc8b8d2ac9b632ee4f97e23684cd2c997801fc7f5b875ffc473533f54c

Observation da961fc7-1d63-4f03-b05f-856774a1b401 · outbound

This paper cites 2026 , doi =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2026 , doi =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.171789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.561203Z digest=sha256:acd0c60cf4ec42a171961c30fafe24312c4ac6e7b11473dd09c15fdc4501631d

Observation da3fd135-f024-46a7-b00b-831a737f4016 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.104353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.676779Z digest=sha256:1af9acdf0126d7c3f9f534040d8ae05efb0d48d2ad1664184bfa7bf359083b07

Pith citing papers

No inbound Pith citation observations are available.