Pith. sign in

Paper Citation Record · LEDGER

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models

As of 16 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2605.26013.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.26013 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T22:44:25.152451Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact13
  • verified fuzzy0
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a36f42a4-0278-4470-847c-ab4a5a28a551 · outbound

This paper cites Stochastic Interpolants: A Unifying Framework for Flows and Diffusions.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Stochastic Interpolants: A Unifying Framework for Flows and Diffusions

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.614664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:9bf503d5eb7cbd579c25775c2b8f102040bd6df2835b47d4d0e7f22bd5dfd218

Observation d83a728d-623f-4bc2-b1ee-0baaaeea8ce5 · outbound

This paper cites Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.620053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:750aa8a91026d43d5ec503c339ba85433c30e1bd1f6c1ef3360b94d95acbe10b

Observation 195e1f9d-4ada-425c-bf82-1066156f70db · outbound

This paper cites Fine-tuning flow matching generative models with intermediate feedback.arXiv preprint arXiv:2510.18072, 2025a.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Fine-tuning flow matching generative models with intermediate feedback.arXiv preprint arXiv:2510.18072, 2025a

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.612457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:34a3376bc02f6da0415fc2892e2242c022acef947002418d6e1f0887ef5c86ee

Observation 4a557aa6-a763-4854-a3a6-bfba55cb2f24 · outbound

This paper cites TempFlow-GRPO: When Timing Matters for GRPO in Flow Models.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.617374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:f8537c1455d59baa342ef3f15b229a093d51338285689e15bf9b9eeef461ec60

Observation ba1d0972-f01e-4102-a6de-ae0d8986d4d9 · outbound

This paper cites CLIPScore: A reference-free evaluation metric for image captioning.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models CLIPScore: A reference-free evaluation metric for image captioning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T22:44:25.152451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:dbcbd6f8437157a2fe3866a09aee07399e2a372b223b22fb1246c6c31e3f9c8b

Observation 4e86341b-76a5-4802-990f-6c6a485c0651 · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Aligning Text-to-Image Models using Human Feedback

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.622449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:baccf70bec7ba9cd9bced97e08ce4534ed959e9a7d30759d26b67a8a8ed99af3

Observation b4459e72-fced-4a42-a3fc-3296ee2c0b17 · outbound

This paper cites MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.609833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:f82934fdeee066361fce80a5efec3a42e61f912ea3153807b3223e589624b651

Observation e0616927-8ee7-4111-930b-8bc43d153589 · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Flow-GRPO: Training Flow Matching Models via Online RL

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.606928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:d1c0b7506946ee0ce57538d7aed962f735427e86111f4c3e439d8e9d0f6d2f04

Observation a64017b0-2a58-48de-a8bd-7b78e7f43515 · outbound

This paper cites Aligning Text-to-Image Diffusion Models with Reward Backpropagation.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.604475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:5924b1913fc62fe3597c811a26c45cdf8bf5ff8f7c0ba5ff6f7de282bd28265f

Observation 87fe6953-3c25-45ca-8ef7-e64179792c6c · outbound

This paper cites Stepwise credit assignment for grpo on flow- matching models.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Stepwise credit assignment for grpo on flow- matching models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.596226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:93a3480e4e7d5eda1ca10381a440a84afb1e2577c8ab47cf43717d4884af1a51

Observation fc3c30c5-b2e3-432d-b3f0-c48a7cb8423f · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.590277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:e0cca93709ce9bbae2c9a726210ee5b6bd03a61fe38c5946af1a5dab5c0ad7ce

Observation d81c8b19-71a5-4086-bdf2-8c03a87b9b62 · outbound

This paper cites Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.593464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:3207d8d75629a0b971deeaa492c94ac2f6054031784685cf1683c8b7768dacce

Observation a9e414b3-e654-4d3e-b1e9-b5345b9c9b03 · outbound

This paper cites Advantage weighted matching: Aligning rl with pretraining in diffusion models.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Advantage weighted matching: Aligning rl with pretraining in diffusion models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.599110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:4696a3f8482ed3cdb36ae01a289ac11a0f2ad5aa31a89e082055066a3cd90891

Observation 6d82f758-7150-47f3-ae3b-571ab367f929 · outbound

This paper cites DiffusionNFT: Online Diffusion Reinforcement with Forward Process.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T22:54:01.601778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:c6ca54f45dfe77499028cfaf80a1145515b90b7966b548485c263932f4d4da5f

Observation b4a22a5f-00fb-4fb4-abeb-d4e5a4d4eb25 · outbound

This paper cites Diffusion reinforcement learning via centered reward distillation.arXiv preprint arXiv:2603.14128, 2026.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Diffusion reinforcement learning via centered reward distillation.arXiv preprint arXiv:2603.14128, 2026

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.587929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:1e3cb8a293e92e90e4574e44168d2abb4af3ef96079cbf391e3c79f55f647d79

Observation 05ed8acd-41ef-4f20-8785-dd71f6052a47 · outbound

This paper cites We report the combined reward, as well as HPSv2.1 and CLIPScore.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models We report the combined reward, as well as HPSv2.1 and CLIPScore

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T22:44:25.152451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:c3f57150749ec4f3d8ea493b4dc0a059b6287c44e020cd7dc1eb614cf6377fb3

Pith citing papers

No inbound Pith citation observations are available.