Pith. sign in

Paper Citation Record · LEDGER

Flow-OPD: On-Policy Distillation for Flow Matching Models

As of 13 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 14 inbound Pith citation observations for arXiv:2605.08063.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.08063 v5

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T23:02:29.120150Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:07:57.812113Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:46:41.044701Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact28
  • verified fuzzy21
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a91f76a8-7138-4aac-9305-ec5515619589 · outbound

This paper cites an unresolved cited work.

Flow-OPD: On-Policy Distillation for Flow Matching Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-07-07T12:13:45.211518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:c309fab28bcf78678199a25b698df3331bc8b469ef8ecf84c56ba25df740c68e

Observation 66c7a2f1-3086-4a6c-b28f-c6f871bc9238 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis.

Flow-OPD: On-Policy Distillation for Flow Matching Models Scaling rectified flow trans- formers for high-resolution image synthesis

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.221791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:316ab919c8ba88fef886d69b07036280cfff4403e1e632bac2499a168aaca4a6

Observation 3e71ed6e-58b6-4d5f-b310-7594f191da4d · outbound

This paper cites Flow Matching for Generative Modeling.

Flow-OPD: On-Policy Distillation for Flow Matching Models Flow Matching for Generative Modeling

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.261733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:34a504e06a2ff8ed39e56a2123afa3123180ca8586fcb60f4a8af97c7abaebfe

Observation 7e74ce14-25f4-4805-af19-8f856fe233f4 · outbound

This paper cites Dualvla: Building a generalizable embodied agent via partial decoupling of reasoning and action.

Flow-OPD: On-Policy Distillation for Flow Matching Models Dualvla: Building a generalizable embodied agent via partial decoupling of reasoning and action

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.244718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:5cfd22a471ab078065a70940691e7d322dd9ad634fe3cb17e470293c155c97f9

Observation 80f96572-4688-4b3b-b18e-a3ec2637fa80 · outbound

This paper cites Vision-r1: Incentivizing reasoning capability in multimodal large language models, 2026.

Flow-OPD: On-Policy Distillation for Flow Matching Models Vision-r1: Incentivizing reasoning capability in multimodal large language models, 2026

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.238354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:e83ac9e097c6d2722bea242f7c4aff22a2326cd1d85bd62c2c2c673b319499ba

Observation 5b39bc5b-4c0d-4f9a-9972-b60439e1d75f · outbound

This paper cites Vision-deepresearch: Incentivizing deepresearch capability in multimodal large language models.

Flow-OPD: On-Policy Distillation for Flow Matching Models Vision-deepresearch: Incentivizing deepresearch capability in multimodal large language models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.248068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:8843284308a15d6fc983d7ef9a8649fd970840b68ba58dcafd75fd4ad42f9031

Observation af800b1c-799f-4a9a-9783-5cd7b12ea13d · outbound

This paper cites Advancing multimodal reasoning: From optimized cold start to staged reinforcement learning.

Flow-OPD: On-Policy Distillation for Flow Matching Models Advancing multimodal reasoning: From optimized cold start to staged reinforcement learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.254996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:cbf51f02773a4f5e01a36614cf08621b839a227aba3717d0946f1d55efd243ba

Observation 3cdaef82-b274-4188-9dea-cc5ae5079c52 · outbound

This paper cites Ares: Multimodal adaptive reasoning via difficulty-aware token-level entropy shaping.

Flow-OPD: On-Policy Distillation for Flow Matching Models Ares: Multimodal adaptive reasoning via difficulty-aware token-level entropy shaping

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.258793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:b1555def5988b858c4992995ef171358ce4a824662aa7ee51f8e5bfbe0397fe2

Observation 95786e95-88eb-428c-9837-f392268102da · outbound

This paper cites Opensearch-vl: An open recipe for frontier multimodal search agents.

Flow-OPD: On-Policy Distillation for Flow Matching Models Opensearch-vl: An open recipe for frontier multimodal search agents

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.240285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:b5ea2435b6611879f097bf239e3e63e6e5e375a5a12bc774316295a7755f4f28

Observation e695a49a-c9d8-43dc-8996-8c72a65d9889 · outbound

This paper cites Unicorn: Towards self-improving unified multimodal models through self- generated supervision.arXiv preprint arXiv:2601.03193.

Flow-OPD: On-Policy Distillation for Flow Matching Models Unicorn: Towards self-improving unified multimodal models through self- generated supervision.arXiv preprint arXiv:2601.03193

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.254845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:0aa5434bae870d60b79e9974f929aaf96df156606264f834027cb67092fd7e21

Observation d965bc1b-f804-4eae-b68a-cb232c911256 · outbound

This paper cites Unify-agent: A unified multimodal agent for world-grounded image synthesis.

Flow-OPD: On-Policy Distillation for Flow Matching Models Unify-agent: A unified multimodal agent for world-grounded image synthesis

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.267827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:7bf8467d28a85e8331584723744dba1ab24f7bfb6b67b87aa3dfb38f274653e6

Observation 223cb247-875e-4f20-b40f-2f0d67f28678 · outbound

This paper cites Gen-Searcher: Reinforcing Agentic Search for Image Generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Gen-Searcher: Reinforcing Agentic Search for Image Generation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.274031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:d8d6d88c7f22a3e30c84caaa7b8f391fc1958cf40a9da628051d5d8e732ac776

Observation b10ec737-893f-4828-8850-e260f3b6d7e4 · outbound

This paper cites Interleaving Reasoning for Better Text-to-Image Generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Interleaving Reasoning for Better Text-to-Image Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.211199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:66efad2c0142324b75814f4fdf846c8e6417acc8c30259ef27e40116588e801d

Observation 64d914c7-fd2a-4158-b40f-67e3dff47bc3 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Flow-OPD: On-Policy Distillation for Flow Matching Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.211077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:c0a2fc02f2096a6418f237a216988cb95d63b4a53b22f887e48dd1bf569f3f5d

Observation a2167dc0-d770-49ea-8253-1c8396fb4168 · outbound

This paper cites Dancegrpo: Unleashing grpo on visual generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Dancegrpo: Unleashing grpo on visual generation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.242642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:11368e0ac856e253950c883f7a64f16a58c321c94226d9ca01260bbf8e4f5364

Observation 468b2dac-b30a-4bec-88eb-4ca16c29c1ce · outbound

This paper cites MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE.

Flow-OPD: On-Policy Distillation for Flow Matching Models MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.197061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:2be5e899454770d39cb6ebe9fdefe96cbd8d353c88d572be6dfb77e7b76b84dc

Observation 25c36539-8dde-4ffd-8247-c4f821d377c0 · outbound

This paper cites GLM-5: from Vibe Coding to Agentic Engineering.

Flow-OPD: On-Policy Distillation for Flow Matching Models GLM-5: from Vibe Coding to Agentic Engineering

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.200138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:8517241107dd96df6f3be841d35e3bb881c85d7fda1db0d9a5740f1f8513b3c8

Observation 28b8d2d3-d86d-46ef-afce-4ca37d5e2b53 · outbound

This paper cites MiMo-V2-Flash Technical Report.

Flow-OPD: On-Policy Distillation for Flow Matching Models MiMo-V2-Flash Technical Report

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.213902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:bd538e241e340da371c8cb273266dfd78f210fbf121b806ee91b5d5fb554b839

Observation c9e74e51-06bb-4a2c-963e-72f57034596e · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023.

Flow-OPD: On-Policy Distillation for Flow Matching Models Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.245210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:2f5d30c438b24b313c4eb32131587b782978bac3a507e22b64600d3251cf0156

Observation bae81ae3-fee2-4eed-a012-325e13db8cf3 · outbound

This paper cites Textdiffuser: Diffusion models as text painters.Advances in Neural Information Processing Systems, 36:9353– 9387.

Flow-OPD: On-Policy Distillation for Flow Matching Models Textdiffuser: Diffusion models as text painters.Advances in Neural Information Processing Systems, 36:9353– 9387

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.234639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:49b4dfe147da2b98e66bfc58c3ac5f7159e2876c4a823b428cb210f51f60fe4d

Observation e1ba4e03-c8ae-4d4f-aeff-73899311fb3e · outbound

This paper cites Training diffusion models with reinforcement learning.

Flow-OPD: On-Policy Distillation for Flow Matching Models Training diffusion models with reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.230681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:a911094321235e5265acc7df2deabc56da5db2ee6e84e4eda5238a8af45e0bdc

Observation 96263acc-8ac5-4ff7-8a8e-d202f0ed1092 · outbound

This paper cites Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models.

Flow-OPD: On-Policy Distillation for Flow Matching Models Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.232686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:0df1ebc8fba9cb0589603c307a920ea833353b2f48a42ba7e3620e44d43f1b29

Observation 59e43cfe-be79-4194-acc0-705e20427db8 · outbound

This paper cites Imagereward: learning and evaluating human preferences for text-to-image generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Imagereward: learning and evaluating human preferences for text-to-image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.236645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:b48eae59ff00bf9c1b18d4b8083245d40e8c799045d2c306fd0aced73f19334d

Observation c11f2208-c858-46c8-9459-dcc829dd2341 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

Flow-OPD: On-Policy Distillation for Flow Matching Models Diffusion model alignment using direct preference optimization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.249882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:0000544f8ef8b4a3f7da1d880f4e542262d460b754f5c32cd9f8da68f3f75e55

Observation e94abee9-1bae-46ed-9689-99c9a237dff2 · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

Flow-OPD: On-Policy Distillation for Flow Matching Models Flow-GRPO: Training Flow Matching Models via Online RL

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.223673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:a47efab13c5bf87f8d6fb90b22124458f5ee4e7a38cceb817db51309b54d3ce3

Observation 38b76edc-9ccf-412d-9a0c-fee4be8d2d46 · outbound

This paper cites AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning.

Flow-OPD: On-Policy Distillation for Flow Matching Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.203930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:e3e57ac9e1922315f9fac8492cbc08d0cb416fc90b3fd41ea7cfddc760e7cc5c

Observation 03c190be-ce76-4ecd-9326-cd2181605489 · outbound

This paper cites Group critical-token policy optimization for autoregressive image generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Group critical-token policy optimization for autoregressive image generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.227667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:32c998b184faecfd116c615a923afe2f49a63b442a4fa793978c8f1f09ec8cf3

Observation 97e58cbd-d4a3-4ee2-9527-598b6f391037 · outbound

This paper cites Stage: Stable and generalizable grpo for autoregressive image generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Stage: Stable and generalizable grpo for autoregressive image generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.264939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:c280afdf892155ea89e6847ca7af32fe7cfc2a5b88d178ca5c4e5edb25199c68

Observation 725afb2d-2343-40b9-90fc-6550de42cc0f · outbound

This paper cites Maskfocus: Focusing policy optimization on critical steps for masked image generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Maskfocus: Focusing policy optimization on critical steps for masked image generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.237873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:60db41418970e2276e4f7b9788be14824204cc60b09f84451ac8ca935ba5085c

Observation 65ff4660-1838-4c73-89d9-21fa16621882 · outbound

This paper cites MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.241028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:94b7902086cba273fd4c093aa7909410e578ca7d9bb54441e3e849fb088e3fa6

Observation 585f42be-5381-4eda-b776-751c12e0a821 · outbound

This paper cites Gdpo: Group reward-decoupled normalization policy optimization for multi-reward rl optimization.

Flow-OPD: On-Policy Distillation for Flow Matching Models Gdpo: Group reward-decoupled normalization policy optimization for multi-reward rl optimization

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.224967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:e3841f2384b77366bb462a6a544e03cf7ad9ff4d24cbe9a9d924a464f5380ddc

Observation 501ac910-cfe6-41bc-b648-abbcaa451632 · outbound

This paper cites On-policy distillation of language models: Learning from self-generated mistakes.

Flow-OPD: On-Policy Distillation for Flow Matching Models On-policy distillation of language models: Learning from self-generated mistakes

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.226872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:740b7bb259d2af5158f764297d90c8a49609e1aa0f057b356f421b128508c314

Observation 8a38b412-4c97-48a0-8d14-1f5fff80410d · outbound

This paper cites Minillm: Knowledge distillation of large language models.

Flow-OPD: On-Policy Distillation for Flow Matching Models Minillm: Knowledge distillation of large language models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.228845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:c84620276dce374d569ad4fbb9e53108d8a874953f921f1b3a1ff276cf70a91e

Observation 2129fdcb-5d2b-4d3e-beb8-6ec1ef47e87d · outbound

This paper cites DistiLLM-2: A Contrastive Approach Boosts the Distillation of LLMs.

Flow-OPD: On-Policy Distillation for Flow Matching Models DistiLLM-2: A Contrastive Approach Boosts the Distillation of LLMs

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.251601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:675e5caa66be3feb0b8b2798d64a4562a90cb3c8e1cf615985f6aa1ddac3608f

Observation 4c5bf3cd-567f-4a87-b520-d44d060aa831 · outbound

This paper cites Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.268007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:68b111663a7d6a4352d4bdcd37ec5e15d65ddfa8ec20ef364ecf90c3a76dd3c5

Observation dc6281a4-cc26-42fe-bfe0-955cc9cad4c9 · outbound

This paper cites Entropy-Aware On-Policy Distillation of Language Models.

Flow-OPD: On-Policy Distillation for Flow Matching Models Entropy-Aware On-Policy Distillation of Language Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.271025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:1f5a0af25a5149ccccfbeb4a6175f844d8ccbcc0a2063eb8e3199f966753305e

Observation f1cda52c-a659-483d-a512-21696893ed63 · outbound

This paper cites Fast and effective on-policy distillation from reasoning prefixes.

Flow-OPD: On-Policy Distillation for Flow Matching Models Fast and effective on-policy distillation from reasoning prefixes

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.214348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:01aa687145b3081e6f46a1843c777d0733912e824229a844e5891c86e510b9dc

Observation f8912e26-cb7a-4788-8e23-19b0577f1fda · outbound

This paper cites Paced: Distillation and self-distillation at the frontier of student competence.arXiv e-prints, pages arXiv–2603.

Flow-OPD: On-Policy Distillation for Flow Matching Models Paced: Distillation and self-distillation at the frontier of student competence.arXiv e-prints, pages arXiv–2603

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.217283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:b83f4042a77f8ad753ff60748e8f71f7d0dd195d4e0bffe5b65d9f82daeb233e

Observation d54ea958-c430-45ba-a40c-2accf4f68eef · outbound

This paper cites On-policy distillation.Thinking Machines Lab: Connec- tionism.

Flow-OPD: On-Policy Distillation for Flow Matching Models On-policy distillation.Thinking Machines Lab: Connec- tionism

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.213144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:aad32f01622cdd3b538a037e3a13f78fdb5d8d1fd8d081eb95c206fc75db535d

Observation b9994932-da38-404c-8a26-3101cc22c0ed · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advances in Neural Information Processing Systems, 36:36652–36663.

Flow-OPD: On-Policy Distillation for Flow Matching Models Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advances in Neural Information Processing Systems, 36:36652–36663

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.215525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:b003b99885690389d37bae3f6fcda5487e7175c679cc50cb826fd26a3a2c0918

Observation 679ee06e-50be-4d76-926b-1e1d53757416 · outbound

This paper cites Teaching large language models to regress accurate image quality scores using score distribution.

Flow-OPD: On-Policy Distillation for Flow Matching Models Teaching large language models to regress accurate image quality scores using score distribution

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.230811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:00ff17da2bd868eba532379f26553d65243b994069dfcf95ab3b8cb45b8be195

Observation ce1a0940-8075-4cd6-9905-3b990f13fd29 · outbound

This paper cites Laion aesthetics, Aug 2022.

Flow-OPD: On-Policy Distillation for Flow Matching Models Laion aesthetics, Aug 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.219746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:b9f5a17094e9b79264864bd8daa6916042e3c4d7f3b817421090f47f90517250

Observation 3532654c-53c8-4187-8034-6511fbcdcd02 · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.248229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:62ade1740b95add5fd9b3a3c6d519186b6cbea886e778eed10393e531c379510

Observation 9aa0d2ce-382c-49d1-ae50-5004a21c1aea · outbound

This paper cites Unified Reward Model for Multimodal Understanding and Generation.

Flow-OPD: On-Policy Distillation for Flow Matching Models Unified Reward Model for Multimodal Understanding and Generation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.223994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:84566d95167e1388a8caa318c879af043d2a0c23a6b9c4c2853a2076f777303f

Observation 52d1b973-000e-4bea-9a2d-09b0dfaea736 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

Flow-OPD: On-Policy Distillation for Flow Matching Models Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.221489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:ad9d7ff8c2a7f21cc99f82cfe441bf9416a58dd663ce771406810c788bbc1a08

Observation d8c02852-2b42-433a-8035-76c56d02c6ee · outbound

This paper cites Qwen3 Technical Report.

Flow-OPD: On-Policy Distillation for Flow Matching Models Qwen3 Technical Report

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.198163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:3e3f3e078eace4a49c1d0a81551eb5b337b5f9cab3c5bf0b16d94fa93d01f919

Observation e8693666-5934-4d0a-bca5-8b80b71f7a5f · outbound

This paper cites DiffusionNFT: Online Diffusion Reinforcement with Forward Process.

Flow-OPD: On-Policy Distillation for Flow Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:05:07.200888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:555a1547804143faaedc57e885432a4ff6f144c61af8d4e154d375bd5b0b789e

Observation 8c6828f1-5ae7-40db-914f-86735c46499a · outbound

This paper cites •3 (Fair):In focus, adequate lighting, but lacks creativity.

Flow-OPD: On-Policy Distillation for Flow Matching Models •3 (Fair):In focus, adequate lighting, but lacks creativity

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.207824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:31c75cc41caef3032c6d90303d1bf5cd444ac2b7fc8aa5bed46ea436dc1637bc

Observation 263bc3a4-643f-487c-9e20-8f88ed910b96 · outbound

This paper cites •3 (Fair):Partially follows, but distorts some important elements.

Flow-OPD: On-Policy Distillation for Flow Matching Models •3 (Fair):Partially follows, but distorts some important elements

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.209997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:20a84f1aecd44e196d8bec04360bfae98e77a6bbcb6b1744a2b1842987e19603

Observation b79f2a78-7c2a-4051-a450-f4c160214d1f · outbound

This paper cites winner-takes- all.

Flow-OPD: On-Policy Distillation for Flow Matching Models winner-takes- all

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T12:13:45.205987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:1939ed2f30f4beee95f49fb8afb21a9abc40cf536d5b8c79fdbe80b8999e0f29

Pith citing papers

Observation a09e644e-bf7c-4ce7-90c5-a46c4235bad2 · inbound

CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation cites this paper.

CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:44:01.967546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T22:34:34.927202Z digest=sha256:ba1cb51684761d409f71f2d39da2fa736fcfa65451a69924299bc1718e287563

Observation ccee52f2-b027-45a1-92ec-1835f22637f1 · inbound

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation cites this paper.

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:36:17.405805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-28T15:14:42.489647Z digest=sha256:e5fbe78fa23b0128b55521be8749a1bdfb429e86b1471feed4751b97622a7a16

Observation 69bf3f1e-3ab5-4125-a429-c120bed6187f · inbound

Qwen-Image-Flash: Beyond Objective Design cites this paper.

Qwen-Image-Flash: Beyond Objective Design Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:26:26.209957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T11:00:04.775189Z digest=sha256:791b92486da122a4efa24d1ea73351619dd8a96cdd2d14b61ac6cf579bff6d7e

Observation f2677ecb-d242-4761-95bc-04ed3107395b · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-04T13:49:51.393566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T04:55:42.018348Z digest=sha256:e60293da164322d9fe4d3a711b53ce17d2ef4652e7e693e9174883cd9ef9c5c2

Observation 8130d905-5353-4844-b283-7e5446b546c4 · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T11:44:54.717393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:44:54.717393Z digest=sha256:138feeb43d8645ae918a310d05946aaaa04b9ee5a1c1aa589a13244c371b58d9

Observation 401bdcef-e218-46a2-ad41-79edc5695a72 · inbound

Qwen-Image-2.0-RL Technical Report cites this paper.

Qwen-Image-2.0-RL Technical Report Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:06:02.859413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T01:11:12.757656Z digest=sha256:0e31c73cd74fc67e70de6f6bc0710fb7fb6d570e0263f936f01faa152cafcf1c

Observation 27f1636a-25a4-47f2-b740-2f494e2457cb · inbound

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators cites this paper.

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:46:41.046168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-10T01:43:20.551939Z digest=sha256:94d7389bc2fc24de591bd9dcd2cdcaf1675cfc5380913722b7c97461ad1e29ee

Observation 6f989ac4-0cf3-4317-815b-03b5eedb3f60 · inbound

ArtChart: Faithful Artistic Chart Generation with Integrated Text Rendering cites this paper.

ArtChart: Faithful Artistic Chart Generation with Integrated Text Rendering Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T21:31:18.478134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:31:18.478134Z digest=sha256:2b9f304bc1226fe89067991fbd0ca514ffba1ed82a30d843a8d525bf3077ecd0

Observation 5190c60f-9532-4e49-ba57-d44ea1828e54 · inbound

FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models cites this paper.

FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-31T12:31:28.249661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T12:31:28.249661Z digest=sha256:c6621e3253544a7631ab2eb1bde6d3c98be4ba2b59c71b59077aa3d3406ba962

Observation ed9b9051-2f5c-4494-b155-3ad0a7b6404e · inbound

Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation cites this paper.

Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T06:35:52.550309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:35:52.550309Z digest=sha256:4a2158263ec70685da652bbf04c244a644511f6148192dfdc9ccf13333aa0afa

Observation 73401ac0-0666-49b5-b1d4-e6e2d1aad336 · inbound

EvoReason: Self-Evolving Reasoning Primitive-Guided On-Policy Distillation for Latent Reasoning in Generative Recommendation cites this paper.

EvoReason: Self-Evolving Reasoning Primitive-Guided On-Policy Distillation for Latent Reasoning in Generative Recommendation Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T15:29:59.022263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:29:59.022263Z digest=sha256:f1cedb3b1881b58901d647491c50370b626ea4a50fd97102db5c3968f669f276

Observation e0e57160-ff4b-4e58-bcc5-34741327364e · inbound

Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging cites this paper.

Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:17:16.066972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:17:16.066972Z digest=sha256:871e423de4bb29e241728b0111608e7215cddc752186c61ca78066033c614af2

Observation ac816fb4-6eaa-44d9-863f-207f2f20cc16 · inbound

Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models cites this paper.

Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T19:19:14.648725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T19:19:14.648725Z digest=sha256:1afaed6e4dcaf8c5a403dac327ff20a15c0c138c60e3b632d807979ae90fb63c

Observation a3312d3c-18f1-4985-b7f4-74cc0859b3b2 · inbound

DreOPD: Degraded-Reference Extrapolative On-Policy Distillation for Flow-matching Models cites this paper.

DreOPD: Degraded-Reference Extrapolative On-Policy Distillation for Flow-matching Models Flow-OPD: On-Policy Distillation for Flow Matching Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T21:07:57.812113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:07:57.812113Z digest=sha256:90fa024033b7780ea049577ce6ec40bd1f88119432b6036ba505db71aefd790b