Pith. sign in

Paper Citation Record · LEDGER

AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2508.06924.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06924 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T20:15:31.498912Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T23:05:07.202113Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8d040a30-2afa-494f-9d02-5257517f6f6f · inbound

RubricRL: Simple Generalizable Rewards for Text-to-Image Generation cites this paper.

RubricRL: Simple Generalizable Rewards for Text-to-Image Generation AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T20:15:31.498912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:15:31.498912Z digest=sha256:b3f925d462e3950adc9dc711ea4b0ebc83d742f53d40ce00e45300e976c094a5

Observation d1f8d1ba-5572-4ce3-ac68-ae72e676f66b · inbound

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models cites this paper.

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-13T19:53:46.276873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:53:46.276873Z digest=sha256:ee581f26ad8453920f4abf71349e29b994d0885bc120b9f2f861b3af68187d02

Observation 12f7e41b-cd0e-4d0a-9666-b3c6503e3207 · inbound

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation cites this paper.

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:40:58.127632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:00:50.105629Z digest=sha256:8d36d50e4580c1d977fe7eedd9705a0881e47d8f6608bdc085fdf0f28d6ec33c

Observation fc4cd62e-0e7c-492b-bf12-121799e99c40 · inbound

Flow-OPD: On-Policy Distillation for Flow Matching Models cites this paper.

Flow-OPD: On-Policy Distillation for Flow Matching Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:00:54.406651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T02:04:16.335479Z digest=sha256:760df7f92d15edb79ff485408f4551bf17aa0caf1d3362e7db931cef48807ad9

Observation 9fd480e6-7bfd-4b00-bdcb-a802e1d36ce9 · inbound

Flow-OPD: On-Policy Distillation for Flow Matching Models cites this paper.

Flow-OPD: On-Policy Distillation for Flow Matching Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:02.848757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T01:17:08.955601Z digest=sha256:c13d5e59fe45f6d9a1f55586f81b866afe5f82f28a457fa6c4a612426a947f2a

Observation 2180e0d1-b876-4f84-a5d2-33845f64c0a8 · inbound

Flow-OPD: On-Policy Distillation for Flow Matching Models cites this paper.

Flow-OPD: On-Policy Distillation for Flow Matching Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:55:05.148260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T05:52:26.736304Z digest=sha256:3130e56c174650aa013aba048fafd559d9c0967c0813f68eab775fe30866f18a

Observation a3227d60-3257-421f-8909-db0864cbb5ab · inbound

Flow-OPD: On-Policy Distillation for Flow Matching Models cites this paper.

Flow-OPD: On-Policy Distillation for Flow Matching Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:43:51.848036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T22:39:32.496303Z digest=sha256:a7faaaf196cf1f2a5c65a366f1e850f42e688c17bd22971e3310d27a4fe32db4

Observation 38b76edc-9ccf-412d-9a0c-fee4be8d2d46 · inbound

Flow-OPD: On-Policy Distillation for Flow Matching Models cites this paper.

Flow-OPD: On-Policy Distillation for Flow Matching Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:05:07.203930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T23:02:29.120150Z digest=sha256:e96b8f4d029cc0ec5264e605f0e7602bc8aab5be854fb622ce3c82c471605d74

Observation a03b8e3f-d539-4b24-bd99-f2867ef7bac8 · inbound

Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping cites this paper.

Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 120

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:16:28.949520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T03:33:40.994346Z digest=sha256:8e203e736ffda7df2d0673ff36dba8895c575d16c8d001be451cfdf8aa7cbd95

Observation c575a3f5-7646-4901-99cd-ab997056f66a · inbound

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models cites this paper.

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:22:48.539798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T21:18:03.005508Z digest=sha256:f1d6872f88f2854842258f8af4a3e4451e98313f4e70ed7e51071f4338220ffa

Observation 30356f5c-f040-403e-a18b-13a3670ba30e · inbound

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models cites this paper.

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T17:41:03.368705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:41:03.368705Z digest=sha256:bffa949cbf63539c7747e62ebdab792ac1b819614094b705636c7aeba4c7d94b