Pith. sign in

Paper Citation Record · LEDGER

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning

As of 13 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2606.26997.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.26997 v2

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T11:52:35.159984Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T09:49:28.700382Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5149cbf7-dd78-4fc9-8aef-3cc9b5dbac56 · outbound

This paper cites AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:0491c236ef70724d26d0f75761ce06b291db67bdc66c26b55111133447a48306

Observation f20a299c-9529-4315-8846-7552fe3b3b62 · outbound

This paper cites Nature645(8081), 633–638 (2025).

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning Nature645(8081), 633–638 (2025)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:65e64043fa6c82d839375730597be1b8c451df853367676a0bb54470a1421bef

Observation c8d59ed5-d847-4c63-880f-f18e47cd6fab · outbound

This paper cites AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:df21bb96daca1927615d73a4fde8163a8d95acd6ab629bfd6f24fcdaab67a1db

Observation d879b52d-318b-498c-9431-b6378b52b818 · outbound

This paper cites In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers).

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:cf4e65ed3e2d31299504a2b17d2437790ff657092ebdc0cbb2b57516e7ab6721

Observation 2a10505b-911d-43f2-ae2d-a140b072a76f · outbound

This paper cites In: Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing: System Demonstrations.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing: System Demonstrations

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:2a04b208f8996753d878989060a2c5776ebe1d8ced8c6f41c61c9b0af9bfca81

Observation dbd7f5f7-7f30-4f50-8c94-4c8f84da2733 · outbound

This paper cites In: Proceedings of the 8th International Conference on Cognitive Computing.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 8th International Conference on Cognitive Computing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:c4e14dd9e1b4b38f70d3db48c5208f409eb90006f78843a2b17b21956df37810

Observation 317c6543-8b85-46ce-9683-45a91e1b69a2 · outbound

This paper cites In: Proceedings of the 9th International Conference on Cognitive Computing.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 9th International Conference on Cognitive Computing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:fd9e78078326b8c6574172687056f55c76bb344cc5344d1c57c052f9e995e445

Observation a9c49bd3-2cb0-47ed-8709-5475382f5952 · outbound

This paper cites In: Proceedings of the 13th USENIX Symposium on Operating Systems Design and Implementation.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 13th USENIX Symposium on Operating Systems Design and Implementation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:c88158e79bbcc6e08c9f1200aeb776b8c505c54bb10d624b78fa451151ec8f0d

Observation e7d5bbb7-071f-406f-8633-205d7cb1bea8 · outbound

This paper cites In: Proceedings of the InternationalConferenceforHighPerformanceComputing,Networking,StorageandAnalysis.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the InternationalConferenceforHighPerformanceComputing,Networking,StorageandAnalysis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:e6628b8b08b7ca0f4bcbb78653009f9fd5099d01d81352518601f586fe93d4ba

Observation 8210c385-9885-4d74-aa7c-a64ad2c01ab8 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Advances in Neural Information Processing Systems

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:879f19d18dc766d260fda7a9b3b1b302d340754fb0e5df43daec95d6e16f2452

Observation becf97e3-52ba-42c8-a7f2-475418e78904 · outbound

This paper cites In: Proceedings of the Twentieth European Conference on Computer Systems.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the Twentieth European Conference on Computer Systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:d3e5f4cac78baf04da4af4b1183f1a4fa523874fbba1caccc13aa4c69b6a9df7

Observation a02d4598-cc45-4561-856a-0f3c056ffb25 · outbound

This paper cites In: Proceedings of the 41st International Conference on Machine Learning.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 41st International Conference on Machine Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:e467a8b14ebee72f65d2d24cc8b5de6c72317fe23c6889f1d324b23fe69f880a

Observation a0521a09-d440-426d-a722-4022c2cebdb6 · outbound

This paper cites an unresolved cited work.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:9b71ecf4d0a245d46fa1aac6f41a2c85ecb77e1a7af6cb74280b07426dbf43fa

Observation fd864d3e-a2e7-4090-9387-a2ef3c8a4995 · outbound

This paper cites In: Proceedings of the 9th International Conference on Cognitive Computing.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings of the 9th International Conference on Cognitive Computing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:a849b14006c484499cbffa627e2cf90fe6b88e5b595284f3d22c8edda320eee3

Observation 95b975a3-86b4-443c-a811-9a560b0a5a5c · outbound

This paper cites In: Advances in Neural Information Processing Systems.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Advances in Neural Information Processing Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:7fdca5d4f4f15e441b135ad28d7b8bfa33dff563b6cc0c4a34cbca45d894f3c2

Observation 1443aa22-f744-4d4d-ad1a-0b7b948733cb · outbound

This paper cites In: Findings of the Association for Computational Linguistics: North American Chapter of the Association for Computational Linguistics 2024.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Findings of the Association for Computational Linguistics: North American Chapter of the Association for Computational Linguistics 2024

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:fbfdf1bec5a900a50693fa6674ad0fa4a9f5fced6f4da92f448899a306e3e797

Observation f2ab2b5b-2519-4fdc-8a89-130a64005db0 · outbound

This paper cites In: Proceedings 16 R.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning In: Proceedings 16 R

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:b830522cb065b3528afff6f4c387528f9f16cab05f3ffeb718d88f38f3d1fac6

Observation acca909e-eb1f-41d0-b378-9dc8aa59e860 · outbound

This paper cites GitHub repository (2025),.

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning GitHub repository (2025),

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T11:52:35.159984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:52:35.159984Z digest=sha256:f33ca1285ad046937be1ccc26d98af66aaa2bd31d17f57811d1b7b82041c3b94

Pith citing papers

Observation 7164737e-8b03-4d42-8053-986528121718 · inbound

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning cites this paper.

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T09:49:28.700382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T09:49:28.700382Z digest=sha256:17bc9a842f2b4b8763305f54a1fd7b01186d0117bceb03999edf781b7507fcde