Pith. sign in

Paper Citation Record · LEDGER

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs

As of 8 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2506.03077.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03077 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:17:37.970701Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T21:05:45.024226Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T21:09:02.660665Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact1
  • verified fuzzy5
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ce1ef4b6-dc6c-438a-9191-552448c946c7 · outbound

This paper cites GPT-4 Technical Report.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.347052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.347052Z digest=sha256:fb9dfecd63f5122d9389326cb6b2b7cd8c86a02da9b056f510451dbc2030b5d3

Observation 8a1bc54c-b3c1-4dc2-b9be-0a48a35291f0 · outbound

This paper cites Gqa: Training generalized multi-query transformer models from multi-head checkpoints.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Gqa: Training generalized multi-query transformer models from multi-head checkpoints

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.397085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.397085Z digest=sha256:e9e40b49cd14bbbb4236f1eff128a9ede63b3253e4a634f689c01abd3ef104cd

Observation d6ac6422-7653-49a6-94ea-aa044d1ae0a8 · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.485503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.485503Z digest=sha256:e188d04011bbbcca25ebfaa888951542d5f7a503ad6c2c428efbb177be8f77d7

Observation 10f603e4-e3fe-44d3-a401-4caa62d4191a · outbound

This paper cites Training Deep Nets with Sublinear Memory Cost.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Training Deep Nets with Sublinear Memory Cost

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.549383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.549383Z digest=sha256:f0a2d86690da67bd36bc55c517ac855f8b60e7c4b13ba4c44b13311841894e3a

Observation 50e68872-b351-4f6b-a2ea-b6cbac7916b9 · outbound

This paper cites QLoRA: Efficient finetuning of quantized LLMs.Advances in Neural Information Processing Systems, 36, 2023.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs QLoRA: Efficient finetuning of quantized LLMs.Advances in Neural Information Processing Systems, 36, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:39.355374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:17:35.603553Z digest=sha256:6bb2f1da2a0d521de71608a3c806cf9b61257717457b27a4104fbc427a485255

Observation 8464f160-98df-44bf-a350-34d01d7099c4 · outbound

This paper cites The Llama 3 Herd of Models.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.688187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.688187Z digest=sha256:e8e1437c0e8c016fad7d46565eaeed154b4ec2909650850d0a17fd11db7e98a0

Observation f7a0a08d-cd12-478e-94a9-8fe0c635447b · outbound

This paper cites Open R1: A fully open reproduction of deepseek-r1, January 2025.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Open R1: A fully open reproduction of deepseek-r1, January 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:39.178562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:17:35.748480Z digest=sha256:ae4c710d214f802c3a0133350b8528e58f7fb3b2e37441543b68f47461982cae

Observation 163b6c39-173b-471c-97ca-30fa4813e820 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.844677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.844677Z digest=sha256:2ab20795c788935d54a0d271e5602a95020210bec4b4b4cdc7601b7a40986871

Observation 137bfbff-2e0c-4154-9c23-4dc4cf1a7223 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs LoRA: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.915418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.915418Z digest=sha256:e2e84787709acbbafe86664cd2938b7c53b455a9074f92025b2834daa16012dc

Observation d0645b24-f71c-4b25-ac34-00e689d4bd09 · outbound

This paper cites Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:35.974206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:35.974206Z digest=sha256:74ef59e49476bcd3f785005caa2ea3112ec4d3337ec27ffd862d4a4b4ed8f2f7

Observation 00f6e65c-b6cf-49aa-b499-8622ddf57d83 · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.055511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.055511Z digest=sha256:95074ef3115c2ca70b314e9dff8f46be31855dbb036b5d090a5ff285e2dbe552

Observation 086568fe-f6f3-4c8c-a795-a0e1505e67d2 · outbound

This paper cites OpenAI o1 System Card.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs OpenAI o1 System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.128133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.128133Z digest=sha256:f6f2c76581424724b7a87161c134e1055f3796c6ff101ebccaa5509949726f44

Observation 37f2a0e6-6d34-467f-a0e2-17d06eb95adb · outbound

This paper cites Adam: A Method for Stochastic Optimization.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Adam: A Method for Stochastic Optimization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.163838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.163838Z digest=sha256:794671adbdc8f1d34c7e0248ef2054f63cc8e8ef450f5fa8959499002aa0abc7

Observation 089afc44-e368-44c4-baf6-206e03341c14 · outbound

This paper cites Reducing activation recomputation in large transformer models.Proceedings of Machine Learning and Systems, 5:341–353, 2023.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Reducing activation recomputation in large transformer models.Proceedings of Machine Learning and Systems, 5:341–353, 2023

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.196477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.196477Z digest=sha256:b393be9c095ccbe21cd78856793db6b9cd08fc1ce040af1d47d0bcc53c522b15

Observation a6807fc2-30f5-4543-adcc-fee13ca3928d · outbound

This paper cites LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.286634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.286634Z digest=sha256:1e8e1a23af2b92e0531f2ff1ccd86597c2c1502540411c58eaa72c5191c988d6

Observation 8316e377-e9ca-4bba-99fc-99db6a301f91 · outbound

This paper cites Sequence Parallelism: Long Sequence Training from System Perspective.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Sequence Parallelism: Long Sequence Training from System Perspective

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.333821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.333821Z digest=sha256:8d51ba48ad2f35e62440646abf1984dc32a2958b7dcbdf5e73b4d01527f8dea1

Observation 887dea13-d8a6-49b2-9740-74bebe2764af · outbound

This paper cites Dora: Weight-decomposed low-rank adaptation.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Dora: Weight-decomposed low-rank adaptation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.385160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.385160Z digest=sha256:71e47c0230a774d2c6b865fd26bced9e6a8840e7c7a384f210c147267cdf4f68

Observation 50d9730b-4d11-47c8-a9b9-7758bdcc25c5 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Understanding R1-Zero-Like Training: A Critical Perspective

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.477447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.477447Z digest=sha256:4aba9e3ed791697eb95f5014fa84cf387bc829dde35e7dee390df000f8fc6d52

Observation 79163203-8efb-4bb8-93d8-a2665bac4e20 · outbound

This paper cites Decoupled weight decay regularization.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Decoupled weight decay regularization

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:39.038667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:17:36.541873Z digest=sha256:5388780edc16b8635890ca7dfe5598d8083f1cb3e34d5d4c8dbea043f585c489

Observation 49f1994c-d8bb-466b-9e92-d43852bea311 · outbound

This paper cites Mini-Sequence Transformer: Optimizing Intermediate Memory for Long Sequences Training.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Mini-Sequence Transformer: Optimizing Intermediate Memory for Long Sequences Training

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:17:38.316476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:17:36.607266Z digest=sha256:88dc1d501e17eb89e7f26b7dac8ee298032189eba0980324475df8d66d6bc819

Observation 01b96a64-3913-4b31-a9bb-d1fc8f587624 · outbound

This paper cites Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:38.876655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:17:36.665209Z digest=sha256:f0007c91d6075d11b983888debaf8bca2f493b37880a9902ab4d814479bd47bc

Observation 78272862-aaf3-446c-b57a-a27a901ed0f9 · outbound

This paper cites BAdam: A memory efficient full parameter optimization method for large language models.Advances in Neural Information Processing Systems, 37:24926–24958, 2024.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs BAdam: A memory efficient full parameter optimization method for large language models.Advances in Neural Information Processing Systems, 37:24926–24958, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:38.708583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:17:36.746379Z digest=sha256:abe051b60e29116ef04d1f9d655b1879595b55dec462b75c2ea4d4d3a4bd3fa1

Observation 34fa238d-a471-43bd-a4d6-a585d826c1dc · outbound

This paper cites s1: Simple test-time scaling.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs s1: Simple test-time scaling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.799677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.799677Z digest=sha256:f30a202592fba5135d58028f292a88ec4d867636844b36bf22c3be64f6f3578e

Observation c193548a-cf57-444e-b2fe-ac0e8e932dbe · outbound

This paper cites Tinyzero.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Tinyzero

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.886052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.886052Z digest=sha256:a05c7e84dcd2e4635053d37e3ff723dff7768df98cdb82493acb41e213de3b8f

Observation 9d289a50-00d3-4dce-b04c-517a15cf9229 · outbound

This paper cites ZeRO-Offload: Democratizing Billion-Scale Model Training.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs ZeRO-Offload: Democratizing Billion-Scale Model Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:36.966763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:36.966763Z digest=sha256:4c22bb852330f5a164e12f0f13ac4b73c42362c11ad90290188466b36f2e68f9

Observation 6697b25b-72d6-4cdc-8f82-5d4f06db5907 · outbound

This paper cites Trl: Transformer reinforce- ment learning.https://github.com/huggingface/trl, 2020.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Trl: Transformer reinforce- ment learning.https://github.com/huggingface/trl, 2020

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.023132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.023132Z digest=sha256:9564144c27bf2165597c301861ae1acc5a5eab652c2401068ae0b38b80f57933

Observation cdfea2b3-f913-47a6-91c9-b5bd5bc89388 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.098700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.098700Z digest=sha256:564a422bd778f8a8a73ea77741dae511a407f22d61b85c5644b8abe586a528a0

Observation 25a6b81e-8e1b-4b63-8f3c-48e89439d978 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.164737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.164737Z digest=sha256:69b89928b2dcbd49fd6fdab02d2234c125adb50d762ccb3c9fdd160ff7b5d572

Observation ec882aa4-946f-4bb0-924e-0ca1319546e9 · outbound

This paper cites Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.235199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.235199Z digest=sha256:92b66f3bde45cb2c1e143e7921334e9d030ce2e2997e3b15755689e4d8fabbb0

Observation fa423f3e-9149-42c9-a053-5cb9e27e2421 · outbound

This paper cites an unresolved cited work.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.353885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.353885Z digest=sha256:0d57377a1afd7b06269206698ba2534b1ac908b73a1a5535a81420102165e519

Observation a44ae5d8-89b5-4f76-af6a-b4533feaaa14 · outbound

This paper cites Qwen2.5 Technical Report.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Qwen2.5 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.412316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.412316Z digest=sha256:ada96a16e21e18ec804cb5c13383a7cc1f59e1b24567ca8683373646e5594a56

Observation 4b920b99-fc98-4c36-a5b0-9bf5f8dff18a · outbound

This paper cites LIMO: Less is More for Reasoning.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs LIMO: Less is More for Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.467105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.467105Z digest=sha256:89b926f59baabb5da59533d533c95652aad955e38617278b932c2474f948436c

Observation 093f2445-c691-4609-9806-fc18ae34f1e2 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.522534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.522534Z digest=sha256:dabaf451b68f0d36ad4b306950a98dbf6ddb20d8c79489653920eeb35adf04ff

Observation 1d3da582-05a5-4454-8db3-fe030f03c885 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.613949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.613949Z digest=sha256:5a6e6789db9d38b10c57f2d312b1c80fb3c6df42fcd3090979f6c1bb94af5d41

Observation 845f3139-907f-4c8e-81c9-188f64adc75d · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.752175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.752175Z digest=sha256:928769b680f1e34e1d7b74e65da879d92893b4d155bdb35f82574d38e4682eac

Observation 3bf4bb70-508f-4448-abfc-112fdcb2806b · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.858509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.858509Z digest=sha256:59ad9a52e2e1d795d3444cbdd7f33c5f5a6a0ef6abdb8bb743f15fb18d2fe714

Observation dd5fd556-5431-438c-8f11-dcfccaef5494 · outbound

This paper cites GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection.

StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:37.970701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:37.970701Z digest=sha256:ddd18204b3b3d499b42087ee48b17f8a700880d3dda05317341be4fa073a1cc5

Pith citing papers

Observation 5ff6b850-2079-4cb7-9d0e-ae335982c396 · inbound

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking cites this paper.

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:09:02.662223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T21:05:45.024226Z digest=sha256:1aded7bb0def8ba5bad79da43ea410abfca1637a9cfd4b895a41044adcfd4894