Pith. sign in

Paper Citation Record · LEDGER

QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2505.17667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17667 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T01:57:53.449669Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:28:18.208570Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e123b75-b515-4470-9423-af84373323fd · inbound

MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent cites this paper.

MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:17:24.509718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T11:17:24.406028Z digest=sha256:eb2f3c2539b68f589e15b99c7f2e54221730f06e83979fa30622d3ae0475ac9c

Observation 8b38a9a7-7814-4433-8295-e3566f2c684b · inbound

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models cites this paper.

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:50:08.524006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T17:50:08.399160Z digest=sha256:a0bbb25c3464396b9242df39e9a370ee46b1f8ef17fc24277dfe3a577073436b

Observation 2cbaf1f8-7acd-4ada-b28d-537dea77f30a · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:53:28.051553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:53:19.148407Z digest=sha256:3dd208619ee19cace61f250b8b67847c134adb2046beee745baabe3b98115b83

Observation 7416320e-45b1-46f1-b594-5636eb518597 · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-13T15:50:49.083652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:50:49.083652Z digest=sha256:bfd0ff48fdead60522cac5d5fd6155673c62964b0afb828402481ed67b294462

Observation 12109965-067d-4e1f-af18-3eafcfa9dc17 · inbound

A Decomposition Perspective to Long-context Reasoning for LLMs cites this paper.

A Decomposition Perspective to Long-context Reasoning for LLMs QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:31:00.039292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:05:34.666937Z digest=sha256:b2d331bb0829e9ff49714b6fc9af1cb1b5c0e4d5d0b7aef05e806913968adf41

Observation 36cfaa94-33ae-4009-9417-ef4af0b4d499 · inbound

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning cites this paper.

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:20:10.482511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T11:17:43.769244Z digest=sha256:a86830b4a8c7f70ac9e4b1be72bb6975f4b8e054ed01381bb35bfb3d8f178a27

Observation a4c46b9d-afc2-44c0-8e74-c1e1e34d5cff · inbound

OPSDL: On-Policy Self-Distillation for Long-Context Language Models cites this paper.

OPSDL: On-Policy Self-Distillation for Long-Context Language Models QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.402410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T06:07:36.830550Z digest=sha256:dc21c84e0d350d07803a67439f48ee2f186cb8a696eba608b5853a3b9d24eed4

Observation 6e5e2916-47b9-4935-9b3a-d04314fb4766 · inbound

StoryAlign: Evaluating and Training Reward Models for Story Generation cites this paper.

StoryAlign: Evaluating and Training Reward Models for Story Generation QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:31:07.683022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:29:13.549559Z digest=sha256:61305643917250503f17fdc3738d2c7a5a561811e4db5b447edf665df42ce6e3

Observation e6d73681-e8af-4b44-bf0c-fa94aad81275 · inbound

A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and Distillation cites this paper.

A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and Distillation QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:32:19.432348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T05:29:00.576006Z digest=sha256:75b8f194534730de686cc0cfbe97f3204034940d208e3f84b89ba8625779292b

Observation c837ddf0-367f-4f48-9c06-6fa5362361eb · inbound

Evidence-State Rewards for Long-Context Reasoning cites this paper.

Evidence-State Rewards for Long-Context Reasoning QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:28:18.210189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-03T13:25:14.844589Z digest=sha256:fa12eecdc5ec54dd8f133ca1a490e49d4e3247addbac47530cf4439ffd17a103

Observation 4e9f7c26-e957-4b75-b019-4d3c7c6d1dbf · inbound

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning cites this paper.

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T12:46:03.589144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:46:03.589144Z digest=sha256:b61143dcd9954509ddb9c0635255253b21a206ec732039ad40c85a5bd4e54cae

Observation 9f3e8f69-c2da-48ed-90f6-f9289cc9115d · inbound

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning cites this paper.

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T01:57:53.449669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:57:53.449669Z digest=sha256:2daeb6d0d1e30aa0db60a6adfe99cffeb36456dd59fbcc30a06b3593c626bf98

Observation ed144155-1032-4489-b5af-d358ca4d980e · inbound

REFACT: Adaptive Fact Restatement for Compact and Faithful Chain-of-Thought Reasoning cites this paper.

REFACT: Adaptive Fact Restatement for Compact and Faithful Chain-of-Thought Reasoning QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T09:18:46.780870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:18:46.780870Z digest=sha256:115392372b786a2a5677b80b004d88f35b81d6acb32ec15f134a56a35eb6ef88