Pith. sign in

Paper Citation Record · LEDGER

PolicyLong: Towards On-Policy Context Extension

As of 22 July 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2604.07809.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.07809 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T17:23:28.939977Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-22T06:31:00.163083+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact13
  • verified fuzzy2
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e72bf671-1ebf-4e80-9136-6bfb28b2cbd2 · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

PolicyLong: Towards On-Policy Context Extension LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:8b7c4b93bc50f2f26135f516441ac7863c3d9ccc48f90b84381c8c3970fd5f8d

Observation ad5657e9-05b0-4bbf-9105-554d8d40d940 · outbound

This paper cites What is Wrong with Perplexity for Long-context Language Modeling?.

PolicyLong: Towards On-Policy Context Extension What is Wrong with Perplexity for Long-context Language Modeling?

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:55:59.998175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:1ac828e894b8ec1a263f01adfab701efa69ab0445e31ca85acec860742c21d60

Observation 26ddc9bf-0b4b-4014-a8e2-123d184df4a3 · outbound

This paper cites Quest: Query-centric Data Synthesis Approach for Long-context Scaling of Large Language Model.

PolicyLong: Towards On-Policy Context Extension Quest: Query-centric Data Synthesis Approach for Long-context Scaling of Large Language Model

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:55:59.923299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:a6c014aceef659daedd9e95ded0eb64ce1976d140708cdda38753d84093fe7a7

Observation 08d6ee09-9f0b-4997-adf9-dd7c549db227 · outbound

This paper cites LongMagpie: A Self-synthesis Method for Generating Large-scale Long-context Instructions.

PolicyLong: Towards On-Policy Context Extension LongMagpie: A Self-synthesis Method for Generating Large-scale Long-context Instructions

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:55:59.889276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:474fd4d70330fead6390cc03ff94f7da5a7cc2f67e412082d6135ab295ff16d2

Observation ce3d8ed4-a950-453c-aead-d15727281add · outbound

This paper cites Reinforcement Learning via Self-Distillation.

PolicyLong: Towards On-Policy Context Extension Reinforcement Learning via Self-Distillation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:29:18.795354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:7734ee348b4969881f7dd11c6a8186b8202de5989de78789de75f54a5c8f2d42

Observation 475a8df5-5fba-425a-be89-411437f1e31e · outbound

This paper cites dots.llm1 Technical Report.

PolicyLong: Towards On-Policy Context Extension dots.llm1 Technical Report

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:55:59.972876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:ad11cb598362484961c4bd6c792447cf95963ad6732faf9359c0df5694e901ae

Observation 9b9cfb45-0072-43f4-a6db-3b39c7d5b6b9 · outbound

This paper cites Entropylong: Effective long-context training via predictive uncertainty.arXiv preprint arXiv:2510.02330.

PolicyLong: Towards On-Policy Context Extension Entropylong: Effective long-context training via predictive uncertainty.arXiv preprint arXiv:2510.02330

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:56:00.029286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:e219902b164c8d83729ab41a2ed2b16c8c695871ed7d9a12e44252d23b35e4c3

Observation 0d2ae284-82b1-45bf-92ea-5239231b1b39 · outbound

This paper cites Dense passage retrieval for open-domain question answering.

PolicyLong: Towards On-Policy Context Extension Dense passage retrieval for open-domain question answering

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T10:41:44.603036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:06fa1ef4e9127cc91043dd6e85de97dcb70c606b6ed7c9dd06cc4d77d0744fbc

Observation ad516e62-a6b3-4cb2-81a0-28040616191d · outbound

This paper cites DeepSeek-V3 Technical Report.

PolicyLong: Towards On-Policy Context Extension DeepSeek-V3 Technical Report

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-11T06:55:59.900611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:95c1e98a19b0f7cb121b65f1cff6ffb64e74b5b4ffd8964dec7659d37219681d

Observation 454faa73-26cf-4259-9d14-7bb8d6ad1570 · outbound

This paper cites Lost in the Middle: How Language Models Use Long Contexts.

PolicyLong: Towards On-Policy Context Extension Lost in the Middle: How Language Models Use Long Contexts

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-11T06:55:59.987957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:71f4b2b9b5c88b173ea9182630e6c27899bfca3f50e5a60405689ea206c48a7c

Observation 7f762f26-4a8f-4a86-9ff4-b8ac2656100b · outbound

This paper cites RocketQA: An optimized training approach to dense passage retrieval for open-domain question answering.

PolicyLong: Towards On-Policy Context Extension RocketQA: An optimized training approach to dense passage retrieval for open-domain question answering

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T10:41:44.600232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:e8726f55f23a143232a0e870676e3bfe2927c35ee33795659ad175237751410b

Observation dab12960-3ec0-434e-bb69-e8c42c87d8a6 · outbound

This paper cites Pope: Learning to reason on hard problems via privileged on-policy exploration.

PolicyLong: Towards On-Policy Context Extension Pope: Learning to reason on hard problems via privileged on-policy exploration

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:55:59.830779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:a18b4f953ef6d0b5f213980e238f910c38541caa0b1ccea8d582705d4c1f8d00

Observation 76009940-ca9a-417d-ab06-90693855083f · outbound

This paper cites Self-Distillation Enables Continual Learning.

PolicyLong: Towards On-Policy Context Extension Self-Distillation Enables Continual Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:52.723793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:200ced5b70650b09dde2c907dfa0589d014db1827431774e670085debe31847a

Observation a768ce4f-81f0-405a-b85e-6d66912a58e5 · outbound

This paper cites jina-embeddings-v3: Multilingual Embeddings With Task LoRA.

PolicyLong: Towards On-Policy Context Extension jina-embeddings-v3: Multilingual Embeddings With Task LoRA

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:55:59.819390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:960f7af69147481d1cde06e577eb886ab3807046506c75ba2785d5987e2a86b1

Observation 360a7722-9923-41b7-aa4a-43901ae9cbcf · outbound

This paper cites Effective Long-Context Scaling of Foundation Models.

PolicyLong: Towards On-Policy Context Extension Effective Long-Context Scaling of Foundation Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:55:59.842670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:b3108e674d241aa887aac1eb21fd54e1e9fa142b3b59d09fb86c108bab75f5e9

Observation 951da3d5-e3e3-4fcf-b748-87e9df62b2ce · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

PolicyLong: Towards On-Policy Context Extension HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:55:59.860500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:35c345b3f318890fb8bcb7b9e231211fe130bf7aa4dfdf0c45e1b8d1296973d4

Observation 4befafb8-c82e-4f37-b188-dd2ac56099df · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

PolicyLong: Towards On-Policy Context Extension Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:54:31.187112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:0ad4a0e56c3b305fc4cbc6255101b70f7fcf9468be3a841720d27ae583fedbaa

Pith citing papers

No inbound Pith citation observations are available.