Pith. sign in

Paper Citation Record · LEDGER

Continual Pre-training of Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2302.03241.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.03241 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:40:23.861194Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:24:01.963863Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 921063f1-6e6e-4ce5-8f3d-19c5d5d8f569 · inbound

Optimization Hyper-parameter Laws for Large Language Models cites this paper.

Optimization Hyper-parameter Laws for Large Language Models Continual Pre-training of Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:45:48.948200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T20:45:31.427677Z digest=sha256:f854f7d01310d2d8d54cfecccfdddf4ee2712abaef75c65113ff701fdd14247b

Observation 9249228c-1fdd-4188-b2ac-11b50a54c596 · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Continual Pre-training of Language Models

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:23.861194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:23.861194Z digest=sha256:374b9ccddf16605098cea4dac7a73c2906426b476fcddcb695b87b47a01065ae

Observation c3a28d15-4050-4a44-a6d8-1de3b97291e5 · inbound

Bisecle: Binding and Separation in Continual Learning for Video Language Understanding cites this paper.

Bisecle: Binding and Separation in Continual Learning for Video Language Understanding Continual Pre-training of Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:09.647212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:09.647212Z digest=sha256:2786c48d525c65b434e99633fb3e940c565bd19174e84c82c63a8c622e99ea0c

Observation e439bcc1-23b8-4871-a4ca-dbc8711b2e83 · inbound

AI-Assisted Fixes to Code Review Comments at Scale cites this paper.

AI-Assisted Fixes to Code Review Comments at Scale Continual Pre-training of Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:29:47.708700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:29:47.708700Z digest=sha256:89da431f1c56dfe7f2486995607c5ee8c97be53b25abcf391ba643dd1564c560

Observation 72de1ada-a28b-4c56-aa12-d5c0e60c7a24 · inbound

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector cites this paper.

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector Continual Pre-training of Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:42:47.455682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T17:39:17.456350Z digest=sha256:1753608d6e44ab609880fe295656e12a5c38fbd0d776c47ed321424d8a36962a

Observation 911fdbd9-5e8c-43f8-aa34-e0a881a701ff · inbound

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization cites this paper.

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization Continual Pre-training of Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:51:30.401187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T14:46:45.598080Z digest=sha256:722864662681dcc5144cd8ee784d5f76aba74c794cd3c91e55e27e16db506d61

Observation e65f9345-c86b-4de6-9f95-d43377b06f5c · inbound

Learning to Discover at Test Time cites this paper.

Learning to Discover at Test Time Continual Pre-training of Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:16:04.138617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T05:16:04.001700Z digest=sha256:8507ce899331369cf92f5eb875c9c7ecd82b88c8901a7ceb6e40ed2809757e21

Observation ae411208-6855-4ce0-93bb-4518282acf59 · inbound

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis cites this paper.

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis Continual Pre-training of Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.664096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T09:34:23.475942Z digest=sha256:90e777d69b1c700bc61b38d279c13bc4d54a3fa0b01e419bd4f408f21c9d5b00

Observation 24e3ec75-1195-4eac-84a1-ad6c5708bdbb · inbound

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration cites this paper.

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration Continual Pre-training of Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:36:58.134724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:36:58.134724Z digest=sha256:15a8499723d8f3650b7c4d054bc0238414c4c3016d210f66f9ad93f122aeb66f

Observation 1a5d88d4-be8b-4722-9808-66c2265a11b0 · inbound

HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning cites this paper.

HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning Continual Pre-training of Language Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:15.049738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-08T11:25:01.780099Z digest=sha256:bb6567441acb9716ff9b18dd62a02470921c4f14953c97d99c92baaeb9d3ddca

Observation 55306f03-a062-40a7-b64f-62c43e096524 · inbound

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm cites this paper.

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm Continual Pre-training of Language Models

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:56:25.597688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:47:54.466097Z digest=sha256:7020b00ebacc01cfe22297e64c737282535a77ec1201a96ef2242672813d6c30

Observation f42279f7-5a37-4751-8db8-b59e099fcdc3 · inbound

MeMo: Memory as a Model cites this paper.

MeMo: Memory as a Model Continual Pre-training of Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:19:43.393031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T03:17:23.202604Z digest=sha256:c0a66f10d7667cfb0b4abb528143468e8dc8f8de45fdb185ea40fd4affb0454c

Observation e5d8e25b-7f4b-4b5a-84c9-9a33b867ef17 · inbound

MeMo: Memory as a Model cites this paper.

MeMo: Memory as a Model Continual Pre-training of Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.612682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T08:36:31.022046Z digest=sha256:f985f64ff23fba93006e4cb9fcd04fb3835e3c1a8680f170ad6000539563e051

Observation 74830d9e-524c-4ee6-909d-e17a937d0c3b · inbound

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining cites this paper.

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining Continual Pre-training of Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:49:49.898714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T07:49:13.043266Z digest=sha256:b1cc28c30affccb65f4c1db068909d59494ecf44f985cb60a45e61949a6ec6ea

Observation a63f0fc3-a9dc-4d0f-8893-f6a0ea715761 · inbound

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay cites this paper.

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay Continual Pre-training of Language Models

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:01.965352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T23:15:45.174086Z digest=sha256:d96d9ab0ee14ed3741ea84bdd4a3cc66df62b7aa5d06c283b831b289595815f9

Observation a6e68a9f-f59c-4174-9689-4b9adf2fb501 · inbound

Scaling Point-in-Time Language Models cites this paper.

Scaling Point-in-Time Language Models Continual Pre-training of Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T15:39:37.977391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:39:37.977391Z digest=sha256:c2a8a4bb02638d2199bd71e2ef63987e5cf7255e252767ef632c32e0bc8dd251

Observation 89c1bf4c-5c67-45d0-a258-d09e19262ba6 · inbound

Learning to Prepare Molecular Ground States with Transformer Models cites this paper.

Learning to Prepare Molecular Ground States with Transformer Models Continual Pre-training of Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T04:45:23.593279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:45:23.593279Z digest=sha256:aa92b19df4cb9227708b1da3121023e1d0da33d02423fd9dbd862712b578a0df