Pith. sign in

Paper Citation Record · LEDGER

Examining Forgetting in Continual Pre-training of Aligned Large Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2401.03129.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.03129 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:55:45.920900Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:48:20.831541Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f6e6bb5d-3215-4633-91c4-4f1b654fd964 · inbound

Large Language Models for Material Property Predictions: elastic constant tensor prediction and materials design cites this paper.

Large Language Models for Material Property Predictions: elastic constant tensor prediction and materials design Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T17:49:14.192656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:49:14.192656Z digest=sha256:51b6a9a6d2c59bc97430261ad32d43e7ec0345122d40fc5ccc7211017074eff9

Observation 86cf26ec-675b-46d1-a72f-691e5dc7c5b9 · inbound

Chained Tuning Leads to Biased Forgetting cites this paper.

Chained Tuning Leads to Biased Forgetting Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T10:38:42.035559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:38:42.035559Z digest=sha256:0b917db1d12e7ee83067c6cf25b1ef5c47b0f5c013df28370f9b84d2df9d0248

Observation 14dfb17d-7526-4be4-bafc-e6ba73337cc6 · inbound

Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging cites this paper.

Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T00:18:57.215751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:18:57.215751Z digest=sha256:37cc07db1d8164a4f85e1dcee7b44cc6ec60892e96a123f7ac5d982991e0c0ff

Observation f642ff60-0197-4dd5-b20a-0de9ba0a6605 · inbound

Full-Parameter Continual Pretraining of Gemma2: Insights into Fluency and Domain Knowledge cites this paper.

Full-Parameter Continual Pretraining of Gemma2: Insights into Fluency and Domain Knowledge Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:55:45.920900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:55:45.920900Z digest=sha256:84fe84943ee6258e9933a21c4c0b8fe128f0d8ef63e9186a9223136985da6ef2

Observation 2a9180e8-9065-4d9c-88cf-e10125918815 · inbound

The Future of Continual Learning in the Era of Foundation Models: Three Key Directions cites this paper.

The Future of Continual Learning in the Era of Foundation Models: Three Key Directions Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T11:10:14.591621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:10:14.591621Z digest=sha256:4bb3c4c973045d7cca9ed8de594d1d62cd1534352000f429930c1cdc0fe5b3d2

Observation aa24de08-ac04-4c7c-afc1-dbd5d3a855b7 · inbound

Software Engineering for Large Language Models: Research Status, Challenges and the Road Ahead cites this paper.

Software Engineering for Large Language Models: Research Status, Challenges and the Road Ahead Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 176

Resolution
unresolved
no resolver link, observed 2026-08-06T21:36:36.500838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:36:36.500838Z digest=sha256:16a13300f54b7fe5615e8b4b1e42b3c97484b83d5062c24c22b047192a068a1f

Observation fa4d997f-22b2-4ef9-a464-7ee36f46cdab · inbound

TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale cites this paper.

TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:07:41.600164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T17:06:12.280460Z digest=sha256:71436990355a3492dd027cd48a60965750b02b23075cffd7087a137696e1a856

Observation ed0aad0a-f824-40c9-918a-3c379bb87050 · inbound

Max-Window Scale Estimation for Near-Lossless HiF8 W8A8 Quantization-Aware Training cites this paper.

Max-Window Scale Estimation for Near-Lossless HiF8 W8A8 Quantization-Aware Training Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T23:04:00.899086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T23:02:56.306685Z digest=sha256:6cffca64944d7b1a93372a4d16a2c92d7a1d90a57c53d879143f69b57c069e21

Observation 7f374d50-fc67-4f01-b9b7-4f0970afb7b8 · inbound

SupraBench: A Benchmark for Supramolecular Chemistry cites this paper.

SupraBench: A Benchmark for Supramolecular Chemistry Examining Forgetting in Continual Pre-training of Aligned Large Language Models

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:48:20.832777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T07:34:11.596337Z digest=sha256:83dd799e76b5fd3983bdeb9a7724990511ec43be07b15f1c1891dafc566a9590