Pith. sign in

Paper Citation Record · LEDGER

Continual Pre-Training of Large Language Models: How to (re)warm your model?

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2308.04014.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.04014 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:37:10.174992Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:40:03.356996Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c5301845-833a-46fd-8dab-cabb8b2d7289 · inbound

Llemma: An Open Language Model For Mathematics cites this paper.

Llemma: An Open Language Model For Mathematics Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T08:17:46.396304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-19T08:17:46.055279Z digest=sha256:30286d39e145badcae5fa5980b3bdd5281b713ba783cc3c20117e3faa20ad3d8

Observation f09bb21c-04c7-4982-b59d-0b0cf85da29c · inbound

Sparse Upcycling: Inference Inefficient Finetuning cites this paper.

Sparse Upcycling: Inference Inefficient Finetuning Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:22:22.001463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:22:22.001463Z digest=sha256:daccde875e878c2f45f41bec37bb1d0addcdd7bb61e08cc10fc2a87f3ca48f8f

Observation 16865331-98b1-4acd-ad22-e5fad726a01a · inbound

ChronoLLM: A Framework for Customizing Large Language Model for Digital Twins generalization based on PyChrono cites this paper.

ChronoLLM: A Framework for Customizing Large Language Model for Digital Twins generalization based on PyChrono Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-10T21:51:13.963790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:51:13.963790Z digest=sha256:f66ec2a1c5fabe762a2389e00c7d6167a868fbac1c3f47e9b3d8ac4dfd311473

Observation e9f3f51e-c284-4f2e-b6ec-0c2672828d61 · inbound

Domain Adaptation of Foundation LLMs for e-Commerce cites this paper.

Domain Adaptation of Foundation LLMs for e-Commerce Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T19:48:29.247682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:48:29.247682Z digest=sha256:3341daa297239c41db9b9cebcfe8a3d45900a123cd3bcc4cd17de05b23801f0f

Observation adff1211-86d9-492c-b82f-6e1c5aa72b46 · inbound

Kuwain 1.5B: An Arabic SLM via Language Injection cites this paper.

Kuwain 1.5B: An Arabic SLM via Language Injection Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:37:10.174992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:37:10.174992Z digest=sha256:a9abad7a8a99d6322c2051349847207df3a165dd8dd4fdb7863493628cd4a61e

Observation 9e75d4c4-e974-4390-91f0-802c0112428e · inbound

DIMT25@ICDAR2025: HW-TSC's End-to-End Document Image Machine Translation System Leveraging Large Vision-Language Model cites this paper.

DIMT25@ICDAR2025: HW-TSC's End-to-End Document Image Machine Translation System Leveraging Large Vision-Language Model Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:47:47.500876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:47:47.500876Z digest=sha256:ce268aef7d37ac1c996fdc310ef6a9ff6c200762947c259113a09d501ee327e8

Observation de5b615b-9b21-455f-b664-3f61993c92e3 · inbound

WenyanGPT: A Large Language Model for Classical Chinese Tasks cites this paper.

WenyanGPT: A Large Language Model for Classical Chinese Tasks Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T05:28:06.108062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:28:06.108062Z digest=sha256:652c84dc1031a60098157c171e40f27fbaf2b1df083a984e70c74dc38749b1a9

Observation 493bde02-200f-426a-9ae8-9b5ec1aabcef · inbound

EnronQA: Towards Personalized RAG over Private Documents cites this paper.

EnronQA: Towards Personalized RAG over Private Documents Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T04:51:30.062833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:51:30.062833Z digest=sha256:e7bdc94b6a73669aa5c5b40386e4264bd795dc787a64b0af6ab78262e94d3343

Observation 601c7a09-2686-4105-9054-b674a8c066a1 · inbound

A Survey on Foundation Models for Personalized Federated Intelligence cites this paper.

A Survey on Foundation Models for Personalized Federated Intelligence Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-05-22T15:34:57.809940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T15:32:15.293888Z digest=sha256:5edc1c413bad48e41c5de68784e48afb7787894adb41a7b3700a58dcde131d7a

Observation 8f197b83-e306-4a0c-9dcb-19b67cd9a90b · inbound

Learning Dynamics in Continual Pre-Training for Large Language Models cites this paper.

Learning Dynamics in Continual Pre-Training for Large Language Models Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T22:14:19.146891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T22:14:19.146891Z digest=sha256:22100de51270769bd2a5bc2f46088ebc469e2791270ec1f1186c1a185da85664

Observation b848ad5b-8d1d-4b90-bc5f-7b1852254379 · inbound

Universal Music Representations? Evaluating Foundation Models on World Music Corpora cites this paper.

Universal Music Representations? Evaluating Foundation Models on World Music Corpora Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:58.797431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:58.797431Z digest=sha256:33653446506f9847c78711877fc767bbafacb1289e57702b89be85dd9b1f1490

Observation 5d159340-43fd-451e-8eb1-9b15311bfe25 · inbound

CultureMERT: Continual Pre-Training for Cross-Cultural Music Representation Learning cites this paper.

CultureMERT: Continual Pre-Training for Cross-Cultural Music Representation Learning Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:07.406000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:07.406000Z digest=sha256:620afc24737726fcb06c21d9d17a048ed50cc65d43c6369956b249e815618785

Observation 4eabadc4-db22-4775-b454-37d42edf1f43 · inbound

Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling cites this paper.

Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:36.670881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:24:36.670881Z digest=sha256:ad919338258ee637a0393e73ba3394591cd6f471ae82f1c5c43d0160b5630ccc

Observation d2d51962-ef7d-42be-9707-e99936108d36 · inbound

From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space cites this paper.

From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:41:04.228501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T12:50:57.603403Z digest=sha256:0de7e581664b072420a62bbd9ddf26bac1eb87b4b4d9cd3823dc069f81d1cc06

Observation 17e2301d-ef36-4a92-8308-7ef0439a43b9 · inbound

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts cites this paper.

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T03:29:21.464973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T03:29:16.555166Z digest=sha256:dfd121eefb9b37acae8b3eb8536222ec86c6cb21aaa4fccddf43dfd86c89cd35

Observation 120a1ce5-b8c9-461f-987a-af4f801ec959 · inbound

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts cites this paper.

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T02:06:15.399440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T02:03:02.654035Z digest=sha256:64191ce2107a094e4980ebe04baca8fcb5f8ab43831990fafb33342c04585c2b

Observation 8720233f-776b-4c41-9c6c-287b845a177e · inbound

ZAYA1-8B Technical Report cites this paper.

ZAYA1-8B Technical Report Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 93

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:26:05.212027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T17:36:37.182196Z digest=sha256:917cdbe159e30a0ab15bdcef22c8f5fae9473b65d265c471445fb6aaa7d3cc0e

Observation cb5e2ed3-df43-49bb-be18-a31c84c2d041 · inbound

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use cites this paper.

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:45:05.604780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T05:44:58.483183Z digest=sha256:75778fcfdfd7097f03a05616bf38bd5eedfeaffcb30b3536eed90efcbbf71fb8

Observation 3e7242d2-46c2-4ab6-98f1-bf1dd8e004de · inbound

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use cites this paper.

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T21:03:46.453219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T21:02:51.889450Z digest=sha256:c6560da1738ce861c573ea998b5f9c42a5edc0196d22240d93f18a0a9450700f

Observation d394e1f9-35de-4c5a-b308-ce39d8361f77 · inbound

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use cites this paper.

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T09:31:22.849196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T09:29:30.950904Z digest=sha256:7a73ceb1786868c1a9916c7374f53c07338b906e4c999fa7e1d88136fc0a9b21

Observation 4483488a-4caf-4876-8634-b89acfb0ad13 · inbound

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay cites this paper.

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T23:24:02.030496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T23:15:45.174086Z digest=sha256:90209f4e7c9b2d9c216cd8560d0a3c1fa78d72b1cc30c65cfe3a4f331128b6d6

Observation 6e0dd535-a6ad-4725-8d4f-875443b23e36 · inbound

Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training cites this paper.

Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:56.694420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T01:53:04.715108Z digest=sha256:26c95dcb90c4c9d7f382b54d825f26a610a060c116ef1154530e947013548127

Observation f108a764-8d7c-4684-bfc9-62af0f54910c · inbound

How Post-Training Shapes Biological Reasoning Models cites this paper.

How Post-Training Shapes Biological Reasoning Models Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:55:31.039643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-01T07:48:31.110861Z digest=sha256:faf79c656b7ece65a273b57defd5784a801b0414a2edacb0b38ee136a57d1a89

Observation 8d235313-7d98-484d-ba09-beda400464bd · inbound

ZONOS2 Technical Report cites this paper.

ZONOS2 Technical Report Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 119

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T18:40:03.358532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-25T22:37:15.072758Z digest=sha256:b2028946ff08e455d96034a48b3ef1ad3f2a697453b08cd9d350d76c8fb8b657

Observation 64b77a19-8ab0-4af1-942a-8d964300f4fc · inbound

ZONOS2 Technical Report cites this paper.

ZONOS2 Technical Report Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 119

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:15:59.016168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-29T02:07:31.791835Z digest=sha256:92d54129beddc45d0913e2cc05285a8abb25bdfc7f2417ba34398e73f672fc13

Observation 14da3bf0-4d5a-4635-9db6-74111616e6a9 · inbound

WSqD: A Horizon-Free Learning Rate Schedule for Large Model Training cites this paper.

WSqD: A Horizon-Free Learning Rate Schedule for Large Model Training Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T08:04:06.432613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T08:04:06.432613Z digest=sha256:c3cb71c53386e16e76a0a68d5cc1804fe8f0a9019b954c102b7338762fe62b64

Observation b510c958-6b20-475c-b8eb-991fa208eee5 · inbound

Scaling Point-in-Time Language Models cites this paper.

Scaling Point-in-Time Language Models Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T15:39:37.195076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:39:37.195076Z digest=sha256:25f140db438902c97c032fe18ea23a4776a4ac14b2a9ab546c7c01ab1ff6a7e2

Observation 8527eb4e-6c61-47ef-93e0-7a084bd673d7 · inbound

ZUNA1.1: A more flexible EEG foundation model for Denoising and Super-resolution cites this paper.

ZUNA1.1: A more flexible EEG foundation model for Denoising and Super-resolution Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-01T09:52:00.140890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T09:52:00.140890Z digest=sha256:410e98bbf53755934a7c0440a62f314f9c711269e552f5280dc48635784ba0c0

Observation 986273d6-f239-4b16-a25a-03c8e668e9e3 · inbound

MedLLM: An Open Medical Language Model at the Sub-Billion Scale cites this paper.

MedLLM: An Open Medical Language Model at the Sub-Billion Scale Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T07:02:04.080396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:02:04.080396Z digest=sha256:3a06decd38e556cc86eea008725a0f932dbd3793e30b7d14182588d3a3d6ff98

Observation 7b87762a-6bd0-46de-b9dd-0e73a90bc0a6 · inbound

Continual Learning in Transition cites this paper.

Continual Learning in Transition Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:24:00.347438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:24:00.347438Z digest=sha256:89902724cb20394fd71f77b49b5da21564f3e34ab9d3dbdfba7ecf0873282239

Observation f134cbd8-d380-47cf-a15a-1ae4601417ce · inbound

Continual Learning in Transition cites this paper.

Continual Learning in Transition Continual Pre-Training of Large Language Models: How to (re)warm your model?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T14:38:35.507340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:38:35.507340Z digest=sha256:39d2a8b2efde7ce112d19ee1d279f476f16552eb081d65c46f9f1fc6cf63fc7c