Pith. sign in

Paper Citation Record · LEDGER

How Does Code Pretraining Affect Language Model Task Performance?

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2409.04556.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.04556 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:32:06.012307Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T17:26:04.409418Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8cb7fc23-b309-425f-8368-199465a71f3d · inbound

Trillion 7B Technical Report cites this paper.

Trillion 7B Technical Report How Does Code Pretraining Affect Language Model Task Performance?

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:32:06.012307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:32:06.012307Z digest=sha256:3a063a102de5cd70a96a738f82c97cf69099b3bd4dacc565e77dbf210fb27be4

Observation 3683302c-82a6-4084-b460-6c3e6781b608 · inbound

Mining Hidden Thoughts from Texts: Evaluating Continual Pretraining with Synthetic Data for LLM Reasoning cites this paper.

Mining Hidden Thoughts from Texts: Evaluating Continual Pretraining with Synthetic Data for LLM Reasoning How Does Code Pretraining Affect Language Model Task Performance?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:19:11.570810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:19:11.570810Z digest=sha256:c9efc1de8c3c1edba54a5663994540a5211e834aa2e3c309c3eeaa6e0ac3ecf9

Observation f626fef6-0c33-4241-9f59-6ac83b318122 · inbound

Transformers Pretrained on Procedural Data Contain Modular Structures for Algorithmic Reasoning cites this paper.

Transformers Pretrained on Procedural Data Contain Modular Structures for Algorithmic Reasoning How Does Code Pretraining Affect Language Model Task Performance?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:14:26.122419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:14:26.122419Z digest=sha256:d0db532631591a4ac405e95e3bd2ffe4b3f22bd6351c6027d24428fdab730ffe

Observation 23b21255-59a2-4805-899b-bae27e1b7322 · inbound

Procedural Pretraining: Warming Up Language Models with Abstract Data cites this paper.

Procedural Pretraining: Warming Up Language Models with Abstract Data How Does Code Pretraining Affect Language Model Task Performance?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T06:55:10.096836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:55:10.096836Z digest=sha256:8afafe5eec28a2d853a777a76c3797aafe5597ec9e791e72020a0851564331f5

Observation 77cd4a89-374a-4fb8-879c-7ee404eaf77d · inbound

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code cites this paper.

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code How Does Code Pretraining Affect Language Model Task Performance?

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:26:04.411843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T17:37:51.790000Z digest=sha256:cdc625e44031b81799084556553f35d4427e4200876c4b18a1adb07d40362229

Observation 2e9da44b-1d20-4420-9713-00a3807fbc6e · inbound

Can Transformers Really Do It All? On the Compatibility of Inductive Biases Across Tasks cites this paper.

Can Transformers Really Do It All? On the Compatibility of Inductive Biases Across Tasks How Does Code Pretraining Affect Language Model Task Performance?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T17:31:48.800897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:31:48.800897Z digest=sha256:f835be132c86e6f027ddf5b58c602204a7093d4ef5a263cb53fdfe89f4e35e61

Observation 1191244d-8c5e-462b-81de-2d793fb0875c · inbound

From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference cites this paper.

From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference How Does Code Pretraining Affect Language Model Task Performance?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T11:19:34.747234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T11:19:34.747234Z digest=sha256:a8ad86d2026439a2c427fb8220088e7c9fad4dde4ff2d9a0d7aa76e94f620385

Observation 53165c4b-6132-4969-ae8c-517aa021b344 · inbound

Can Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora? cites this paper.

Can Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora? How Does Code Pretraining Affect Language Model Task Performance?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T19:28:06.308321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:28:06.308321Z digest=sha256:8b149ef068888f2374659ced3f661f95f03717af743e47532e13e96fee989ae9