Pith. sign in

Paper Citation Record · LEDGER

Training Trajectories of Language Models Across Scales

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2212.09803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.09803 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:41:46.230224Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eca6dce9-0940-423c-bb4c-a39f29078657 · inbound

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling cites this paper.

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling Training Trajectories of Language Models Across Scales

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T17:45:17.903448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-15T17:45:17.540282Z digest=sha256:b15215d1dbde0e78838fe1be7578b0c2bb67cd069914664eb104414237f17caf

Observation 50b8710f-c728-4e18-9c14-b2f87676ecc5 · inbound

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling cites this paper.

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling Training Trajectories of Language Models Across Scales

Reference 180

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T17:45:17.674417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-15T17:45:17.540282Z digest=sha256:1e95728cfa22bb8e815a58d1e67975cd36435a59bc1871069b89c43469313747

Observation 7e2a865d-2e55-4b95-a984-98098a30b8f6 · inbound

Scaling Data-Constrained Language Models cites this paper.

Scaling Data-Constrained Language Models Training Trajectories of Language Models Across Scales

Reference 131

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:35:21.479135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T01:35:21.150772Z digest=sha256:a493aea50f981f3f5edb3efb29f051749f1a97c8f4cc67cbb7872038e6d75f92

Observation 02c817c7-1e25-4563-88a2-6c11c8826ec6 · inbound

Predicting Emergent Capabilities by Finetuning cites this paper.

Predicting Emergent Capabilities by Finetuning Training Trajectories of Language Models Across Scales

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T13:41:46.230224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:41:46.230224Z digest=sha256:6cd4eb24e201e7c122a5489840df1c58ee43e992571bbd5bf500ddca3f8339da

Observation 177a0653-5804-4c5b-88cc-80f483b72fff · inbound

Predictable Emergent Abilities of LLMs: Proxy Tasks Are All You Need cites this paper.

Predictable Emergent Abilities of LLMs: Proxy Tasks Are All You Need Training Trajectories of Language Models Across Scales

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T19:11:48.516890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:11:48.516890Z digest=sha256:92db60282dc3a215e3cc4ad239c715183014a3d02c5b2ce63ae5dd9e6b378c6d

Observation 9bd4b689-2dfa-4773-81db-8605e8572abc · inbound

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model cites this paper.

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model Training Trajectories of Language Models Across Scales

Reference 219

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:02:23.501537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-19T08:02:23.002090Z digest=sha256:4ec1c0e24644a68131135a744fcc36dd2e3bdd47b3afa994332e58fde24c89c3

Observation 50842ce7-c236-41b2-a326-fb7d0b61af28 · inbound

How much do language models memorize? cites this paper.

How much do language models memorize? Training Trajectories of Language Models Across Scales

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.259204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.259204Z digest=sha256:e37050b3f8627e59c12d4ee9e7efefebe6f841f64f57249ab97ee36e58f1b8dd

Observation 472b5137-f1ae-45ce-90c0-e4df3b90f4dd · inbound

Fairness Dynamics During Training cites this paper.

Fairness Dynamics During Training Training Trajectories of Language Models Across Scales

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:39:52.924492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:39:52.924492Z digest=sha256:b29ba47ad276814865a8933f5ee40ec912237ad93947cd286239d9cf1e9fd014

Observation 87b1c8c5-13f9-4c96-aac4-392dacadbba9 · inbound

Capability Salience Vector: Fine-grained Alignment of Loss and Capabilities for Downstream Task Scaling Law cites this paper.

Capability Salience Vector: Fine-grained Alignment of Loss and Capabilities for Downstream Task Scaling Law Training Trajectories of Language Models Across Scales

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:02.700113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:02.700113Z digest=sha256:f4b142a64ca4d3c6abe0408715bdb89b9f9043ceededc492d28732bab0be5a7b

Observation e71ca646-0868-4ca5-873d-b501270a77b0 · inbound

Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent cites this paper.

Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent Training Trajectories of Language Models Across Scales

Reference 226

Resolution
verified exact
arxiv_id, observed 2026-05-20T01:32:55.873440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-20T01:29:14.555216Z digest=sha256:ecbb33616dd323dce66211f5372f7b3f2b75c4e9a06ff9ea8f96b79464204d6d

Observation 0a19b94e-d449-4f28-9d41-7f6768ef0bdb · inbound

Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent cites this paper.

Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent Training Trajectories of Language Models Across Scales

Reference 226

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:40:24.922360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-25T06:39:16.246591Z digest=sha256:3d1a82c41a3ba1295e7a2a223edcca9915da49e242440039a8e2bc2fb4803443