Pith. sign in

Paper Citation Record · LEDGER

Efficient Online Data Mixing For Language Model Pre-Training

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2312.02406.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.02406 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:33.532810Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0bead57a-b683-4937-8140-0549c6ce35a4 · inbound

DataComp-LM: In search of the next generation of training sets for language models cites this paper.

DataComp-LM: In search of the next generation of training sets for language models Efficient Online Data Mixing For Language Model Pre-Training

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:58:16.837009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T22:58:16.523267Z digest=sha256:9f29fc832830d7e5a531f09d80aee62daa7158e976520c7ae0bdaac3bc51e3a0

Observation 4bd057bc-eb32-4353-914d-1e15184bd867 · inbound

DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks cites this paper.

DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.417636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T03:58:48.967122Z digest=sha256:804d25e81b530c622be163e20e0ac6a4d9c436185f5cf19baad76d7aa8881fc8

Observation de4514b8-5107-4316-909a-45c8d58ea7e1 · inbound

Merge to Mix: Mixing Datasets via Model Merging cites this paper.

Merge to Mix: Mixing Datasets via Model Merging Efficient Online Data Mixing For Language Model Pre-Training

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:33.532810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:33.532810Z digest=sha256:c2f96df0d7f39465fc7457211e6074303b71c95393efbfb1cd5de4aad8a73238

Observation 4f784259-f93c-48de-97e0-b32d974601ba · inbound

GRAPE: Optimize Data Mixture for Group Robust Multi-target Adaptive Pretraining cites this paper.

GRAPE: Optimize Data Mixture for Group Robust Multi-target Adaptive Pretraining Efficient Online Data Mixing For Language Model Pre-Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:24.571162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:04:24.571162Z digest=sha256:6b745acf0128d3532ab6455d0db676ffea90f6cf86cad966bd1805ebec2a0b06

Observation 62f76901-e1cd-4dd5-b557-a4ee9abbe3a4 · inbound

Rethinking Data Mixture for Large Language Models: A Comprehensive Survey and New Perspectives cites this paper.

Rethinking Data Mixture for Large Language Models: A Comprehensive Survey and New Perspectives Efficient Online Data Mixing For Language Model Pre-Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:48.487199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:48.487199Z digest=sha256:e105d5844e9d1631d86f8387f7db0d8a8867599f155641da47142bb6e9418f1a

Observation f3e214dd-d14f-4ad1-af5b-4b898e03b140 · inbound

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning cites this paper.

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:19:26.845188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:19:26.845188Z digest=sha256:ad33439768bdd918751fc3be4c7a270c0795401ab670ee46ae6ddfb44ccdfc37

Observation 44a757b4-b7af-4777-8df1-10dfab028887 · inbound

AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs cites this paper.

AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:26.935097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:26.935097Z digest=sha256:ffe42a4c056cffeb54e0846b574b6f8ad583ee0cee233ce581dca5e9b15f04bd

Observation 15210643-f738-44b6-9b55-0efb7dff02c0 · inbound

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text cites this paper.

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text Efficient Online Data Mixing For Language Model Pre-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:42.562902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:29:42.562902Z digest=sha256:626772fcf8e33eb067dc02c0486d1dc9f54eec13e3e1f65c8aebf5edf49d8e1c

Observation 917c7a5c-e1c5-43cc-ba0e-1d04c9a4f981 · inbound

Language Models Improve When Pretraining Data Matches Target Tasks cites this paper.

Language Models Improve When Pretraining Data Matches Target Tasks Efficient Online Data Mixing For Language Model Pre-Training

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:53:07.823581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:53:07.823581Z digest=sha256:107979ebcf91d34025427435d3e30502d32e3d7ffd22a376fd8b80259d8c8bea

Observation 7bda05d1-b601-47aa-ac2f-10d9cc52fb0d · inbound

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining cites this paper.

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining Efficient Online Data Mixing For Language Model Pre-Training

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:08:12.635236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:07:29.544613Z digest=sha256:3696c9917f80127dcfaf8ca29532bc1855a13fce839741fd0d2b5e2ea2664642

Observation 29625642-4fda-4ab9-a4d1-8621cc8980ab · inbound

Data Mixing for Large Language Models Pretraining: A Survey and Outlook cites this paper.

Data Mixing for Large Language Models Pretraining: A Survey and Outlook Efficient Online Data Mixing For Language Model Pre-Training

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:25.751611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:56:04.958757Z digest=sha256:360d49c0ce101f708774db392d04814e7d5717b949ce5b727c81d041c776c3a6

Observation a2866833-2709-4d6f-bfaf-a598fcbefb75 · inbound

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling cites this paper.

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling Efficient Online Data Mixing For Language Model Pre-Training

Reference 268

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T03:08:59.654681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T03:05:36.871497Z digest=sha256:65971a8536e6151cdc44d3ece10759f0fbc4d5ea3d85ff290206ff6215e933d1

Observation 978961dc-3dc4-47d2-9e4f-f5450bcf27c9 · inbound

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them cites this paper.

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them Efficient Online Data Mixing For Language Model Pre-Training

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:22:46.209942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T23:19:22.355753Z digest=sha256:96211e2ec89591860998e81e4d2ed080acbc8a2e64ba995de68dc829b952cd74

Observation 80073e1f-33d6-402a-9663-849359df43e1 · inbound

Explaining Data Mixing Scaling Laws cites this paper.

Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:37:22.526097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T20:20:53.599064Z digest=sha256:afa8452460a092ecd37b582de4482565e54946d05336c9bfe93135205a2167f1

Observation a3e95546-417f-496d-b6ee-c29f887664a7 · inbound

Explaining Data Mixing Scaling Laws cites this paper.

Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-15T10:54:16.434702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:54:16.434702Z digest=sha256:1903666af22a8cfc6063e86215cf053aac6d5a7893430c83c5fa7ea9e9841873

Observation fcbbb887-5ef7-4809-91a6-a7d7a2c0f8ee · inbound

Explaining Data Mixing Scaling Laws cites this paper.

Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T12:13:45.639396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:13:45.639396Z digest=sha256:aa94ff9121d4a0aaf53dff1723abb324c4018ce1e30025c169ac219b49d94bc9

Observation 7b1c2692-3879-48e7-8864-d1c738009978 · inbound

DRIFT: Refining Instruction Data via On-Policy Data Attribution cites this paper.

DRIFT: Refining Instruction Data via On-Policy Data Attribution Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T19:18:54.754640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T01:57:05.784589Z digest=sha256:73cafd2ca3fdb9939c6ba32bbdab1e6e5c1ac561d0fc8a8aa51cbdd64292eca0

Observation 2e0072df-b689-447b-a45f-cd91b1e4564a · inbound

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning cites this paper.

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:19:57.827829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T00:46:13.587037Z digest=sha256:1c315f048cdee7bbdc0fa625581e5eca89cb6cad4c043c892d699ff7ec4f88a2

Observation 4c3cf89b-205f-46bc-b4ad-a8646687d9e1 · inbound

Smooth Scaling Laws Hide Stepwise Token Learning cites this paper.

Smooth Scaling Laws Hide Stepwise Token Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:14:26.553661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T06:08:22.652557Z digest=sha256:e4b5e9c28c674d01df3ee62ee2b082ba19330f85fb82c39932015f2d914debb5

Observation c6a2181e-ab22-42fa-94b5-ecccc8901c7d · inbound

Smooth Scaling Laws Hide Stepwise Token Learning cites this paper.

Smooth Scaling Laws Hide Stepwise Token Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-13T07:15:03.029543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:15:03.029543Z digest=sha256:296227ea7e2b44bf17e58e4bb5542f3246dc4c5064f76441768a63760c57aae3

Observation 64423c53-a41e-4643-9677-4e878e0b15e2 · inbound

WARP: Weight-Space Analysis for Recovering Training Data Portfolios cites this paper.

WARP: Weight-Space Analysis for Recovering Training Data Portfolios Efficient Online Data Mixing For Language Model Pre-Training

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:58:46.729349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T17:54:31.856386Z digest=sha256:213bde0e7228bf28405419cc9743c0d87a906bd58a66d02b5f58fded3aafb1f6