Pith. sign in

Paper Citation Record · LEDGER

DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2305.10429.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.10429 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T04:17:27.127336Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:10:09.134017Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ddfd013d-5f9e-44ca-b6eb-4fb10d7b5e35 · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:46:40.480274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:c924203dae5d80e2c259ed593a77cd1e63ce9193b403f57afe2370c5b9973347

Observation f908cae6-8cff-4bac-8b2c-effc1ca44a04 · inbound

Llemma: An Open Language Model For Mathematics cites this paper.

Llemma: An Open Language Model For Mathematics DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 197

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:17:46.419903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T08:17:46.055279Z digest=sha256:61bc89e43397882c59d66c2a78e591ae660b795f76cf497c8cc31b4b8dfcf756

Observation 41496677-dd96-420c-a33b-7e258923b5e6 · inbound

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models cites this paper.

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T22:41:53.033134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T22:41:26.047957Z digest=sha256:e8065f62b4e5a84e8cf126dad959ca47a193979bc6e9c476f0b6d34c04c42f61

Observation d1cb1622-3927-4ef0-8546-d8c72aa83fd5 · inbound

Multi-Task GRPO: Reliable LLM Reasoning Across Tasks cites this paper.

Multi-Task GRPO: Reliable LLM Reasoning Across Tasks DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T04:17:27.127336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:17:27.127336Z digest=sha256:a0dd645ae122f2b0be41d381fd4ae90f521d1ac6af8b8527d3be5b409bd59e83

Observation 3fd118da-1829-4887-8de4-76839afb8443 · inbound

GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning cites this paper.

GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T21:05:26.534644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:05:26.534644Z digest=sha256:084b3356658d07c3ba01e766119a1970dcc9c397b39bbdd04e76c2d2f04ad588

Observation dbceb4dc-a267-431b-bb0f-2776d01a6a1e · inbound

Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings cites this paper.

Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:19:27.466265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:17:26.661595Z digest=sha256:1c9a40b43638d1461a2cdaba67d10ac18fd39b957ae40043bf26190fb4c82692

Observation 38133877-a792-4574-a582-bc635b36b31a · inbound

Ambient Diffusion Policy: Imitation Learning from Suboptimal Data in Robotics cites this paper.

Ambient Diffusion Policy: Imitation Learning from Suboptimal Data in Robotics DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:37:57.184917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T09:52:38.167538Z digest=sha256:106123b416300b63318ac4e2764777dd0d732e8cf9f0e7b4bedce48dc3a6c5ea

Observation a4e4bb3b-96a9-41fa-b64a-769f3af999c7 · inbound

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining cites this paper.

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T21:10:09.135670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-25T19:04:11.976747Z digest=sha256:46b2842ac8988d2741ffc40399c6ecb9777a3193fbb837b5d6476f3ed69bf55b

Observation c4ea9473-5c1b-4fbe-83dd-040053d71e34 · inbound

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures cites this paper.

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:48:39.445070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-03T16:44:41.720388Z digest=sha256:8c2b00541227edd92433e2ca00f95d338df34df73d52b13435041fd19470df0d

Observation 1fc5d40f-f130-4a86-9e41-b585abd7572d · inbound

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications cites this paper.

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T03:38:20.625282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:38:20.625282Z digest=sha256:2314b2d741ccf88897ce87ea0ee07e77fcb6afd6411b7c23f971e5ff26d99134

Observation 66f4baed-efb2-4bff-8050-86e767bad1c8 · inbound

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement cites this paper.

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement DoReMi: Optimizing Data Mixtures Speeds Up Language Model Pretraining

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.227138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T01:15:07.227138Z digest=sha256:82574e64d94e330a4886497cef979b7e241f6b42b5a3e8a6927103556880ed3c