Pith. sign in

Paper Citation Record · LEDGER

MAmmoTH2: Scaling Instructions from the Web

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2405.03548.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.03548 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:56:46.878443Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:32:33.164935Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b5358e2-adf0-4273-b65d-bdc78f894573 · inbound

MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark cites this paper.

MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark MAmmoTH2: Scaling Instructions from the Web

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:51:06.893449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T15:51:04.674346Z digest=sha256:3bd548ebfaa5ca131eae3682fd96e6487815b1b5a22afbdf624b6f5fd4d02489

Observation 48c3cc84-7f43-4c68-a089-361f6a39a680 · inbound

DataComp-LM: In search of the next generation of training sets for language models cites this paper.

DataComp-LM: In search of the next generation of training sets for language models MAmmoTH2: Scaling Instructions from the Web

Reference 210

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:58:17.348467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-17T22:58:16.523267Z digest=sha256:d1b0d1f7a86e7980608d40d19651e2bd9373a0db756c51c1206156077c05809b

Observation 923df388-c29d-4125-bfce-3ad63a44236e · inbound

Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs cites this paper.

Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs MAmmoTH2: Scaling Instructions from the Web

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-18T23:58:29.245776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T23:58:29.040819Z digest=sha256:70088189eddcc4d6f7d2ef394e624379a8ce5c2290b821560d80239eed58113f

Observation 3931144b-d780-4865-bbc2-90aa840b8f3f · inbound

On Domain-Adaptive Post-Training for Multimodal Large Language Models cites this paper.

On Domain-Adaptive Post-Training for Multimodal Large Language Models MAmmoTH2: Scaling Instructions from the Web

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T05:56:46.878443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:56:46.878443Z digest=sha256:602b14542e0f1470af19d16fb64cab8523dac3829860e85978897b00d5c70bcd

Observation 081081a9-a94d-42be-83bd-99b9c40cf147 · inbound

Towards Adaptive Mechanism Activation in Language Agent cites this paper.

Towards Adaptive Mechanism Activation in Language Agent MAmmoTH2: Scaling Instructions from the Web

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T05:09:47.993215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:09:47.993215Z digest=sha256:55682e43b2defd2ee540ece2f93b97a597626642928bda0820ca28e068dfdf92

Observation c261e836-d204-40ea-8271-d6f959f73968 · inbound

Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models cites this paper.

Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models MAmmoTH2: Scaling Instructions from the Web

Reference 230

Resolution
unresolved
no resolver link, observed 2026-08-11T22:57:02.141855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:57:02.141855Z digest=sha256:ff644177c148af1bce035e11d7eb591105df6551ac562e0449700cd0f3351f6b

Observation ead3c10e-3ba9-4af1-bade-d79474c01fdb · inbound

Evaluating and Aligning CodeLLMs on Human Preference cites this paper.

Evaluating and Aligning CodeLLMs on Human Preference MAmmoTH2: Scaling Instructions from the Web

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:53:15.343854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:53:15.343854Z digest=sha256:a450b73f953d1d8c3c7b7ec7e7783c25f034aeb26a0c60de3e485f6fb987bdc6

Observation 0c97adbc-f256-419b-9e92-7da13b58b7e4 · inbound

Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models cites this paper.

Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models MAmmoTH2: Scaling Instructions from the Web

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T13:58:43.794129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:58:43.794129Z digest=sha256:7d5448967f728919622dd516831cb0f94bdef8aeeea82162ea5373759d3a920d

Observation 630ecdc3-b899-4d5d-9015-2605c8c0af01 · inbound

Offline Reinforcement Learning for LLM Multi-Step Reasoning cites this paper.

Offline Reinforcement Learning for LLM Multi-Step Reasoning MAmmoTH2: Scaling Instructions from the Web

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T10:50:19.106362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:50:19.106362Z digest=sha256:75a33a99dad068ddea854dc195985d9decbbd0efc2afe41b079e878a0f7c177b

Observation 14e08398-944a-4a7c-9328-6ce04b2aabd3 · inbound

Visual Large Language Models for Generalized and Specialized Applications cites this paper.

Visual Large Language Models for Generalized and Specialized Applications MAmmoTH2: Scaling Instructions from the Web

Reference 293

Resolution
unresolved
no resolver link, observed 2026-08-10T22:08:10.037128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:08:10.037128Z digest=sha256:2b4881c006a78c6654eebc6349ce4f2e271580a9e5307c57f91817965330814e

Observation e2f1b138-f52b-4824-b168-1d3473c6ddab · inbound

CoddLLM: Empowering Large Language Models for Data Analytics cites this paper.

CoddLLM: Empowering Large Language Models for Data Analytics MAmmoTH2: Scaling Instructions from the Web

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-09T19:30:38.659184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:30:38.659184Z digest=sha256:fcc4f1cffe8c5bc09ae6153b8c5e99bda6b891d191862f04c6a16d82eec69bce

Observation 3a8e8894-dc67-41c4-ba23-63b27de10a6e · inbound

Process Reinforcement through Implicit Rewards cites this paper.

Process Reinforcement through Implicit Rewards MAmmoTH2: Scaling Instructions from the Web

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:23:30.879781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-11T20:23:30.763794Z digest=sha256:7d569e74e9ba8c4ca98060ba253fadc1beb88a13e5b9fb9a26a84e3afc8339c5

Observation 9e4c3ad5-16f6-4a88-8bac-4043bca9582a · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning MAmmoTH2: Scaling Instructions from the Web

Reference 243

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:33.168685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:01f24e3a0ad72b2c286fc13d26faad197f76abb29f449c2b6fb7f213ae4ea158

Observation f37a60ce-af3d-4dc3-b3d7-07bfa999b013 · inbound

CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction cites this paper.

CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction MAmmoTH2: Scaling Instructions from the Web

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T13:11:51.786167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:11:51.786167Z digest=sha256:2665a7b039f6bd53c5321cfc9040d944c8c1668d588a0861a511906c9b8fa9dd

Observation 5892fd4e-f7e1-4f70-b5ee-23efcc2377c5 · inbound

One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs cites this paper.

One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs MAmmoTH2: Scaling Instructions from the Web

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T11:01:03.524223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:01:03.524223Z digest=sha256:cb0787f003526385c901f2685c47284458eab5a8229116c468fed4ea3ecda14a

Observation 67f5e408-faec-4e24-b5f0-10d70d4cbaf9 · inbound

Beyond Templates: Dynamic Adaptation of Reasoning Demonstrations via Feasibility-Aware Exploration cites this paper.

Beyond Templates: Dynamic Adaptation of Reasoning Demonstrations via Feasibility-Aware Exploration MAmmoTH2: Scaling Instructions from the Web

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:12.491981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:12.491981Z digest=sha256:f4dbda732869dd2c9922d9a63c640dc331a240e4f252b23d41d5eeccd33678b0

Observation cf98c9ff-b42a-496e-9dde-b54d27aff825 · inbound

Scaling Reasoning without Attention cites this paper.

Scaling Reasoning without Attention MAmmoTH2: Scaling Instructions from the Web

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:13:28.299855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:13:28.299855Z digest=sha256:e30a77280a4df921f41406eee3781fefd93cadcccfb9f64e6470bd746b32c76c

Observation 2eca3266-6d6a-4d49-8891-57e20a7381df · inbound

A Survey on Large Language Models for Mathematical Reasoning cites this paper.

A Survey on Large Language Models for Mathematical Reasoning MAmmoTH2: Scaling Instructions from the Web

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:47.552410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:14:47.552410Z digest=sha256:208b7d97060772df6f9560fd077c6f9e3160267ce2c40439f6579bf7244359b7

Observation 33721b64-c380-41c5-a00a-e4e9b6b54103 · inbound

IFEvalCode: Controlled Code Generation cites this paper.

IFEvalCode: Controlled Code Generation MAmmoTH2: Scaling Instructions from the Web

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:34.126248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:44:34.126248Z digest=sha256:c26ca25703a1974e32de55f0c31cbdaceb4c2d4aa2571a23d4d5d8282a5d8326

Observation aae23b8d-5928-4e2a-9202-60cc7bc8c165 · inbound

Coupled Variational Reinforcement Learning for Language Model General Reasoning cites this paper.

Coupled Variational Reinforcement Learning for Language Model General Reasoning MAmmoTH2: Scaling Instructions from the Web

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T16:43:57.892407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:43:57.892407Z digest=sha256:2f9dcf6ac7ba8ede90f72486c3e0c5b7ed7ec623ffcc61f8f1e129a05e6ef053

Observation fecf1d13-a05c-4f37-a98a-64dbe28ed678 · inbound

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models cites this paper.

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models MAmmoTH2: Scaling Instructions from the Web

Reference 14

Resolution
malformed identifier
arxiv_id, observed 2026-05-11T08:21:00.336915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:41:37.207910Z digest=sha256:736a29803004ba221e8d5f1d39919f55b22fd6668eeeeda7c6ab5b1254644c36

Observation a0168488-637e-4804-bf57-5ccbbeb55544 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay MAmmoTH2: Scaling Instructions from the Web

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:b8b314f5f63885d9ca644a001ea996c162604ced49ce45cc99f5fa44fa025f3a

Observation 21481048-a15e-45ca-a546-1ff1ed5e919a · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay MAmmoTH2: Scaling Instructions from the Web

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:39.565899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:39.565899Z digest=sha256:72ab5ca076523b7e611d9bddb53d1df7d5b0d205dd0a1d21c0bdc86940d3fd34