Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2402.09668.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:03.600455Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 3d59eb21-f235-4807-978f-cb47c3275d11 · inbound
A Survey of Large Language Models How to Train Data-Efficient LLMs
Reference 237
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 792957dc-e756-4ca0-bad0-21339c6cd949 · inbound
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 149a544f-2fdb-4df2-bf3a-ca4639ca9b4b · inbound
DataComp-LM: In search of the next generation of training sets for language models How to Train Data-Efficient LLMs
Reference 157
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e7b6d1c-21d2-49e5-8281-2fa89251ff4d · inbound
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 008d596f-f874-4fec-ae02-f2b8d97f73f7 · inbound
Enhancing LLMs via High-Knowledge Data Selection How to Train Data-Efficient LLMs
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75691a37-0af0-4eb6-8cbb-6578d75e4297 · inbound
FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain How to Train Data-Efficient LLMs
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7be11a93-fc4f-456a-b1ce-c09431db1d5d · inbound
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining How to Train Data-Efficient LLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e5f661e-1df7-4c21-b0f1-6e0f8d337875 · inbound
Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets How to Train Data-Efficient LLMs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a2bb00-27e7-413c-aad1-04bfb0365e04 · inbound
Judging Quality Across Languages: A Multilingual Approach to Pretraining Data Filtering with Language Models How to Train Data-Efficient LLMs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baf9518d-8283-4d8e-9693-627a73a6bf4c · inbound
Truly Self-Improving Agents Require Intrinsic Metacognitive Learning How to Train Data-Efficient LLMs
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d290ed7-dc77-45a3-b4e6-29d20b1418fa · inbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation How to Train Data-Efficient LLMs
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a1f4957-80ec-4cd3-a2fd-c1a72c8777ca · inbound
Assessing the Role of Data Quality in Training Bilingual Language Models How to Train Data-Efficient LLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98bac85f-4f08-4101-acf8-8a2b1e0ef229 · inbound
Disentangling the Roles of Representation and Selection in Data Pruning How to Train Data-Efficient LLMs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1040ee3-a77b-4e43-8c1c-ea8522bbea7f · inbound
Efficient Training of Deep Networks using Guided Spectral Data Selection: A Step Toward Learning What You Need How to Train Data-Efficient LLMs
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d867eea-4249-4856-8592-b3db67339c98 · inbound
Sub-Scaling Laws: On the Role of Data Density and Training Strategies in LLMs How to Train Data-Efficient LLMs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6c6eb5c-08c4-4a4b-a6c8-30527e315b02 · inbound
Beyond Traditional Algorithms: Leveraging LLMs for Accurate Cross-Border Entity Identification How to Train Data-Efficient LLMs
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8490cd5d-1c4f-4bab-95f1-f8eca72731b9 · inbound
Language Models Improve When Pretraining Data Matches Target Tasks How to Train Data-Efficient LLMs
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a86b5a3-61a3-4750-a132-2cb73815b823 · inbound
LAMDAS: LLM as an Implicit Classifier for Domain-specific Data Selection How to Train Data-Efficient LLMs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7860f5b1-4c1e-4780-86e1-89b7fa135eb3 · inbound
An Empirical Study on Influence-Based Pretraining Data Selection for Code Large Language Models How to Train Data-Efficient LLMs
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b912b53-402d-4d00-bb74-0e15c36c04ac · inbound
Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts How to Train Data-Efficient LLMs
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b4a269f-1cbb-4eaf-b0d9-3a16f4405fe2 · inbound
KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates How to Train Data-Efficient LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6601df6f-9c38-41f9-ab46-d3096b43ba4e · inbound
DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models How to Train Data-Efficient LLMs
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98535e94-9a8d-4b8a-8662-34780db3a8c3 · inbound
Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence How to Train Data-Efficient LLMs
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af26c0cc-ad3e-43fe-9683-2353c9ae1b0d · inbound
Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence How to Train Data-Efficient LLMs
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fe4cc44-6781-46ce-ab2f-60a76e197dfd · inbound
Accelerated Relax-and-Round for Concave Coverage Problems How to Train Data-Efficient LLMs
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad0f918b-ca1e-4b43-a0a9-df8466e9b472 · inbound
Reflections and New Directions for Human-Centered Large Language Models How to Train Data-Efficient LLMs
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10576518-8575-43d8-9ea9-7263c128c47b · inbound
Efficient Test-Time Finetuning of LLMs via Convex Reconstruction and Gradient Caching How to Train Data-Efficient LLMs
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 741108f7-c40c-483a-9b45-dd06febbe00e · inbound
Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them How to Train Data-Efficient LLMs
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.