Pith. sign in

Paper Citation Record · LEDGER

LESS: Selecting Influential Data for Targeted Instruction Tuning

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 80 inbound Pith citation observations for arXiv:2402.04333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04333 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 80 of 80 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:29:51.783032Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cfa146e5-5021-43a3-8ee1-770efa704979 · inbound

Retrieval-Augmented Generation for AI-Generated Content: A Survey cites this paper.

Retrieval-Augmented Generation for AI-Generated Content: A Survey LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 130

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:32:17.369856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T13:32:17.177021Z digest=sha256:581508e3ef1d61f557120b1a6ee3b8efd9ab06abbbc1c4c85ff058e234098287

Observation 7052df9d-d093-454c-86e8-712b6a15f902 · inbound

BPO: Towards Balanced Preference Optimization between Knowledge Breadth and Depth in Alignment cites this paper.

BPO: Towards Balanced Preference Optimization between Knowledge Breadth and Depth in Alignment LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T19:14:37.068426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:14:37.068426Z digest=sha256:e7e34a6cbcfcca2218b1207223d9f71097f85e5bbda82c5cc3aedd49759c7a08

Observation 91040c02-dd05-4c54-9924-57e20f97a625 · inbound

Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning cites this paper.

Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T15:58:03.299039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:58:03.299039Z digest=sha256:280ea8a718c53f390e51ad5754cbb684e24d75b2529dec21191161da31dbe54a

Observation 60badec1-eec3-4c83-841d-ecd57ffc2dbc · inbound

ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning cites this paper.

ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T05:13:41.578831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:13:41.578831Z digest=sha256:8ded68db4cc5a643657bcfa277001f98eefee2fc91396996db5ab2c2d8b081a8

Observation b9b02ac0-e4d5-459e-b5fa-e9caeb710396 · inbound

Mastering Collaborative Multi-modal Data Selection: A Focus on Informativeness, Uniqueness, and Representativeness cites this paper.

Mastering Collaborative Multi-modal Data Selection: A Focus on Informativeness, Uniqueness, and Representativeness LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T19:54:52.058857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:54:52.058857Z digest=sha256:7c47527b1675057104373af5f96f2e57747a11fb7e91380ad7aa59983bbe0bde

Observation 5a998296-4a9a-4b9f-9064-4db61e7354f1 · inbound

LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning cites this paper.

LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T14:02:43.932718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:02:43.932718Z digest=sha256:3dc8eea30added32e610ed33f78abe9e1405293e93b93b9b6ed97e0705f1f967

Observation 506b9688-31f3-41b4-acab-8cf59a1f6c4c · inbound

How to Synthesize Text Data without Model Collapse? cites this paper.

How to Synthesize Text Data without Model Collapse? LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:05:48.554007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:05:48.554007Z digest=sha256:b5ca784c0061626b32e5c64c72bf01b0f2a505a2ba88eccd61b5dfb2da6b2b81

Observation 6f5105b7-fb81-4494-bec3-1497ff23ea11 · inbound

ResoFilter: Fine-grained Synthetic Data Filtering for Large Language Models through Data-Parameter Resonance Analysis cites this paper.

ResoFilter: Fine-grained Synthetic Data Filtering for Large Language Models through Data-Parameter Resonance Analysis LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T11:59:28.468647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:59:28.468647Z digest=sha256:cc86f429d161fd2654ef6e522989ab28a10b22107fd8901a9f1f0e83715af11f

Observation ee33c58c-b071-4cec-9b85-655741650ca0 · inbound

RobustFT: Robust Supervised Fine-tuning for Large Language Models under Noisy Response cites this paper.

RobustFT: Robust Supervised Fine-tuning for Large Language Models under Noisy Response LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T11:51:02.962175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:51:02.962175Z digest=sha256:7f399eb7c2b070acb80a86468b7c33cd14f56a7a9de572a3bb08b63d6ec177a7

Observation c249a302-567d-49ea-a5f7-c1bb3654e406 · inbound

Error-driven Data-efficient Large Multimodal Model Tuning cites this paper.

Error-driven Data-efficient Large Multimodal Model Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T11:18:35.618509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:18:35.618509Z digest=sha256:911f89e1f6c11adda8f1c8e3a93c3552f61bbcf09e5d846dcb653bcfdff00f8e

Observation 84e65fbf-ae62-405c-87bc-d30ba4b33810 · inbound

Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights cites this paper.

Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:00.908041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:00.908041Z digest=sha256:20063b4420c6d4cd7149b135f6b7ac8a44a1850934d2d98cbf29be36a24526f7

Observation 026e0e28-e136-4351-8f29-181c598f4b92 · inbound

Foundations of Large Language Models cites this paper.

Foundations of Large Language Models LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 269

Resolution
unresolved
no resolver link, observed 2026-08-10T20:14:59.398548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:14:59.398548Z digest=sha256:edf217d7915984b9626cc5631e9299b15735b49b1c8e6cd68d95e4b2847515e9

Observation 233e79e0-c43a-4119-be56-18093fa30945 · inbound

Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities cites this paper.

Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T17:34:52.508784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:34:52.508784Z digest=sha256:b70ee6c28d9d9865b6e5168cc3e3446131c8a2ca9cc307a764fcb6c11c9f8402

Observation 843288a1-07cf-45c3-a358-a839a71c8e79 · inbound

R.I.P.: Better Models by Survival of the Fittest Prompts cites this paper.

R.I.P.: Better Models by Survival of the Fittest Prompts LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.441499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.441499Z digest=sha256:d1f49dc16a9fdab12fb3686b8e91e3c152b4e568ecef0a5f20df1135c493182e

Observation 26a3a64c-2735-40c2-a3ad-a760e24418ad · inbound

Ensembles of Low-Rank Expert Adapters cites this paper.

Ensembles of Low-Rank Expert Adapters LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-09T20:29:52.691742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:29:52.691742Z digest=sha256:04260fd7507c4586caf78b1fc55a7373f939c9f4c126c12babf6629795851136

Observation 45275c11-40ac-4f3e-8466-13e544982167 · inbound

DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks cites this paper.

DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.502624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-23T03:58:48.967122Z digest=sha256:76c176113d6044baaa1f33631cfdde2c1138d72f3139c2ee655b190acabe494d

Observation 11ccd4aa-1365-4999-b946-c001122c4f81 · inbound

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation cites this paper.

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T17:31:59.904624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:31:59.904624Z digest=sha256:142a2ef24d75d20154f452119a1a7bc3522066b51a2d50e9a769af9f64784345

Observation bc7838af-4d9d-4ff1-ba19-5ec589007993 · inbound

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model cites this paper.

Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 223

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T08:02:23.531239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-19T08:02:23.002090Z digest=sha256:5cd44a5f283983c6698d6017d4c447ff1b5ba618005dbca9e2825a8a2aacbf5d

Observation 1c5560d7-8c6b-472b-ad48-2e96db58f920 · inbound

Data-efficient LLM Fine-tuning for Code Generation cites this paper.

Data-efficient LLM Fine-tuning for Code Generation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:51.783032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:51.783032Z digest=sha256:fb65a9496b1e378cb5dff718a6ab89b1897444e0fffef2f7a5e59d257996884a

Observation 81f2f5b7-1049-4dce-b663-bcf934e9c5e3 · inbound

DONOD: Efficient and Generalizable Instruction Fine-Tuning for LLMs via Model-Intrinsic Dataset Pruning cites this paper.

DONOD: Efficient and Generalizable Instruction Fine-Tuning for LLMs via Model-Intrinsic Dataset Pruning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:10.993352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:10.993352Z digest=sha256:2a7480beeb39792aa5a11175f55d1cfd0d5915e476676362cec7cbda5a65fc97

Observation 432ae583-2d6a-4763-9dc0-3ca9eee05459 · inbound

R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training cites this paper.

R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T04:49:32.984118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:49:32.984118Z digest=sha256:e13afdc8d94f7ddbc7535370c7503c760a7059c61fc96a7f5fdeacb7a15d2c72

Observation c6de88d6-fa5e-4c34-9016-6309d0cb2f41 · inbound

Adversarial Cooperative Rationalization: The Risk of Spurious Correlations in Even Clean Datasets cites this paper.

Adversarial Cooperative Rationalization: The Risk of Spurious Correlations in Even Clean Datasets LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T01:09:03.312234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T01:09:03.312234Z digest=sha256:cbee9065f00aa7604a274046646d0cb818c3a3576f3be281eede4b110fb99d78

Observation 2fca783e-3bdc-445d-9915-86eaa1797c73 · inbound

RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection cites this paper.

RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T23:12:39.076022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:12:39.076022Z digest=sha256:b5934bb2250d40fdcce64f7d14b1da747ca113034cc8332d6e20804d137ce653

Observation 987fb47c-4b71-49fe-a94a-b6cade2d47d0 · inbound

UFO-RL: Uncertainty-Focused Optimization for Efficient Reinforcement Learning Data Selection cites this paper.

UFO-RL: Uncertainty-Focused Optimization for Efficient Reinforcement Learning Data Selection LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:37:48.240990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:37:48.240990Z digest=sha256:5cb2b72233b9d3a70022a06de5c6d82e9dbc1ff111748394cdd13eead0ec32c8

Observation 58a20d5e-0717-4bd0-adc8-4543c0628660 · inbound

IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment cites this paper.

IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T20:32:06.112224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:32:06.112224Z digest=sha256:b523166a193c5c07f49bd00dbc4c4949a6e7f5306db8fe7057a7c1951e7a5ea1

Observation 44bd3a5a-7149-4c66-9ae8-22063bb3aeab · inbound

Merge to Mix: Mixing Datasets via Model Merging cites this paper.

Merge to Mix: Mixing Datasets via Model Merging LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:38.203646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:38.203646Z digest=sha256:dee2bf7eb23d5954b840c0c4783835efb7148839ae9397cea231b5407355c90e

Observation 6c0c0df7-4c2c-461c-8c2f-c99bf41c04c8 · inbound

Generalizing Large Language Model Usability Across Resource-Constrained cites this paper.

Generalizing Large Language Model Usability Across Resource-Constrained LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 159

Resolution
unresolved
no resolver link, observed 2026-08-15T22:08:55.882570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:08:55.882570Z digest=sha256:991804d553a205e23dc940f3ec6b4028cbca611eb856754492621521da055dab

Observation 8fbc55e1-4d0d-4961-b324-4f89de4aa9fe · inbound

Efficient Data Selection at Scale via Influence Distillation cites this paper.

Efficient Data Selection at Scale via Influence Distillation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:17.393085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:25:17.393085Z digest=sha256:c666fd46ca8f60f049cfe8a72b96167529608d13f58867758b5c195b4cd1e96c

Observation c9cb0f3c-ed17-41e1-a95f-d38187aef1da · inbound

Daunce: Data Attribution through Uncertainty Estimation cites this paper.

Daunce: Data Attribution through Uncertainty Estimation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:56:13.859762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:56:13.859762Z digest=sha256:077434ca6e374354ff53fd5013d7de1bc6b7a89b21f650f7879f7a32db854cdf

Observation 3208fa00-3ea2-4407-a70c-a64010242d19 · inbound

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs cites this paper.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.108486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.108486Z digest=sha256:b991cb8362d64092ea15ed08b90126b1a283fb98addb6c4cf0aa4f0d55b64ea3

Observation e0929f37-826d-4ac9-9b84-95733f5edd1e · inbound

Data Pruning by Information Maximization cites this paper.

Data Pruning by Information Maximization LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:45:03.259826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:45:03.259826Z digest=sha256:926bd7b3c77e1cc0ed6f8e9fdfc7e8897fdfe39b2d53456a8837ab602e2471b8

Observation 1fd56d91-44f8-4727-9c8f-17c273d6c01f · inbound

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis cites this paper.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.274238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.274238Z digest=sha256:a02b52e1855e90bfcd03805ead4ae46adb3f76c7e23fac51737009b9002b5455

Observation c2e64d93-be90-45dc-b914-c03dbf45b9ce · inbound

EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation cites this paper.

EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:49.728829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:52:49.728829Z digest=sha256:f2864ba033623d2c2f75c486a9cba609bf13b8b1991c4f93a2476fc334131fc8

Observation 875d1ffe-848a-4f5a-ae4b-74fde5b61074 · inbound

Towards Efficient and Effective Alignment of Large Language Models cites this paper.

Towards Efficient and Effective Alignment of Large Language Models LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 191

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:43.112708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:43.112708Z digest=sha256:70315e0277f58213db13b003379c2878a94e6efaf3a29c333718bb5e0fa118bd

Observation 86930eeb-2581-4feb-8996-c825a2b0a21e · inbound

Approximating Language Model Training Data from Weights cites this paper.

Approximating Language Model Training Data from Weights LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:03.711431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:59:03.711431Z digest=sha256:fa9217ae56dff738b4d8bfab2b16f688fe4cc04f5d5f7588a21a27f45aa17642

Observation 6089f827-694c-45b9-b258-ef70f3d67c75 · inbound

CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning cites this paper.

CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T18:39:18.992970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:39:18.992970Z digest=sha256:6bf7aae131bde03b39ee069cfdeec14d2bea5a066f6f8f95e2eaf265acc79f33

Observation fd07e8cb-16c9-47e2-aeb9-a5b319f6dc64 · inbound

Data Diversification Methods In Alignment Enhance Math Performance In LLMs cites this paper.

Data Diversification Methods In Alignment Enhance Math Performance In LLMs LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:28.792871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:43:28.792871Z digest=sha256:714e82313190f5ea61a12a0528b07dcdc024426c0479713f6a93ad5fe441240c

Observation a4b53896-a980-4e21-bb80-20510d0c1b40 · inbound

Attributing Data for Sharpness-Aware Minimization cites this paper.

Attributing Data for Sharpness-Aware Minimization LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:02:10.809020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:02:10.809020Z digest=sha256:38d8706c0e527d07458f4a31fc653a2b3db5919398bbd6044eef7515d0ecc031

Observation bea76b64-afea-4e70-826e-b5e3baf675e8 · inbound

Class-Proportional Coreset Selection for Difficulty-Separable Data cites this paper.

Class-Proportional Coreset Selection for Difficulty-Separable Data LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:25:16.024557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:25:16.024557Z digest=sha256:b9d140ccc59b439f8f7da5d2be3f576a34aa49b44932ca2b20995bef8dc2b8f6

Observation 39380165-d2f1-4710-8921-760b50e32e49 · inbound

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap cites this paper.

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.542379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T23:46:24.208438Z digest=sha256:d243c1fd32dab99eca4286d66bcbfa0ec75484bf29672b77d3b7962e7abda2f1

Observation e2651289-819d-4323-8f74-b7faa459aef5 · inbound

ADMIRE-BayesOpt: Accelerated Data MIxture RE-weighting for Language Models with Bayesian Optimization cites this paper.

ADMIRE-BayesOpt: Accelerated Data MIxture RE-weighting for Language Models with Bayesian Optimization LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:00:05.206082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:00:05.206082Z digest=sha256:bddd2932081408e91ee90cc13421ebdf805cc8bb6cdcad6aa56f3831c5b0d3a2

Observation 17e8a213-d68d-4a21-b5bd-71b06b1fe91a · inbound

Vision-G1: Towards General Vision Language Reasoning with Multi-Domain Data Curation cites this paper.

Vision-G1: Towards General Vision Language Reasoning with Multi-Domain Data Curation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T17:24:08.165737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:24:08.165737Z digest=sha256:7afd084d6b6b8ac30edd8af6889ce12a85d8a93f893b4810087baf7564925eb3

Observation b10d7fa2-8da9-4dd2-bb51-0fc5e4b821df · inbound

Understanding Data Influence with Differential Approximation cites this paper.

Understanding Data Influence with Differential Approximation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T18:28:49.220498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:28:49.220498Z digest=sha256:1d0639b983515e80cfc27d27174e4a50c22440cbcb772911d6b8cdbc059461dd

Observation a7660945-0e4b-49f6-b351-f168c06c3188 · inbound

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models cites this paper.

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:41:52.993698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-18T22:41:26.047957Z digest=sha256:2c54457b643292b8db3776a6e67c0b0da61b3f18db95eede9263c4dded6aa003

Observation d9755a9f-9825-455a-8593-269958c0d50e · inbound

Influence-driven Curriculum Learning for Pre-training on Limited Data cites this paper.

Influence-driven Curriculum Learning for Pre-training on Limited Data LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T17:56:37.005434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:56:37.005434Z digest=sha256:4982a82481f863d06da8f2316344540087c7faabeddde5b6bcea5902fb18b4a9

Observation 552ae6ce-9406-47c8-b42a-fade4dc9b29b · inbound

Active Domain Knowledge Acquisition with 100-Dollar Budget: Enhancing LLMs via Cost-Efficient, Expert-Involved Interaction in Sensitive Domains cites this paper.

Active Domain Knowledge Acquisition with 100-Dollar Budget: Enhancing LLMs via Cost-Efficient, Expert-Involved Interaction in Sensitive Domains LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T17:02:40.537745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:02:40.537745Z digest=sha256:5425ab7eeb83fc39344cee6f4f912e1932568cb51788ab2352a82f2a29a9cff9

Observation a3b348a0-105a-462f-b5a8-f6b9c5595c07 · inbound

BLISS: A Lightweight Bilevel Influence Scoring Method for Data Selection in Language Model Pretraining cites this paper.

BLISS: A Lightweight Bilevel Influence Scoring Method for Data Selection in Language Model Pretraining LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T11:16:14.978521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:16:14.978521Z digest=sha256:a2af8eb7e8fceb9a596bc935d5f3defab0d8a1190004cd429007edeb49244a0b

Observation b3e846b1-8718-4172-9390-914fe4cd7430 · inbound

LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning cites this paper.

LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:57:46.857879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T10:54:22.183741Z digest=sha256:265bb067ba5b85747e6e750f1e36f644672d889f906993aa6efda6cb914ef096

Observation fe16efd4-61ae-429c-9b36-2720b842818f · inbound

GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning cites this paper.

GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T21:05:26.433575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:05:26.433575Z digest=sha256:51d55b1b21efd62df3e502dadb63c6733f55caac8bcf68faaf546eb106157c71

Observation 87d181e7-8f54-4883-974c-dda1de54edeb · inbound

An Empirical Study on Influence-Based Pretraining Data Selection for Code Large Language Models cites this paper.

An Empirical Study on Influence-Based Pretraining Data Selection for Code Large Language Models LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:16:04.614363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:13:24.750244Z digest=sha256:a1acce05d0f012de5335b93716fa2cc741f42733db2a680f36b50144dd1415a9

Observation 685f7df8-22e0-422a-88e0-8836ddee743c · inbound

Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts cites this paper.

Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:15:58.957553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T17:42:31.465077Z digest=sha256:88a50247d7169033b04463c58a3b967f5c1902415744d23daa174c1ac938097c

Observation ccf6d8b0-fb66-49a8-b0c6-0057937bdbc6 · inbound

Selective Contrastive Learning For Gloss Free Sign Language Translation cites this paper.

Selective Contrastive Learning For Gloss Free Sign Language Translation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:31:10.152768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T11:42:20.114297Z digest=sha256:1eccbc12c7bf24381ab97c5dc634b1bcd1a7e305e6ecf19e960298fdd8fba69f

Observation 6c9d08b1-8c79-44eb-b94f-c6155e29257f · inbound

Rigorous Interpretation Is a Form of Evaluation cites this paper.

Rigorous Interpretation Is a Form of Evaluation LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:31:13.532088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-08T15:37:53.477706Z digest=sha256:eea3e1dc09ca904cb6d8577f38f03304646fbcbcd101be6ceb4b5a62dce66969

Observation b0c16eca-6102-4996-a681-820d1221266b · inbound

Let the Target Select for Itself: Data Selection via Target-Aligned Paths cites this paper.

Let the Target Select for Itself: Data Selection via Target-Aligned Paths LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:51:17.987635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T02:47:55.649231Z digest=sha256:49df69199b6bb099d8115760a6ea04c7dcc8da9d790fa794e16b80150e93b8c1

Observation c5d4e810-3b1f-4dd4-aefc-94145e55996d · inbound

Toward Communication-Efficient Space Data Centers: Bottlenecks, Architectures, and New Paradigms cites this paper.

Toward Communication-Efficient Space Data Centers: Bottlenecks, Architectures, and New Paradigms LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:52:52.481222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T19:50:51.545684Z digest=sha256:abd6b6b95b8b7a78615082723ee0f6eb53d693990fdcb863e251df473cea709d

Observation c15f55db-b6e3-465c-a7e3-d176f638882d · inbound

Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning cites this paper.

Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:07:54.541665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-14T20:02:58.318276Z digest=sha256:d8c28eb31517d5ebc29918a4b8565e2d83f9ed1bc024f9921e38cefbf3271b64

Observation 38050bd9-b388-4dfa-a9ab-b90a733208bb · inbound

PRISM: Preference-Aware Influence Function Based Data Selection Method for Efficient Fine-Tuning cites this paper.

PRISM: Preference-Aware Influence Function Based Data Selection Method for Efficient Fine-Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:04:56.612576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T17:01:21.521025Z digest=sha256:685cd86222a1e814f644f6748bde405599bcfe9969cda3fadd1f7e0a256a4d3e

Observation fffa1b45-079a-4bab-a7b5-6f7eec05379c · inbound

Unified Data Selection for LLM Reasoning cites this paper.

Unified Data Selection for LLM Reasoning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T05:34:40.117198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-22T05:33:20.930156Z digest=sha256:771f7b3bec732f5be4d0806390ceeed26a2d01cbdda741725277fc8cd99d3238

Observation 22045372-c9d9-46be-8336-63469e73ee5b · inbound

SLAP: Stratified Loss-based Pruning for On-Policy Data-Efficient Instruction Tuning cites this paper.

SLAP: Stratified Loss-based Pruning for On-Policy Data-Efficient Instruction Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:05:05.687911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T22:02:30.217607Z digest=sha256:75b46c9c8604f275b4ceb57d4f61390286d91f72c9754bfce10c730293681fbd

Observation 81ebf038-a562-469c-92a9-ba547a2ec863 · inbound

Single-Rollout Hidden-State Dynamics for Training-Free RLVR Data Selection cites this paper.

Single-Rollout Hidden-State Dynamics for Training-Free RLVR Data Selection LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:13:30.072550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T14:08:40.968105Z digest=sha256:95698fd79b6cb522763ad9e0a61cbd3242c96ca096ac023063e51ed09f9d301c

Observation b05c5d5a-316f-4ef5-af46-570b1d7322c6 · inbound

CODEBLOCK: Learning to Supervise Code at the Right Granularity cites this paper.

CODEBLOCK: Learning to Supervise Code at the Right Granularity LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T08:17:45.517082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T10:52:00.147391Z digest=sha256:b7a5d2865191b34069dd1b03787a19388cfe979f37638d6818d3cdaa9ab37a3d

Observation 72067d73-c7bb-4b23-a19e-e6d195fa543c · inbound

DRIFT: Refining Instruction Data via On-Policy Data Attribution cites this paper.

DRIFT: Refining Instruction Data via On-Policy Data Attribution LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T19:18:54.743967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T01:57:05.784589Z digest=sha256:109f6f354c067dd6afb083b73351dac0289e5880a4cd25f68de568619e28474d

Observation 42197a11-2828-4ea1-836a-5fc49503a4cc · inbound

Predicting Mergeability of Parameter-Efficient Fine-Tuning Updates cites this paper.

Predicting Mergeability of Parameter-Efficient Fine-Tuning Updates LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T00:49:17.660072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T21:01:04.043286Z digest=sha256:e759295002a13a3b3ca91e00314324d9c844d620441312b1b80ff4ddd406da43

Observation 1b4ce2c7-c125-45c1-b9ff-5a310395af18 · inbound

Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior cites this paper.

Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T08:49:14.974070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T08:45:34.884703Z digest=sha256:e75b74c7916090b57efe7a84e207d6e6fdd16482cca2f212cb1bdef03fd67a53

Observation 5d0dc1af-7531-4af3-bd5a-db33d86d8c9a · inbound

Data Selection Through Iterative Self-Filtering for Vision-Language Settings cites this paper.

Data Selection Through Iterative Self-Filtering for Vision-Language Settings LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 197

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:49:44.950900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T09:22:47.537137Z digest=sha256:1b8be44010d879f13cdc97398473f2996573d52794fc2143a0befe8e2df4049e

Observation f5187601-fab7-4ef1-8fca-a6d3cd15ade9 · inbound

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR cites this paper.

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:30:01.570899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-25T22:47:09.330723Z digest=sha256:a75e6c8b4d3856c977f55808f7b74767b9b94e1bc2a5258e497ae6d68cc34c73

Observation 9ba4dd02-1047-4f55-a843-acde45ce4f63 · inbound

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR cites this paper.

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.437339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T09:55:19.804589Z digest=sha256:8a10197effd91d17c8aee63f171df45a0f81308e165a1c088e253f5066764752

Observation 45c43e9a-76f2-40d8-a7e7-3e39de747679 · inbound

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity cites this paper.

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 218

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T20:50:12.684409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-25T19:23:56.452083Z digest=sha256:99941a30cbbd5e8ffc7f9d9951e36e2865d3915a1fde71d89159bfaeb8817fc5

Observation babf0a09-e7f6-4dea-828d-665ffdb55bc4 · inbound

When Does Generating More Help? Disentangling Fixed-Source Synthesis from Source Expansion in Synthetic Data Scaling cites this paper.

When Does Generating More Help? Disentangling Fixed-Source Synthesis from Source Expansion in Synthetic Data Scaling LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T15:28:33.802412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-03T15:20:51.474398Z digest=sha256:cf52d64cd3fde8f908948eb06bc309dec6c247d2c42c8cb7a276a6dad6bd3758

Observation 6f1c2684-6d63-4df4-a498-b5b16a5118cb · inbound

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures cites this paper.

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:48:39.404908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-03T16:44:41.720388Z digest=sha256:48f02b9b7d5fdef8ee6af202c6adbf287b79e70aa5054b12ff645e8d276eb06e

Observation 65992ae1-e984-42f8-b12b-e0c037d7dcc9 · inbound

DataShield: Uncovering Risky Fine-Tuning Data Across LLMs Through Consensus Subspace Alignment cites this paper.

DataShield: Uncovering Risky Fine-Tuning Data Across LLMs Through Consensus Subspace Alignment LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T00:17:36.191807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:17:36.191807Z digest=sha256:3216e8fe16bf41206416bc9caa57bc3dce9d8a1aaed178bef83d0608eb1d1aae

Observation d3d50ac5-3475-4a71-a7fc-109dec7d815d · inbound

DataPrep-Bench: Benchmarking LLMs as Training Data Preparators cites this paper.

DataPrep-Bench: Benchmarking LLMs as Training Data Preparators LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T13:43:29.432403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:43:29.432403Z digest=sha256:ba66c5b30eb89a91a7e8896cfb30df8db4ab948c1e9091ac87ecbd9272ddb101

Observation c2c8288b-7575-47fa-8bed-2d1642d7603f · inbound

DomainPilot: Domain-Level Loss-Guided Two-Stage Data Mixture Optimization for Efficient Language Model Fine-Tuning cites this paper.

DomainPilot: Domain-Level Loss-Guided Two-Stage Data Mixture Optimization for Efficient Language Model Fine-Tuning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T06:19:33.174874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:19:33.174874Z digest=sha256:5aa2f5e1e39ce32e1ac72cc1afe28a9962ad4b9103dc6c092eecb422882a56f2

Observation 35f2a815-325f-4514-b402-e323b95f7b8a · inbound

Less Data, Better Alignment: Data-Centric Multi-Evaluator Agreement for Preference Optimization cites this paper.

Less Data, Better Alignment: Data-Centric Multi-Evaluator Agreement for Preference Optimization LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T00:30:21.553602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T00:30:21.553602Z digest=sha256:8158a7a90aee42906ec456cc03942e7b1621c8704c2598cc8d08d844e5865842

Observation be57078a-a0bc-4538-a878-545a405a753d · inbound

Bridging Compute- and Data-Optimal Pretraining cites this paper.

Bridging Compute- and Data-Optimal Pretraining LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 180

Resolution
unresolved
no resolver link, observed 2026-08-01T03:02:08.841217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:02:08.841217Z digest=sha256:17a2ebadd490e271dad578b405f135a3590560e172fd8222973b63429e047141

Observation 02092546-02ad-403a-b3fe-a213b719dbbe · inbound

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement cites this paper.

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.235536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T01:15:07.235536Z digest=sha256:55093564e1ec30039ad93a8a24024f5af09c42dc051d4a2ab29e29bfe0c48538

Observation ae92028b-9090-43f1-b779-6ca9cca80a4b · inbound

SDO: Structure-Aware Data Organization for Efficient LLM Post-Training cites this paper.

SDO: Structure-Aware Data Organization for Efficient LLM Post-Training LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T10:52:18.942456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T10:52:18.942456Z digest=sha256:b8d7281753bf3cdd7b8021e9fe59a1e1042a229e07c981bd78a27d8272620cda

Observation fc11e768-8335-42cf-b92d-6c6c9617c245 · inbound

SafeBuild-Bench: A Temporal-Robust Construction Safety Benchmark with Graph-Enhanced Data Mining cites this paper.

SafeBuild-Bench: A Temporal-Robust Construction Safety Benchmark with Graph-Enhanced Data Mining LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T01:29:18.422598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:29:18.422598Z digest=sha256:024816f6a096c946584106d9da0cf7dc22a57df090c2ac29af4416aaf7e8ca7f

Observation 8bffff43-b5e4-40cb-996b-9ea6ae5014aa · inbound

CODS: Iterative Bellman-Residual Data Selection for Reusable Offline Reinforcement Learning cites this paper.

CODS: Iterative Bellman-Residual Data Selection for Reusable Offline Reinforcement Learning LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T00:27:49.959430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:27:49.959430Z digest=sha256:8430dbc546d6b3fac89f08b305bc307515264f17ab2216bcdc14f9104ec38446

Observation 0e714bce-1beb-405e-8ec6-dd1145f17663 · inbound

LittleLearner: Language Models Under Pedagogically Controlled Knowledge Exposure cites this paper.

LittleLearner: Language Models Under Pedagogically Controlled Knowledge Exposure LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:12.534998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:12.534998Z digest=sha256:1589b265c3cb58b67763b506d3eddf2704c0d87b55221bababbd5f4b2e30639a