Pith. sign in

Paper Citation Record · LEDGER

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

As of 20 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2607.07504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07504 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-09T09:01:10.366890Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T17:35:06.786385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact11
  • verified fuzzy1
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1ef08d5-8e60-440e-8a12-9177e2374d08 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.401920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:c9cb69055858ab03783d32991405b2880a3b605948647c4a30b4d50ed9d030bd

Observation 09d0a9de-c414-4766-aba1-75a1a4f54bf9 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.404650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:226c09f363c07afa1311e8b6c6b2b874c97da6b436ea6514d08bdf60734d7ebb

Observation 93953701-11d7-4430-96e5-7d9e96237f6b · outbound

This paper cites SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.130866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:dde0812f07471efcd6c11af614b546e1cea272c13e09ae5ae679c2b909feb3d9

Observation f22a2e72-0f09-4d72-b003-a7c628239b6f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.408625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:91f603f5163909fb49485ba1d630a1c62043dcd6e4721d6e2c20d3f0cacf1d7e

Observation 7f27b98f-55ae-40d6-acc9-06c9a92c1442 · outbound

This paper cites DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.133830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:19c048f5554c1b75a303c26a9b530c648dfaa414cdab9b25159fc0e838586221

Observation 86898c1e-ad37-47cd-bbd7-4350a2efb67e · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:2fd31e65ce3d2275d97d8ab4df3d67fdf3ee8c397867ef44f06cb5e75434528a

Observation b0419213-ee9d-4b80-a42f-e4c338735113 · outbound

This paper cites Advances in Neural Information Processing Systems33 (2020), 9459–9474.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Advances in Neural Information Processing Systems33 (2020), 9459–9474

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T09:06:06.399724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:1c0df51d29c612ee24391d40ee8b3c6f4fc3fc70e918e8c2b63ae5fdc449cd56

Observation ee6846d7-14ac-4868-99d4-8a12d99aa43f · outbound

This paper cites AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.143880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:f41948a56b8728ea47aac650ca78bc1458d54cb679f85387a241ce5808723386

Observation d15ac8f1-18c3-4d69-bd29-4b83bf4449f8 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.117269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:cead6264fdce68b5177725b99810be2f2ae2cca4eca2a11e576db3b74fd8c592

Observation a79288f6-49c0-4932-bb08-4e4c562b752f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:9b2f0a1762667d0bf9b05613bb94220e0e2be1440caae98ff6a3df5894e4ee3c

Observation 1f671541-6cea-4542-a5ae-64519f0f1e16 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.141613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:a85ef32b76d9511c2f7d29d5ea32f45ac0b378dc9b53ddd172d90d0a1bc06982

Observation 05041e25-47f6-4e87-8b1a-1b26a33f5c0b · outbound

This paper cites Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.154704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:013e0b1eac2c38f7bee2961b876c7de06e126551175c416d8df5675c2e0e2ddf

Observation 0625b5d3-5f77-41e8-af3c-dfb249bf919f · outbound

This paper cites LLM4DS: Evaluating Large Language Models for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows LLM4DS: Evaluating Large Language Models for Data Science Code Generation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.145336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:3371be1f425538f22b142c49ffbef4ec415c445d4863099b08d16d773a0c4105

Observation 910858c5-2644-4638-8c81-5c2cbfa4ebd4 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-09T09:06:06.127569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:182710ae39250ff9671db66e39836aef1549ff905702746477aa6cc906b2b43f

Observation 3e0f0e24-9e40-4397-8d8b-3ee7a87b32df · outbound

This paper cites PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.149504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:a63abc3d50430ca2784e1ebb3dae7e2a5ad4bbde67adaaa6143a2d1e960b53bf

Observation 736d9a04-7a50-4ce0-8288-8d3905cf1aa5 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.396644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:df86cc356363278cde6771757639cb29d83ff1da7c998b598ffa6e0414f863cc

Observation fb8e1859-f952-49df-b828-53cf5e68d1ea · outbound

This paper cites DataSciBench: An LLM Agent Benchmark for Data Science.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DataSciBench: An LLM Agent Benchmark for Data Science

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.147665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:b5ca420e87f383f9d8357e7116340b0bf1127ceee0d5a633a2c3b0be9ff726f2

Observation c39d3934-296e-4a00-b58d-f937b4bb400f · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.151311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:3e57e818e16f5cfd6c52fbc58fd1831b218a502fd255fd76108de6b1c30a5ef1

Pith citing papers

Observation a02860c8-bd7b-44f9-bd72-d4b19e054b75 · inbound

Is Progressive Disclosure All You Need for Long-Context Agents? cites this paper.

Is Progressive Disclosure All You Need for Long-Context Agents? Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T17:35:06.786385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:35:06.786385Z digest=sha256:f0147ba9bcd49864c5c68cf43cb5e29aba3f6b7a08ceeb50cb5349b29672da24

Observation 5198b524-eec3-48cd-9bc4-6faa16cc494a · inbound

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents cites this paper.

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T04:31:35.698995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:31:35.698995Z digest=sha256:5106799c9507d1769f7f5f633cb6f2d461ab79e5ae0d24241d87414ae002b904