Pith. sign in

Paper Citation Record · LEDGER

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

As of 7 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2607.07504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07504 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-09T09:01:10.366890Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T17:35:06.786385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact11
  • verified fuzzy1
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1ef08d5-8e60-440e-8a12-9177e2374d08 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.401920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:e4023395c39606d113582bb7db81967b724615b07572f6592bddd9406d250f01

Observation 09d0a9de-c414-4766-aba1-75a1a4f54bf9 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.404650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:09f02c1bfb7d9d45f6c6928425bb24f370f9a7ced8f4aec900ce891e4f3edb6f

Observation 93953701-11d7-4430-96e5-7d9e96237f6b · outbound

This paper cites SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.130866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:c72d4e31e589f3988e5bb8d40ea01940b95d26e9245afba8d35aa942411abed2

Observation f22a2e72-0f09-4d72-b003-a7c628239b6f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.408625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:d4d05b9d9359110d0d64bda6c495421196a9055e507ec6415314d8cd11ebdd8f

Observation 7f27b98f-55ae-40d6-acc9-06c9a92c1442 · outbound

This paper cites DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.133830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:ae8ad4de90bb2942eb8490241fc0a0a7ed0d8982fb36972d7356abc3e22754a5

Observation 86898c1e-ad37-47cd-bbd7-4350a2efb67e · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:45fccce28c3530dcb20fedbf1dcb1309cf3c625505053815770435ab7d9682e9

Observation b0419213-ee9d-4b80-a42f-e4c338735113 · outbound

This paper cites Advances in Neural Information Processing Systems33 (2020), 9459–9474.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Advances in Neural Information Processing Systems33 (2020), 9459–9474

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T09:06:06.399724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:07acde543ea02f6bfa40e89242c3a0f952f51effa0e573c02cb8d18afd2c008f

Observation ee6846d7-14ac-4868-99d4-8a12d99aa43f · outbound

This paper cites AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.143880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:f8d5589d2a3416b34ee02e7d9decef299e3ef4c11c7ce2715ea6825377169318

Observation d15ac8f1-18c3-4d69-bd29-4b83bf4449f8 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.117269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:1d582177b78c22c6bf0c6abfb58727aaa52496182051f1ca8be881b9b207cd14

Observation a79288f6-49c0-4932-bb08-4e4c562b752f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:86c0d3a4b0050769dc0ed2f8f177d9f90190684ef0e9636ee3147c8b96e382c6

Observation 1f671541-6cea-4542-a5ae-64519f0f1e16 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.141613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:22d4d2a2fb263af1d6e149a2bf5ee71ef84a9c700fd4fabc7affc07604f93d7f

Observation 05041e25-47f6-4e87-8b1a-1b26a33f5c0b · outbound

This paper cites Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.154704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:577d4fa6fc3ca1731fb34570e7a8cfb3e5ddb43dbd055612a8b40cf1a35d8c79

Observation 0625b5d3-5f77-41e8-af3c-dfb249bf919f · outbound

This paper cites LLM4DS: Evaluating Large Language Models for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows LLM4DS: Evaluating Large Language Models for Data Science Code Generation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.145336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:250b0728b93edb2d7f01ae730fe5883b8d24a39f31a542f53b10dff2d98ccba9

Observation 910858c5-2644-4638-8c81-5c2cbfa4ebd4 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-09T09:06:06.127569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:e509af4668aa9a9b6b13912706724093ba9848baea823108028c527bfbe7d90b

Observation 3e0f0e24-9e40-4397-8d8b-3ee7a87b32df · outbound

This paper cites PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.149504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:76dab5e4b0694be487c20b41d3d873f4f98a504f60ce16c5c3dbbe3361d66016

Observation 736d9a04-7a50-4ce0-8288-8d3905cf1aa5 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.396644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:04192093d091f57e242e97fbf90eb513fc418805effde585b63bae1d2f015475

Observation fb8e1859-f952-49df-b828-53cf5e68d1ea · outbound

This paper cites DataSciBench: An LLM Agent Benchmark for Data Science.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DataSciBench: An LLM Agent Benchmark for Data Science

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.147665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:41262886d5fb45bb9fe3ca8e61b362b4b480d63fdfb94e304e605eb82314a9af

Observation c39d3934-296e-4a00-b58d-f937b4bb400f · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.151311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:8d6d5cea6ce582f26221f2abf469bb26e8ad87cf9830a88af03460e1288e49b5

Pith citing papers

Observation a02860c8-bd7b-44f9-bd72-d4b19e054b75 · inbound

Is Progressive Disclosure All You Need for Long-Context Agents? cites this paper.

Is Progressive Disclosure All You Need for Long-Context Agents? Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T17:35:06.786385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:35:06.786385Z digest=sha256:cea5de7c4ae000bb4067ae6ad0f37fc4dabe75a07c8dab1088fc00f7ae9a4fc2

Observation 5198b524-eec3-48cd-9bc4-6faa16cc494a · inbound

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents cites this paper.

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T04:31:35.698995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:31:35.698995Z digest=sha256:e4e808b8c4f746b4d1ff2ffa48cc99faeb44e122f52e60cab8543f9f0f24dab3