Pith. sign in

Paper Citation Record · LEDGER

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

As of 7 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2607.07504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07504 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-09T09:01:10.366890Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T17:35:06.786385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact11
  • verified fuzzy1
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1ef08d5-8e60-440e-8a12-9177e2374d08 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.401920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:dc6345d082831cdc3b4dc5463df7c114a390b95586ae3994671688fdb0ca5edf

Observation 09d0a9de-c414-4766-aba1-75a1a4f54bf9 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.404650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:91a724d1c987d9fe4bde792bf90c02bc732f6a627cc117410a1664dc37a684c4

Observation 93953701-11d7-4430-96e5-7d9e96237f6b · outbound

This paper cites SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.130866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:5b4f644d54d5f5ac3c700910ac31c71aba6dbf0a4f157027fbb6fa98180d4f22

Observation f22a2e72-0f09-4d72-b003-a7c628239b6f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.408625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:cd995c7d02bfb9b0c82d80df803db1a44eaaeec57f836c6abad5a0fb9dbdebfc

Observation 7f27b98f-55ae-40d6-acc9-06c9a92c1442 · outbound

This paper cites DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.133830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:e60d1c71e99bed25356970027928495acf976b7bd6eb25c533c14b150698e059

Observation 86898c1e-ad37-47cd-bbd7-4350a2efb67e · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:85670fe1729ee6f127279110b241c921af097c406b6875771a33fec41671847a

Observation b0419213-ee9d-4b80-a42f-e4c338735113 · outbound

This paper cites Advances in Neural Information Processing Systems33 (2020), 9459–9474.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Advances in Neural Information Processing Systems33 (2020), 9459–9474

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T09:06:06.399724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:b7e13320e96a71bd3a5e106fef676750e62a0bf40aa6e928de1f73d3f8d90226

Observation ee6846d7-14ac-4868-99d4-8a12d99aa43f · outbound

This paper cites AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.143880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:b493c40468805bbbc29454752ad3ec43d99bc96215f2fcff1c3b79b5ef215d6e

Observation d15ac8f1-18c3-4d69-bd29-4b83bf4449f8 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.117269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:1dfeb4dcaf22c28eed340f5823853615e650b2022d98b16050b77591228eff5e

Observation a79288f6-49c0-4932-bb08-4e4c562b752f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:d539a2582c1a9049ac6486a775594fad26bf0dd2192ea50378dd0448f5c8078f

Observation 1f671541-6cea-4542-a5ae-64519f0f1e16 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.141613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:bb59ac1ca8020f08b8fe30be4e7e7952bf3b391aa1a7b594f71b3d6a6e434bb3

Observation 05041e25-47f6-4e87-8b1a-1b26a33f5c0b · outbound

This paper cites Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.154704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:db7767f31bdce2b1ae95e010178c5565ba411ed8530eed5b9635c58513793cec

Observation 0625b5d3-5f77-41e8-af3c-dfb249bf919f · outbound

This paper cites LLM4DS: Evaluating Large Language Models for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows LLM4DS: Evaluating Large Language Models for Data Science Code Generation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.145336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:53132c417b66817fa7ff5506b3fe3ba0fa52ecbe36efd85c4aaf1c9895d8a5d8

Observation 910858c5-2644-4638-8c81-5c2cbfa4ebd4 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-09T09:06:06.127569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:a3745b7c7598554b0a60b72c8787832dfac3aa814b3f926bed2b7bd0d1e9f992

Observation 3e0f0e24-9e40-4397-8d8b-3ee7a87b32df · outbound

This paper cites PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.149504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:c10020a3a861055589bfac332199690999873fd5214717fe682c7172fcfb81d7

Observation 736d9a04-7a50-4ce0-8288-8d3905cf1aa5 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.396644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:174af6dc709e62b6cec9bace02ddf4aac11f4abd6c357eef77966f96f604c5d1

Observation fb8e1859-f952-49df-b828-53cf5e68d1ea · outbound

This paper cites DataSciBench: An LLM Agent Benchmark for Data Science.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DataSciBench: An LLM Agent Benchmark for Data Science

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.147665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:90d1fe924c4a0374ebd751c7201cd6c9f944d92617f9ae77dd30697e2c90147b

Observation c39d3934-296e-4a00-b58d-f937b4bb400f · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.151311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:54f0d3b37f1d80030b0897b261907d975205406247df8df77a308044de6ce742

Pith citing papers

Observation a02860c8-bd7b-44f9-bd72-d4b19e054b75 · inbound

Is Progressive Disclosure All You Need for Long-Context Agents? cites this paper.

Is Progressive Disclosure All You Need for Long-Context Agents? Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T17:35:06.786385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:35:06.786385Z digest=sha256:cea5de7c4ae000bb4067ae6ad0f37fc4dabe75a07c8dab1088fc00f7ae9a4fc2

Observation 5198b524-eec3-48cd-9bc4-6faa16cc494a · inbound

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents cites this paper.

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T04:31:35.698995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:31:35.698995Z digest=sha256:e4e808b8c4f746b4d1ff2ffa48cc99faeb44e122f52e60cab8543f9f0f24dab3