Pith. sign in

Paper Citation Record · LEDGER

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

As of 10 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2607.07504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07504 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-09T09:01:10.366890Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T17:35:06.786385Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact11
  • verified fuzzy1
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1ef08d5-8e60-440e-8a12-9177e2374d08 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.401920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:7ab5725a89aa90f25b4a93a888b95eb0e3fae514ef9bcab4ffa024655ee8180b

Observation 09d0a9de-c414-4766-aba1-75a1a4f54bf9 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.404650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:3f5183c7c269e008dc2c5907d675d02c1dbfe936559b81424f023f99bedc01ce

Observation 93953701-11d7-4430-96e5-7d9e96237f6b · outbound

This paper cites SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.130866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:a18e24ea8001e65348fa1941daff50abe486ae3de15839da48adbc4185b10bc7

Observation f22a2e72-0f09-4d72-b003-a7c628239b6f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.408625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:22547b622c4a22a14039d988743c72433afc0282a89dd0b2c6ab6882fca75b00

Observation 7f27b98f-55ae-40d6-acc9-06c9a92c1442 · outbound

This paper cites DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.133830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:fdaf7eac55345732aa6c16253c32b19e1b12cdbb793a605c0fc872cc063e96e7

Observation 86898c1e-ad37-47cd-bbd7-4350a2efb67e · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:9a60e382aa213850daea456e213dca675ef986b03138e096d9bcf23416eb7258

Observation b0419213-ee9d-4b80-a42f-e4c338735113 · outbound

This paper cites Advances in Neural Information Processing Systems33 (2020), 9459–9474.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Advances in Neural Information Processing Systems33 (2020), 9459–9474

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T09:06:06.399724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:0fef7ebb7f77f62a6b75476a7b41200a2996d2801901dad3e68ee8228a649c50

Observation ee6846d7-14ac-4868-99d4-8a12d99aa43f · outbound

This paper cites AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows AutoDCWorkflow: LLM-based Data Cleaning Workflow Auto-Generation and Benchmark

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.143880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:e5e7a1f6b507e832cd0a44c576688c686653efecb5910f80504a66b07fbfe65f

Observation d15ac8f1-18c3-4d69-bd29-4b83bf4449f8 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.117269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:fd6a2f29051cfb7ee2060b27f349b8e07a3bb1887ee1a64a0a1a536a45e7eaa9

Observation a79288f6-49c0-4932-bb08-4e4c562b752f · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.406788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:c460130218820601e9e92ad2807720a6e20d9cad2e1e18177e76d532c6cc96ce

Observation 1f671541-6cea-4542-a5ae-64519f0f1e16 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.141613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:803e0256a332d7f0b5f8b782760b2f7ff38c134067f262c8f4ed5ab3a258989f

Observation 05041e25-47f6-4e87-8b1a-1b26a33f5c0b · outbound

This paper cites Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.154704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:c6df9094fa740696985950f89852b084f1e9be15483b94645edf43c558e4c0df

Observation 0625b5d3-5f77-41e8-af3c-dfb249bf919f · outbound

This paper cites LLM4DS: Evaluating Large Language Models for Data Science Code Generation.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows LLM4DS: Evaluating Large Language Models for Data Science Code Generation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.145336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:ca97d7b9c293cdf1698ecea56d30a1b838fca71aae0ea0e193189956c3344dbb

Observation 910858c5-2644-4638-8c81-5c2cbfa4ebd4 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-09T09:06:06.127569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:ad5bf5a24a4cbe2513a84a95747e5dd453a1178d6fb7014fd6056e1004332bba

Observation 3e0f0e24-9e40-4397-8d8b-3ee7a87b32df · outbound

This paper cites PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.149504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:ebdde82b3fbd9ac8e4e3f7992acab119f1a80abe3cffe3ce0a17acd09bd4d6d4

Observation 736d9a04-7a50-4ce0-8288-8d3905cf1aa5 · outbound

This paper cites an unresolved cited work.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-09T09:06:06.396644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:c5e906a635b39ba40294209053837235dc9ca5b206610178ab4d20ad60e306e1

Observation fb8e1859-f952-49df-b828-53cf5e68d1ea · outbound

This paper cites DataSciBench: An LLM Agent Benchmark for Data Science.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows DataSciBench: An LLM Agent Benchmark for Data Science

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.147665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:dad142b3e1d8d0d10e122fbc3ff294fdea1f17d06e3f8e4af68a30618ae99ebd

Observation c39d3934-296e-4a00-b58d-f937b4bb400f · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:06:06.151311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T09:01:10.366890Z digest=sha256:d6210b88e5459a35e9e77fc031fd3532e7551321590a09abe7279adaa86aae97

Pith citing papers

Observation a02860c8-bd7b-44f9-bd72-d4b19e054b75 · inbound

Is Progressive Disclosure All You Need for Long-Context Agents? cites this paper.

Is Progressive Disclosure All You Need for Long-Context Agents? Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T17:35:06.786385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:35:06.786385Z digest=sha256:9d910c5109004bab2b012eb9184bff758b24d688e2868e745220094d158f18cb

Observation 5198b524-eec3-48cd-9bc4-6faa16cc494a · inbound

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents cites this paper.

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T04:31:35.698995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:31:35.698995Z digest=sha256:3c6f4531f7a88dc46d33af149154c3ff6bebe8ffba34999e6954ddd051706717