Pith. sign in

Paper Citation Record · LEDGER

SciCode: A Research Coding Benchmark Curated by Scientists

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2407.13168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.13168 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:33:13.132899Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 357ddd75-67dc-4f11-8700-e9dae2fa006f · inbound

RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts cites this paper.

RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts SciCode: A Research Coding Benchmark Curated by Scientists

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T14:33:13.132899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:33:13.132899Z digest=sha256:e4ec4f22370e52d9bac1d2d0fe548a39201aa32cf8aaad62b881aba2e115d2ee

Observation 26ea0b99-71b6-4221-9c4f-e5a8a3ff6b75 · inbound

The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning cites this paper.

The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning SciCode: A Research Coding Benchmark Curated by Scientists

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:58:33.516755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T15:58:33.219451Z digest=sha256:4f62a2c323b2d634d1127a3db9bedb599f7e98398d136672dac30e52ede57710

Observation f9ec80f5-56a6-423b-b50f-8a1c5c61ebc0 · inbound

A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications cites this paper.

A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications SciCode: A Research Coding Benchmark Curated by Scientists

Reference 268

Resolution
unresolved
no resolver link, observed 2026-08-07T00:48:36.931506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:48:36.931506Z digest=sha256:8d3a68259116ffacaeba302687597f9983947ceb0e4548fc872a929fedf5b776

Observation 800280f9-3e7f-4189-bc35-f3c1d4f2e981 · inbound

Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers cites this paper.

Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers SciCode: A Research Coding Benchmark Curated by Scientists

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:28:58.706378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:28:58.706378Z digest=sha256:99cca7b423a7711a8dd30aff7ea8ddf88b4862725540399d260859f15f224fac

Observation ae47f8bd-341b-4e9d-b23f-311c2f666365 · inbound

ChatVis: Large Language Model Agent for Generating Scientific Visualizations cites this paper.

ChatVis: Large Language Model Agent for Generating Scientific Visualizations SciCode: A Research Coding Benchmark Curated by Scientists

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T11:07:33.400361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:07:33.400361Z digest=sha256:31c35dd42c356a932b50ea12b00254c4f8c7aba3bc94cdd875e602098d738aa0

Observation e295e3be-ce7d-4a97-9302-ae35b64346f3 · inbound

How Far Are AI Scientists from Changing the World? cites this paper.

How Far Are AI Scientists from Changing the World? SciCode: A Research Coding Benchmark Curated by Scientists

Reference 163

Resolution
unresolved
no resolver link, observed 2026-08-06T10:55:15.053795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:55:15.053795Z digest=sha256:2ef7409e85129abcce7a0a9a019bd4c8d2f226a427f5bb47f28933681e8177d9

Observation 8e7113de-0d70-49f1-aeda-073e81f86333 · inbound

Symmetry-induced magnetism in fullerene monolayers cites this paper.

Symmetry-induced magnetism in fullerene monolayers SciCode: A Research Coding Benchmark Curated by Scientists

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T16:35:14.166427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:35:14.166427Z digest=sha256:9d699c1db6fc79820d791e8240c0f359a0bb60d504f1090c960aa36f4d0a5028

Observation c6c3e951-9187-446f-8974-1a3d1d2f5592 · inbound

ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation cites this paper.

ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation SciCode: A Research Coding Benchmark Curated by Scientists

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T14:52:49.846361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:52:49.846361Z digest=sha256:a2c591c395a88eae7089fbb363820fbd6ece6c508d50cc21592dcd003e404280

Observation 429efa30-d1ac-4a38-a32f-e2d181d4a552 · inbound

Agentic Exploration of Physics Models cites this paper.

Agentic Exploration of Physics Models SciCode: A Research Coding Benchmark Curated by Scientists

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:31.686179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:31.686179Z digest=sha256:3cca78f04671ad29df598a4d831a16c5e788cfd6216bf68606774e7215cb18de

Observation 6eb69df1-2e08-4527-a41a-0de411762877 · inbound

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding cites this paper.

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding SciCode: A Research Coding Benchmark Curated by Scientists

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:31:06.956134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T08:28:59.161675Z digest=sha256:8c46745d3cb4a0a97ccfcbae7f8f021fb675b78118a5eb95f920d2049297ccca

Observation 7f7ee4fc-074c-4d6b-ab84-f9ccd8cd171e · inbound

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations cites this paper.

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations SciCode: A Research Coding Benchmark Curated by Scientists

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T14:45:42.228598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:45:42.228598Z digest=sha256:ff69267c204dc79d8825069cc2b982c827be476fd35dc8c0ff6ee9c0950a789f

Observation ed0b687e-cc6b-49be-92af-aff48f2bc8f9 · inbound

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments cites this paper.

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments SciCode: A Research Coding Benchmark Curated by Scientists

Reference 178

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:23:27.277193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T01:20:03.181903Z digest=sha256:38344715cfa280f6cf3d248fc5bcfa3f1c86836cf31dd6ea68b809df2bef04e4

Observation 6f049bba-1a2d-402a-a552-2bf68bdd4445 · inbound

Towards Verifiable and Self-Correcting AI Physicists for Quantum Many-Body Simulations cites this paper.

Towards Verifiable and Self-Correcting AI Physicists for Quantum Many-Body Simulations SciCode: A Research Coding Benchmark Curated by Scientists

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:38:22.712994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T22:33:21.677835Z digest=sha256:f305d8df7ed3e7aea962ef19d3904a9e8ce65d345f5720fb0e8c161cd2b3a711

Observation 9a3caa96-7fb9-48ad-9eaa-3db28ee5371c · inbound

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows cites this paper.

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows SciCode: A Research Coding Benchmark Curated by Scientists

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:09.902343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T08:12:07.708441Z digest=sha256:9956a28e331f2f4ad5f8ef3d323a9a77e837c6d05319f565fdf4d683393e1868

Observation 749c09ca-c726-466e-bfc4-37015cd96f9b · inbound

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction cites this paper.

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction SciCode: A Research Coding Benchmark Curated by Scientists

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:05:06.667656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-15T06:04:03.605898Z digest=sha256:6fa40187916bc29aad8221c4fc70345f89906e80d13a3011f912743681699199

Observation 30a26250-82a9-4582-9e30-92048bc76ffe · inbound

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science cites this paper.

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science SciCode: A Research Coding Benchmark Curated by Scientists

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:38:12.125490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T10:36:09.724234Z digest=sha256:c0f522b5074b9fbdf7c2e1d1c8fc089278a2232256c989474c8ec17d105b1615

Observation 3597d05c-7c21-4363-9498-43fc22e9abea · inbound

AI for Auto-Research: Roadmap & User Guide cites this paper.

AI for Auto-Research: Roadmap & User Guide SciCode: A Research Coding Benchmark Curated by Scientists

Reference 204

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:33:12.560011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T10:30:50.256635Z digest=sha256:aee8b094a57cb4bdbadb8e5946c7332ea55c89949b8ac4305f5718734729ed48

Observation 8461a9a1-6228-4f84-bb37-e12aa9854a6f · inbound

AI for Auto-Research: Roadmap & User Guide cites this paper.

AI for Auto-Research: Roadmap & User Guide SciCode: A Research Coding Benchmark Curated by Scientists

Reference 203

Resolution
unresolved
no resolver link, observed 2026-08-02T13:43:48.326864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:43:48.326864Z digest=sha256:ee163cfa96ca7b9e3d5c05ca20d4f52dfbce55186614dc3f4c8dbc89dffd5deb

Observation 594472bb-ef4d-4262-91ca-611201ccb2c1 · inbound

Open-World Evaluations for Measuring Frontier AI Capabilities cites this paper.

Open-World Evaluations for Measuring Frontier AI Capabilities SciCode: A Research Coding Benchmark Curated by Scientists

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:39:43.760942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T06:38:51.427985Z digest=sha256:3c562f60f19aef78b177d5b939d6ed43cfb06e246045c01022d7125858ec2c06

Observation fe0e2b63-1e39-4c6a-8be7-353d8050740e · inbound

Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation cites this paper.

Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation SciCode: A Research Coding Benchmark Curated by Scientists

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:13:45.141431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T16:36:03.557894Z digest=sha256:a8ec4e1a61c3f44a2726d8303ba1c5487087fb7e2c8c1373d9a75d62550a200f

Observation 7dd2ce34-215a-455f-888a-04f918c13cd8 · inbound

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence cites this paper.

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence SciCode: A Research Coding Benchmark Curated by Scientists

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:23:24.667564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T12:13:58.111299Z digest=sha256:dfd07f47c8ce43e48d41edb19a712918b823a7e694ecfab9da2d7160e439b292

Observation aa669625-4e45-48c8-876c-282fd3192d40 · inbound

Algorithmic algorithm development with LLMs: A Case Study on LLM-Usage for Contraction Order Optimization in Tensor Networks cites this paper.

Algorithmic algorithm development with LLMs: A Case Study on LLM-Usage for Contraction Order Optimization in Tensor Networks SciCode: A Research Coding Benchmark Curated by Scientists

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:16:23.648169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T14:35:24.936368Z digest=sha256:b7419ed580cf28d0079b35318ab35c3d72253ee82423d87c096687fd3fa217f1

Observation 89a7cee2-b121-4937-8705-eabbf6158743 · inbound

Matter to Mechanism: A Benchmark for AI Co-Scientists in Materials and Battery Research cites this paper.

Matter to Mechanism: A Benchmark for AI Co-Scientists in Materials and Battery Research SciCode: A Research Coding Benchmark Curated by Scientists

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-06-28T12:12:07.890457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T12:08:10.552789Z digest=sha256:31f95adbb4d992fee7ccac451c2ac07821f93a2a1c933ad73b6b74f8cc083a9d

Observation e06e1321-082c-49de-92e7-7e99f29feeab · inbound

InquiTree: Evaluating AI Agents in the Scientific Inquiry Loop with Paper-Derived Research Trees cites this paper.

InquiTree: Evaluating AI Agents in the Scientific Inquiry Loop with Paper-Derived Research Trees SciCode: A Research Coding Benchmark Curated by Scientists

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:07:37.239543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T14:06:59.471772Z digest=sha256:a7926e1a7c7a6210153176a80a6a2c600b489fc3e317d417e7140e45d6c525f3

Observation c8187079-cfac-4b82-a5ec-4605193f6714 · inbound

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation cites this paper.

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation SciCode: A Research Coding Benchmark Curated by Scientists

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:56.399127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T01:07:49.603969Z digest=sha256:1facf371982e9e61314739453356fdcd55b91e4e0a770abca5c142eebd2bdd9c

Observation 466d84a5-23b2-4baf-a4db-9ca0daa3f42f · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity SciCode: A Research Coding Benchmark Curated by Scientists

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.115724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:f8e963ec51e72320aa1670a28fb538577f19cce11b58a1901b7c0091330f4376

Observation 30066e6f-b907-413c-86de-03e67bc40b24 · inbound

LLMoxie: Exploring Agentic AI for Scientific Software Development cites this paper.

LLMoxie: Exploring Agentic AI for Scientific Software Development SciCode: A Research Coding Benchmark Curated by Scientists

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T07:37:37.032195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T07:37:37.032195Z digest=sha256:416237e755296556f6e305cf7803c6a903ea2a1ada608cd8226e4d21557f943f

Observation 41d2b364-9ca8-404d-9205-a685dd5fa552 · inbound

From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis cites this paper.

From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis SciCode: A Research Coding Benchmark Curated by Scientists

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T14:26:52.999650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T14:26:52.999650Z digest=sha256:48db6287c2d9d200fc9c2024d260fcfd18adfd82dbf9c2e08d1683f6e8df683b

Observation 4f56e5f3-0eb6-4bff-bf67-53c535cb0799 · inbound

CLVisc Agent for autonomous relativistic hydrodynamics studies cites this paper.

CLVisc Agent for autonomous relativistic hydrodynamics studies SciCode: A Research Coding Benchmark Curated by Scientists

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T00:52:32.159176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:52:32.159176Z digest=sha256:f371be801aab6c2d3d056b10818e9289180bfb82ac57c1806c29cefa2d324c97