Pith. sign in

Paper Citation Record · LEDGER

InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2306.14898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.14898 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:23:50.664728Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T11:23:20.895534Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4e88b8a0-fefd-425c-8f95-4aee49b6af53 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.507302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:70bdb9a1685373544ad3b6e59f371c916976377254a62d5c4e19fde2fd1cc3ee

Observation c7fad0a8-726e-4417-be4f-a7eba076ae48 · inbound

Gemma 2: Improving Open Language Models at a Practical Size cites this paper.

Gemma 2: Improving Open Language Models at a Practical Size InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:11:16.397576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T12:11:16.326752Z digest=sha256:fb5cf2a98b07bd697733092aba93f3a4b5dcb6e4ee699d86b245765861b88d02

Observation 539ad4c5-a155-4bb9-aa70-e2f6f2e740f2 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 147

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T12:04:10.693845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:3f871c9420d68a5052463806e2916f7d220b2b7cd9358230159c50854760c075

Observation a05762d5-d970-4558-a807-414ce6e31ab7 · inbound

OS-ATLAS: A Foundation Action Model for Generalist GUI Agents cites this paper.

OS-ATLAS: A Foundation Action Model for Generalist GUI Agents InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:29:27.725374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T09:29:27.173784Z digest=sha256:10b8aff5d8d6a0d4ad43934fec7cceb3eb5a38ce3a9ec66224b9f8544fc14e4d

Observation fe4b2975-5b95-4c5c-9a92-7fd4b40436d2 · inbound

OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization Problems cites this paper.

OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization Problems InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:23:50.664728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:23:50.664728Z digest=sha256:901e753a45c46fd73a1b094c16620531e39902bbe8b600e820aa1ec10cecdb8d

Observation dc6a2e0e-4271-4bd0-820f-bdcb8a7cdd84 · inbound

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities cites this paper.

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:52:07.763204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T05:48:02.828938Z digest=sha256:192523d059a3c7fb02add633f49a99d3e8f74f2e5eda85faa93e967ac5601840

Observation 421b26c3-468b-4eec-bfc1-9e04165ceaef · inbound

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation cites this paper.

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.731715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T04:37:33.942379Z digest=sha256:ddc1dc01dd368a835764f5e0097756886bb991e861cfaf692fb8a0b88a77dd43

Observation 4f30a2d7-2449-48b5-a025-1fc9b2f9f650 · inbound

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security cites this paper.

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:25:15.844022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:25:15.844022Z digest=sha256:83dff68414c39fab8b2322ed048bde90f098456b594f7db6dd35bb78edbfc4fa

Observation ff2ff9f9-012e-4dad-932e-cec5c038a972 · inbound

Feedback-Driven Execution for LLM-Based Binary Analysis cites this paper.

Feedback-Driven Execution for LLM-Based Binary Analysis InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:44:37.936889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T10:40:32.133423Z digest=sha256:012bc6d5aae42007360cfa78b4ecdcbd01519b1741e14373b219dd2b58a9b31d

Observation 27b5ac5f-7a8c-49a3-a988-b6d093ff4ecb · inbound

Towards Optimal Agentic Architectures for Offensive Security Tasks cites this paper.

Towards Optimal Agentic Architectures for Offensive Security Tasks InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:16:03.139155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:02:04.359269Z digest=sha256:65186ac74fa91f43b223dda77a8161df8cf9fb3d8cf8b119b34af5ef72c50694

Observation 4842b0c6-41f6-44f9-9f67-e5f2811cbac0 · inbound

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models cites this paper.

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:54:48.809796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T00:50:27.257888Z digest=sha256:39a985e16bcc990eb7f8ea1e791e26080a595f2ceab6d23effefc3173870b0f6

Observation 4759661e-1d2d-4fcd-b639-63c929da966d · inbound

Toward Scalable Terminal Task Synthesis via Skill Graphs cites this paper.

Toward Scalable Terminal Task Synthesis via Skill Graphs InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:51:16.938897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T16:13:50.484665Z digest=sha256:25cb0c9e58879a37c49d1b0d83777a0f11adfce42d38752a2bc4c564c31950c8

Observation c5a1fe64-f8c3-442e-bf6b-15e7ce40cc5b · inbound

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces cites this paper.

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 141

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.775762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T02:57:15.521594Z digest=sha256:a3fbf65cbef860b0d79f0be6a2565acf0891c9920e96ef481e9de08e95db0e4e

Observation bedc2437-7737-4586-b6fb-561fb2ff7552 · inbound

CrackMeBench: Binary Reverse Engineering for Agents cites this paper.

CrackMeBench: Binary Reverse Engineering for Agents InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:56:43.456149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:44:13.078520Z digest=sha256:bde7e39c7ab10958db033a44c9a5684c4580eb2c25581077b1265777157bd39c

Observation cfefc68d-e3a0-4e4b-8aa8-a6f859724145 · inbound

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback cites this paper.

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:23:10.555911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T09:22:06.285118Z digest=sha256:ea0242625e551482485609fda564faeec2dc91a2f5e3929e026527f3fd9e04bb

Observation 06905fc6-0fae-4c00-9649-0a9350f6218b · inbound

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning cites this paper.

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T11:23:20.897179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T11:19:38.959705Z digest=sha256:1f1261603843b418b671c33a84f84a2c89a17dac2fabf44db3cfcdf8f3a9802a

Observation cde89d7d-a132-4ca7-81a8-df4fc26dbda3 · inbound

Token Reduction Is Not Cost Reduction cites this paper.

Token Reduction Is Not Cost Reduction InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T06:45:15.404937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:45:15.404937Z digest=sha256:f1c86f1ec1f0af99ad70ca9441c198a945c2fea00327a1ea2462b081f84cd5b4

Observation b5f905ca-bc27-4185-9528-d5ddceb2e1d9 · inbound

Token Reduction Is Not Cost Reduction cites this paper.

Token Reduction Is Not Cost Reduction InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T04:23:32.473137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:23:32.473137Z digest=sha256:07166af1824d96128cfe33127d9ac3c48216b436200275a7ef98fc34d40efeaf

Observation 523c8696-edfe-4ff8-9149-b511a4f7c8d4 · inbound

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play cites this paper.

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T02:31:26.014099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:31:26.014099Z digest=sha256:01b28e725b08d3b395df286f3f327eb3d935f7984c1730154e0681c918577a61

Observation e1d6829c-1f02-40f2-920f-7b5caa53f733 · inbound

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates cites this paper.

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T00:48:33.449476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:48:33.449476Z digest=sha256:210cf5db95d4be81a608c3832274fa4e198a8f035e20358c88f1338044834417