Pith. sign in

Paper Citation Record · LEDGER

Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2308.07921.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.07921 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T15:30:23.485228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:05:47.722524Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f045e8db-7a38-4741-ba4c-2e149467e263 · inbound

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning cites this paper.

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-17T23:46:39.679051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-17T23:46:39.330438Z digest=sha256:16bf6a267ff730c830b088866eb3340ce34a119f65f4ed551d81f4e82647b635

Observation 347f1a27-5ce3-441a-bc87-350987c6a856 · inbound

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving cites this paper.

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:19:36.202672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-19T09:19:35.918151Z digest=sha256:2081df72875fcfbf99e9f4a78a498fd1e9a1e287bd565e5252b9a92f6adb2032

Observation a2fa2715-fbc3-4580-9f60-aa9bdc76923a · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:27.558572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:31918d088bb6ea631b3c874cb977ec1a315f09d94842763c351ca74d2f0dcf58

Observation 2c6470d1-797e-4621-bdf4-c35bfb147b42 · inbound

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models cites this paper.

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:03:26.818551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-17T03:03:26.723464Z digest=sha256:94fe2fecc4624f0b8d20752076af8dceec1f3d0cb9fb6a8d17d57f29eed8aa91

Observation e7facf7c-f0db-41e4-b826-a2b3d64a6350 · inbound

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? cites this paper.

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:29:30.194704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-17T01:29:30.032408Z digest=sha256:221b988620f9e7385526338cf59af2b63bbf9e3966f4bb738384d76a525f231f

Observation f0975690-70c8-4c1d-991c-b99feeaa2ca5 · inbound

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark cites this paper.

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:05:47.726018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-23T20:03:38.336841Z digest=sha256:4ebdd3517b5fb42ef7257fdcd2a789dfe8e65b4e7380738516b9ea280bbb09db

Observation 4195cfc8-f499-430d-9dda-4b8b96997c42 · inbound

NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise cites this paper.

NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:56:07.223575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-08T16:54:17.663989Z digest=sha256:67aac02971ccf2a0ea79bd3bfdbaca8163571fc003b832f9e60882f728140408

Observation cdacb0a2-0273-4001-bb04-df4c0a37a362 · inbound

Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation cites this paper.

Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T15:30:23.485228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T15:30:23.485228Z digest=sha256:39d4467f0bfe5a3ede6a50bf2bc4db8c033a9aa48170fde0f248ec1680da36ba