Pith. sign in

Paper Citation Record · LEDGER

Evaluating Language Models for Generating and Judging Programming Feedback

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2407.04873.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04873 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:25:55.263057Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T05:30:23.456663Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
pith, observed 2026-08-10T05:30:23.456663Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43bacf51-4b71-465c-83df-e9a3d512c10a · inbound

The Role of Generative AI in Software Student CollaborAItion cites this paper.

The Role of Generative AI in Software Student CollaborAItion Evaluating Language Models for Generating and Judging Programming Feedback

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T15:25:55.263057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:25:55.263057Z digest=sha256:0fd505de76a2767787ae3816d275c5cb027b0a207c3abd2fdd5ab4d2b3e1573d

Observation 292ac296-b475-4dea-aeff-e10867768c1e · inbound

BitsAI-CR: Automated Code Review via LLM in Practice cites this paper.

BitsAI-CR: Automated Code Review via LLM in Practice Evaluating Language Models for Generating and Judging Programming Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T14:40:07.146136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:40:07.146136Z digest=sha256:d2efc7fc51d5cec8208939a3e77f9e34448bea60d8acf7ec60830d7872d6145d

Observation 3a97827e-db70-4142-99df-94f726687d04 · inbound

FEAT: A Preference Feedback Dataset through a Cost-Effective Auto-Generation and Labeling Framework for English AI Tutoring cites this paper.

FEAT: A Preference Feedback Dataset through a Cost-Effective Auto-Generation and Labeling Framework for English AI Tutoring Evaluating Language Models for Generating and Judging Programming Feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:14.392716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:12:14.392716Z digest=sha256:d534e5d472beefb688838870c736c216a0f86b4c8062e46531778b2549d0ad7a

Observation ee3c2bad-fc0d-4f81-99af-0d5c3ca826fd · inbound

That's Not the Feedback I Need! -- Student Engagement with GenAI Feedback in the Tutor Kai cites this paper.

That's Not the Feedback I Need! -- Student Engagement with GenAI Feedback in the Tutor Kai Evaluating Language Models for Generating and Judging Programming Feedback

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:52:28.522144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:52:28.522144Z digest=sha256:9cf16cc8017228f0d3e0a920c77aba6384e211770ad014ca7189f568d6651404

Observation d9ccdde0-ae2c-48d6-859f-7bd278f7055b · inbound

On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization cites this paper.

On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization Evaluating Language Models for Generating and Judging Programming Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:10:23.715529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:10:23.715529Z digest=sha256:0d0768567fda30780f2f2cc1a3162f1a6ec1f774f5c9544055b3a536f36b73e2

Observation 7928f3cf-f799-45de-b4f7-186211f748b5 · inbound

Wisdom of the Crowd, Without the Crowd: A Socratic LLM for Asynchronous Deliberation on Perspectivist Data cites this paper.

Wisdom of the Crowd, Without the Crowd: A Socratic LLM for Asynchronous Deliberation on Perspectivist Data Evaluating Language Models for Generating and Judging Programming Feedback

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:48:35.462202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:48:28.615358Z digest=sha256:897eab1c6e83e9904f9cc269d8ef5f3b431f628863e7e2759d0f389ee07f9dd9