Pith. sign in

Paper Citation Record · LEDGER

ECBD: Evidence-Centered Benchmark Design for NLP

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2406.08723.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08723 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:01:31.533556Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T11:01:17.101876Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6ecbfa7b-ec7d-406d-8cb7-f60d42cbee48 · inbound

BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices cites this paper.

BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices ECBD: Evidence-Centered Benchmark Design for NLP

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T17:01:31.533556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:01:31.533556Z digest=sha256:630c43d3d71f4068d1195c7be00ba23a4311664a56043f09a2bc390ef1d3e61d

Observation 82b09957-73fb-468f-b5a9-02b6a6fbb6ef · inbound

More than Marketing? On the Information Value of AI Benchmarks for Practitioners cites this paper.

More than Marketing? On the Information Value of AI Benchmarks for Practitioners ECBD: Evidence-Centered Benchmark Design for NLP

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:41:51.676325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:41:51.676325Z digest=sha256:76dd4945aa660b7a4a12d9d1702e7e76af3abb3a96e3f0f0147417becfc092af

Observation 392c0350-9da5-427f-a879-abb4cde1c1c3 · inbound

Measurement as Bricolage: Examining How Data Scientists Construct Target Variables for Predictive Modeling Tasks cites this paper.

Measurement as Bricolage: Examining How Data Scientists Construct Target Variables for Predictive Modeling Tasks ECBD: Evidence-Centered Benchmark Design for NLP

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:06.102199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:06.102199Z digest=sha256:b94fc198b791db95f4d709638d28598dc27994bf20f4375224f0b07c85fa9525

Observation f650c0da-1250-4734-b8d9-f5d51213e267 · inbound

Toward Valid Measurement Of (Un)fairness For Generative AI: A Proposal For Systematization Through The Lens Of Fair Equality of Chances cites this paper.

Toward Valid Measurement Of (Un)fairness For Generative AI: A Proposal For Systematization Through The Lens Of Fair Equality of Chances ECBD: Evidence-Centered Benchmark Design for NLP

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T19:47:46.544145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:47:46.544145Z digest=sha256:d8d4e56ba71a316428b6e074a34daf50ae6b9514215ea6708a87a3e6c6cedfa9

Observation 08c6ec39-2564-4980-a553-244719dbf551 · inbound

Towards Real-World Validity in Generative AI Benchmarks: Understanding and Designing Domain-Centered Evaluations for Journalism Practitioners cites this paper.

Towards Real-World Validity in Generative AI Benchmarks: Understanding and Designing Domain-Centered Evaluations for Journalism Practitioners ECBD: Evidence-Centered Benchmark Design for NLP

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:01:17.103979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T10:59:16.139525Z digest=sha256:054f2682516f51a11895a30a6c605ca8ed88af30ff6eedf6ddf38da42723afdf