Pith. sign in

Paper Citation Record · LEDGER

SPHERE: An Evaluation Card for Human-AI Systems

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2504.07971.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07971 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:52:00.830672Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T05:41:23.886026Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2f157041-c928-402a-ba3f-bda948d890d1 · inbound

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy cites this paper.

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy SPHERE: An Evaluation Card for Human-AI Systems

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:00.830672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:00.830672Z digest=sha256:797b308a37098a977158a928a90d99c46de3c33010fc8b8561eb511ec87234a9

Observation f94f7301-ce97-49d9-8ff7-15993292d181 · inbound

Can Agent Benchmarks Support Their Scores? Evidence-Supported Bounds for Interactive-Agent Evaluation cites this paper.

Can Agent Benchmarks Support Their Scores? Evidence-Supported Bounds for Interactive-Agent Evaluation SPHERE: An Evaluation Card for Human-AI Systems

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T02:16:07.662383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T05:05:55.592359Z digest=sha256:fe382a5f0b924117e9103529c1c1c8f3f21c796b6beef8b3e40ff55224050511

Observation 846035ee-f750-44b8-abfe-56b15411a040 · inbound

L2-Bench: An Evaluation Benchmark for Measuring LLM Capabilities in Second Language Education cites this paper.

L2-Bench: An Evaluation Benchmark for Measuring LLM Capabilities in Second Language Education SPHERE: An Evaluation Card for Human-AI Systems

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-13T06:19:53.291826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T06:19:53.291826Z digest=sha256:a0591a2d9f5500fc5d68086c3866b6f9853c24c343ddb5bb535b3c1824baa434

Observation b0da8545-2cb9-4422-8eb7-bf8b34241df2 · inbound

L2-Bench: An Evaluation Benchmark for Measuring LLM Capabilities in Second Language Education cites this paper.

L2-Bench: An Evaluation Benchmark for Measuring LLM Capabilities in Second Language Education SPHERE: An Evaluation Card for Human-AI Systems

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-02T07:47:36.590139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:47:36.590139Z digest=sha256:bb0fb1f2134e94d508393a14f25d5d61175149d28cb2018f39c05a4e9559730e