Pith. sign in

Paper Citation Record · LEDGER

On Benchmarking Human-Like Intelligence in Machines

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2502.20502.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.20502 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:40:38.848439Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T22:40:43.385937Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c6fa9a29-8b5d-4f61-90c6-ee414cc17cf0 · inbound

Are Large Language Models Reliable AI Scientists? Assessing Reverse-Engineering of Black-Box Systems cites this paper.

Are Large Language Models Reliable AI Scientists? Assessing Reverse-Engineering of Black-Box Systems On Benchmarking Human-Like Intelligence in Machines

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:40:38.848439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:40:38.848439Z digest=sha256:c966938459b98cc9363ae81a54cf43d9fb31e901050255281d8c87446fa66081

Observation f0712092-7ecb-455b-847e-294c7a6321be · inbound

Belief Attribution as Mental Explanation: The Role of Accuracy, Informativity, and Causality cites this paper.

Belief Attribution as Mental Explanation: The Role of Accuracy, Informativity, and Causality On Benchmarking Human-Like Intelligence in Machines

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:48.245967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:17:48.245967Z digest=sha256:c826137a906030852f36fd10e9543377d42d2572582f8b5115f9179f732c6c17

Observation 2bef12fa-7b3c-42ec-b865-707b211ee34a · inbound

Using LLMs to Advance the Cognitive Science of Collectives cites this paper.

Using LLMs to Advance the Cognitive Science of Collectives On Benchmarking Human-Like Intelligence in Machines

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:40.641054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:00:40.641054Z digest=sha256:156217290f6008f90b842e85d5d30e0d9cd23561e98c98c84c348831d74cdf5e

Observation f1bfedf2-d615-4aa1-914d-daa2980f72f4 · inbound

What's in the Box? Reasoning about Unseen Objects from Multimodal Cues cites this paper.

What's in the Box? Reasoning about Unseen Objects from Multimodal Cues On Benchmarking Human-Like Intelligence in Machines

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:26:12.008828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:26:12.008828Z digest=sha256:0c2e3bc78555539605001d9ff2f0b07b7db2b757f2ed88efc53b284ddb37c927

Observation 7a926ee0-43e9-42a9-a26a-69b226da4aa3 · inbound

Modeling Open-World Cognition as On-Demand Synthesis of Probabilistic Models cites this paper.

Modeling Open-World Cognition as On-Demand Synthesis of Probabilistic Models On Benchmarking Human-Like Intelligence in Machines

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T16:50:50.547105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:50:50.547105Z digest=sha256:8f8945b10ab14b5bebf9071c0f5992752d0ca2f364c5ef3819a1b0a9eb127f03

Observation ed2d6c68-bea8-4dbe-a71b-ffac6ef24556 · inbound

Measuring and mitigating overreliance to build human-compatible AI cites this paper.

Measuring and mitigating overreliance to build human-compatible AI On Benchmarking Human-Like Intelligence in Machines

Reference 131

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:40:43.388238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T22:37:37.267715Z digest=sha256:89b1d9b3a40efb3b4a68e487d79815ce916fa92b6b29b4ed78e741cc0bcaaa0e

Observation ec556cb3-cda5-4a1c-8b00-db7de7bddacd · inbound

Predicting Biased Human Decision-Making with Large Language Models in Conversational Settings cites this paper.

Predicting Biased Human Decision-Making with Large Language Models in Conversational Settings On Benchmarking Human-Like Intelligence in Machines

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T10:11:07.889923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:11:07.889923Z digest=sha256:ef91da4453515b7c8c052095d8b7fdc6007404bd3b462e6cc3c40a9b14054f64