Pith. sign in

Paper Citation Record · LEDGER

The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2305.07141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.07141 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:50:53.238710Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T00:02:50.462922Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 72a3e848-8351-4812-94db-87525415f8dc · inbound

From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark cites this paper.

From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:53.238710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:53.238710Z digest=sha256:59d8981855577b127603272e62ced92173b29a3e005d4b63eefa5d28084cdee6

Observation f710b1dc-d717-48f5-b6e8-1b3226d3f386 · inbound

EXP-Bench: Can AI Conduct AI Research Experiments? cites this paper.

EXP-Bench: Can AI Conduct AI Research Experiments? The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:20:45.976185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:20:45.976185Z digest=sha256:6e45a6b4466f362ca4bff04743822d48aedd292f244d4b51171c4d2c74ab2ce4

Observation 1a3a8214-6db7-4854-8c46-91ef7b765378 · inbound

Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey cites this paper.

Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T10:46:20.148999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:46:20.148999Z digest=sha256:358d38e8f796a9e3b7927699986ecc490db1013f7803820daeabba7719739877

Observation 398426f2-88c9-4b91-bf59-11acab7cceec · inbound

Less is More: Recursive Reasoning with Tiny Networks cites this paper.

Less is More: Recursive Reasoning with Tiny Networks The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:53:09.551996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T04:53:09.465781Z digest=sha256:c270fd21a67999c6c42f461426f7d3cc656cfa564d7bccac16f7f5861fe015a9

Observation d9965a2c-7799-4554-9e43-271b2abce97e · inbound

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator cites this paper.

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T21:06:30.009675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:06:30.009675Z digest=sha256:7f67c15d25a6fe0a3a42a7f07925ea8a237815c96d514f1f5eeb6b2ed20f62ff

Observation c72d0830-192b-4969-986c-eba9665bff35 · inbound

Gradient-Based Program Synthesis with Neurally Interpreted Languages cites this paper.

Gradient-Based Program Synthesis with Neurally Interpreted Languages The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:56:08.113351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T04:29:33.858344Z digest=sha256:c4f6a3fef19a7e98f408f6391977455163b212c816454794b5899caf37b5cfc9

Observation f27b2a9a-8254-461e-be74-f227a4f12d9a · inbound

RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs cites this paper.

RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:50:59.841006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:47:30.890858Z digest=sha256:e538f3125c289944a086c35b8caa8fc9d25fa2392ad8ca1800a128cd6307576e

Observation bee69c28-d92e-430f-827a-d52915de6276 · inbound

VisAnalog: A Diagnostic Suite for Visual Concept Transfer on Natural Images cites this paper.

VisAnalog: A Diagnostic Suite for Visual Concept Transfer on Natural Images The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:20:24.987449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T05:17:45.034348Z digest=sha256:939f29f77be9e446e0c428feb7b00850f4347abf84e7980a77129a563cc0192a

Observation 30d609d9-cbad-4152-9ff9-699239c859e1 · inbound

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning cites this paper.

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.464616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T23:15:56.598968Z digest=sha256:dd129911d761f57151f5d52fdc471ef324b93314f8d24960290e898eb5d4e2f7

Observation ffc1d1bb-0548-4580-a222-f889f709314a · inbound

Interpretable GOHR Agents via Sparse Autoencoders cites this paper.

Interpretable GOHR Agents via Sparse Autoencoders The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T00:31:49.692564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T00:31:49.692564Z digest=sha256:87780edb5f6c01c9d787fa057fe7427b395cbd4b6dd29d0bdbb91de9719af4fa