Pith. sign in

Paper Citation Record · LEDGER

Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2304.11164.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.11164 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:08:08.025968Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T22:28:59.707323Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3c39ff05-dc27-4795-82cd-3a5d395a57ef · inbound

Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations cites this paper.

Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.025968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:08:08.025968Z digest=sha256:092e4ed0ee20a5c3cae5e2268b70247eb9c53bde21bb61e7fde395bc5fb59697

Observation be08a52a-dfe6-45da-bd06-bf7d3d97ef18 · inbound

From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark cites this paper.

From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.103242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.103242Z digest=sha256:e0a7aa00af01f97135d4151f59fe2e82161633d065e8d3b6fc1dfd731f04fbb1

Observation 17eaa702-14d6-46c0-b172-226e24bd879a · inbound

MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models cites this paper.

MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:40:01.425301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:40:01.425301Z digest=sha256:45bc0b82e48ab967e02209d898f02084a7dafc27a99553ba6170d884f3464653

Observation 6ad35b72-3ee1-4c6d-b81f-bbdcb287468a · inbound

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi cites this paper.

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:18:13.594209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T11:18:08.326304Z digest=sha256:855c353d596496dfab480bfdfce819309c93437146f075d4e02fa8582ef9ce7c

Observation 09b2c60a-4f3c-41ca-9161-cba22ef04fd3 · inbound

GS-QA: A Benchmark for Geospatial Question Answering cites this paper.

GS-QA: A Benchmark for Geospatial Question Answering Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:34:31.671810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T02:33:47.304356Z digest=sha256:7829c85bb529cf058ab309174da5eb18cd00eed70eac3abdf4c6b3ec32f8ca4e

Observation 7119106f-d003-40e8-92ad-4debd76a80a3 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.484476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T05:41:46.057986Z digest=sha256:38e97345744af052d68a50459a9190a7f1a85b2984409b0111c03b741c8aee85

Observation 828d2b03-da5c-487d-91ec-f89c92fb5c1b · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:28:59.709846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-03T22:28:24.036122Z digest=sha256:724e5160bb23916844d84e3eea6527c2ddd266c0c61b962cf5a9c044c8370920