Pith. sign in

Paper Citation Record · LEDGER

Evaluating the Robustness of Analogical Reasoning in Large Language Models

As of 17 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 12 inbound Pith citation observations for arXiv:2411.14215.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14215 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:29:23.569643Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:09:59.546494Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:58.093447Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f6797330-2437-46b7-9056-dfa6ea9855cf · outbound

This paper cites Shortcutted commonsense: Data spuriousness in deep learning of commonsense reasoning.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Shortcutted commonsense: Data spuriousness in deep learning of commonsense reasoning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.688800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.274831Z digest=sha256:b5fcb42a026ead22aaa127f40c201e2ad6b2e4dbd02001bce0f8ab255eb06d43

Observation 0a6294b0-65b0-4038-aa88-006428ab3531 · outbound

This paper cites Understanding the source of semantic regularities in word embeddings.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Understanding the source of semantic regularities in word embeddings

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.673195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.281026Z digest=sha256:cf8f6342c1b37bf2ece14bc91cf39409cb2c90ead5972036e6079a92fabdec3f

Observation d9cf3482-c79e-462f-893d-66a766ee6417 · outbound

This paper cites Leaping across the mental canyon: Higher- order long-distance analogical retrieval.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Leaping across the mental canyon: Higher- order long-distance analogical retrieval

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.645656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.287753Z digest=sha256:259e64d756bc99ad2d60f483379c16b8bc97a4ece1064fb65d947249eff2cb93

Observation 7ee07644-1b0d-4b5e-bb5b-31a30da256f5 · outbound

This paper cites Hwang, Soumya Sanyal, Sean Welleck, Xiang.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Hwang, Soumya Sanyal, Sean Welleck, Xiang

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.623390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.293017Z digest=sha256:2e8edc6c2590af37d4ca54f7ebf27da7199aa9872e2c805e78098cda2520c2b7

Observation 8bdba1c0-b678-4fdb-b90a-92837e2bebb0 · outbound

This paper cites an unresolved cited work.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:29:24.597307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.304451Z digest=sha256:fdec9e52e2d17507cbf41441682708b2ad6427a38d57ef6f4cb0587c94264147

Observation 1b79503a-869b-4f86-8982-5a55da24204b · outbound

This paper cites Geake and Peter C.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Geake and Peter C

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.579803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.310214Z digest=sha256:d66b4e43901a27fb9290d0894cd1e34f1e74316f4919586c242c09c59ceceea0

Observation 413fbef8-e7ff-40f0-9972-591f6b891f69 · outbound

This paper cites an unresolved cited work.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.321430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.321430Z digest=sha256:9bcdc033b772753854bff5e51502b93414b0279b0efaf10b79e9fa9e11f7c27d

Observation 0e9e5f03-d05d-4dd7-996e-ef2d58b25163 · outbound

This paper cites Response: Emergent analogical reasoning in large language models.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Response: Emergent analogical reasoning in large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.327020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.327020Z digest=sha256:5b51d191c952b42a419a4f1413e5b7c01e91278f09f5c43f854b6bf3e26a3f5e

Observation 6eb15d4e-8f80-41cb-8e92-618b11c1b600 · outbound

This paper cites Hofstadter.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Hofstadter

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.562169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.338307Z digest=sha256:ad9d0a70debf68c2a08680a592a4b634ae79a45d4dc1144af313c8d6f58b8122

Observation 1ee12367-7925-4bd9-a07b-e1caf3fb42d4 · outbound

This paper cites Hofstadter and Melanie Mitchell.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Hofstadter and Melanie Mitchell

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.539408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.343017Z digest=sha256:09d18c8d8ae66090fd866ba4a2ff9f01a226747b498cc3d8e675f02b79ad0c71

Observation cf364608-c4ac-4868-a8dc-109150435dba · outbound

This paper cites Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.349868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.349868Z digest=sha256:ced1298a68f5df41cc86bf4d4fe6e8edcb9fe04f8829f593b99ba07b17229064

Observation 2bf5af2c-5eb7-4b98-a555-24819af0c8f5 · outbound

This paper cites Towards Reasoning in Large Language Models: A Survey.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Towards Reasoning in Large Language Models: A Survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.360740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.360740Z digest=sha256:5b531a82200e64f3e87d7090203e2579e237abe787e6a08b6ca6aee356c92912

Observation d9e7432b-2ccc-4fc9-a76a-d020d274200a · outbound

This paper cites Toward best research practices in AI Psychology.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Toward best research practices in AI Psychology

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.367341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.367341Z digest=sha256:d5653e03d4b8fd05de9b77d7fe6331bab47e20b15ea2de459a7ea0ece9481882

Observation 5251ec88-cee2-4e28-9107-1dc43845f42a · outbound

This paper cites A Peek into Token Bias: Large Language Models Are Not Yet Genuine Reasoners.

Evaluating the Robustness of Analogical Reasoning in Large Language Models A Peek into Token Bias: Large Language Models Are Not Yet Genuine Reasoners

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.374654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.374654Z digest=sha256:53e29af6a3c1da6e16cc35172e2c774a6e395418093926268172e14210b5f55e

Observation e7fac589-e40a-4f1b-ad57-e081f0df073a · outbound

This paper cites an unresolved cited work.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.382494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.382494Z digest=sha256:b50b8be22369b6490c8edf99a92b427fda0e8c2516612d26909698bd91e87bdf

Observation ee2d7aab-9bd3-4556-80a5-43e8d69b5792 · outbound

This paper cites Can LLMs really reason and plan? Com- munications of the ACM , 2023.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Can LLMs really reason and plan? Com- munications of the ACM , 2023

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.515714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.391772Z digest=sha256:ab894d7d8303ca276b6740c1a3b0ac11449e3e1c0cb657e0b5e0506a8872ed2f

Observation 0f88c4df-1292-4684-ace9-85ebef76491b · outbound

This paper cites Lampinen, Ishita Dasgupta, Stephanie C.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Lampinen, Ishita Dasgupta, Stephanie C

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.499995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.396721Z digest=sha256:307971b8a35453daa436d81f7eb8a96a2529ca6eef07dd61e47f34f706a623ee

Observation 921a8974-88ac-4261-9a11-5ad6e242ecef · outbound

This paper cites Event-related potential responses to letter-string comparison analogies.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Event-related potential responses to letter-string comparison analogies

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.483307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.400818Z digest=sha256:ee63566ababc3a0317f3cc658d52d493d91c89d4ec32136cf698aa341374c1f5

Observation 26aa35a6-543b-4f0c-ad09-7f153e4316b3 · outbound

This paper cites Recreating Raven’s: Software for systematically generating large numbers of Raven-like matrix problems with normed properties.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Recreating Raven’s: Software for systematically generating large numbers of Raven-like matrix problems with normed properties

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.460576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.407146Z digest=sha256:a46b3de0a1cf2b8546fbfb1d613199ef9f87a5b6ad00616c47d4ed6cd24d2df6

Observation b7ca1d4a-21fc-4766-8c0b-feb876c65462 · outbound

This paper cites Thomas McCoy, Shunyu Yao, Dan Friedman, Matthew D.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Thomas McCoy, Shunyu Yao, Dan Friedman, Matthew D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.425204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.417558Z digest=sha256:c580bfb9281542de1190108777951b81fa8900b2c99f4f4ea498d4a5a4359693

Observation be6f9df6-38f0-4270-be46-67bdd28de7cf · outbound

This paper cites When a language model is optimized for reasoning, does it still show embers of autoregression? An analysis of OpenAI o1.

Evaluating the Robustness of Analogical Reasoning in Large Language Models When a language model is optimized for reasoning, does it still show embers of autoregression? An analysis of OpenAI o1

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.422926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.422926Z digest=sha256:8b864bf920d599aff0f819d4e28dfc75ee9d0c49744b83f8a21d890303d6b61e

Observation 8b410f02-3a27-468c-8a20-692042a8134a · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Evaluating the Robustness of Analogical Reasoning in Large Language Models GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.430844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.430844Z digest=sha256:85d2dc754f8b000c4d8e1ebe7de08767be1c3b7fa05e39ee1b80d28842d37ea8

Observation 70eff4da-04d3-41a2-8034-0e4711e4aa8f · outbound

This paper cites Analogy-Making As Perception: A Computer Model, chapter 5.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Analogy-Making As Perception: A Computer Model, chapter 5

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.396773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.441500Z digest=sha256:d9a504d71ed1569671a6d018644d5ab37f6ef778ea69964ff906458ddad2817b

Observation fa2579d0-b19d-47bf-9a75-40679fd91524 · outbound

This paper cites Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.456184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.456184Z digest=sha256:4c17247a5881dd41752b633a79cdcbee767f6899909eb96222df42ba70bea466

Observation 9d295830-9e8a-46d2-be46-80c5dbac89ad · outbound

This paper cites Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.465067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.465067Z digest=sha256:5e02ea25d3924b166e667e1c2da5a6e13bdd5fc0a28ce334bfa8b770f9cab2af

Observation dad1a983-37b6-43b8-9ef3-ddb614f9340b · outbound

This paper cites Learning to reason with LLMs, 2024.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Learning to reason with LLMs, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.379644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.472699Z digest=sha256:6a06f0eec7f979df5c86799d403d67df6936e520d4b2baba2630a6c8b82d5f22

Observation 4dd219b2-251e-4ba7-9035-ae809d4c81c0 · outbound

This paper cites Deciphering the Factors Influencing the Efficacy of Chain-of-Thought: Probability, Memorization, and Noisy Reasoning.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Deciphering the Factors Influencing the Efficacy of Chain-of-Thought: Probability, Memorization, and Noisy Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.481760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.481760Z digest=sha256:6af9a5694fdfe4eb53a43e646757d34c2a7e3e25f9557d75a422662e07b9a0b9

Observation 639b0fad-dd04-49be-a507-2130a56e0866 · outbound

This paper cites an unresolved cited work.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:29:24.357036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.492907Z digest=sha256:9bafcf5dcf8c52882dadf335580b21ea11aa7d3b11a66be4949f5c87427e189f

Observation c716b057-9349-4d2c-ab59-86076098dc17 · outbound

This paper cites Logan IV, Matt Gardner, and Sameer Singh.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Logan IV, Matt Gardner, and Sameer Singh

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.328237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.498841Z digest=sha256:197d8d84f27839dc18611f8695ecb9917a0556d841ba1ffe8c20c460f5724007

Observation bbc9c4a7-997e-4aa2-b948-23f2ec1f15fa · outbound

This paper cites ARN: Analogical reasoning on narratives.

Evaluating the Robustness of Analogical Reasoning in Large Language Models ARN: Analogical reasoning on narratives

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.299505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.507021Z digest=sha256:e1b638028ecfea1e7b01dd1bb8c33b5721cc92414154d55e304d6ec54766bf36

Observation 03d3a4d9-d130-4c9b-8da9-0eac78933716 · outbound

This paper cites Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.513851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.513851Z digest=sha256:b80b5417276c61d20be4a0902c9a1279d6bcfc2ea6ea32f61fa762116c4204d6

Observation faf66272-57bd-4858-aeea-fad22e5b86c6 · outbound

This paper cites Stevenson, Alexandra Pafford, Han L.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Stevenson, Alexandra Pafford, Han L

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.520685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.520685Z digest=sha256:cf7a13b3a2e438fc330d7abe04497f3b87490664ae742b3bec4376634582c31e

Observation fea2a709-2ef7-4aeb-acae-8bc9b0f13de9 · outbound

This paper cites MMLU-Pro+: Evaluating Higher-Order Reasoning and Shortcut Learning in LLMs.

Evaluating the Robustness of Analogical Reasoning in Large Language Models MMLU-Pro+: Evaluating Higher-Order Reasoning and Shortcut Learning in LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.528914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.528914Z digest=sha256:8033fbcf6a72c22f2dd8fd389db188fe77fbe830ca4887d12ab0f979afd94544

Observation bf4e7ee1-04ae-4dc5-9c3a-6804da0f50e6 · outbound

This paper cites Dual process theories: A metacognitive perspective.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Dual process theories: A metacognitive perspective

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.237128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.536751Z digest=sha256:ec83754a42a79b66559fb75d294802c93ab42ac341a525e23fbdb6f4b212f863

Observation 48a18c36-5358-45b6-85ba-b617c86cac5e · outbound

This paper cites Holyoak, and Hongjing Lu.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Holyoak, and Hongjing Lu

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:29:24.220397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.548319Z digest=sha256:a4c85d00c95ce79d02140ca642f5d1c2dc519978771b371982cb7d14641c26fe

Observation a2154751-175d-4849-9781-679d4cee47d6 · outbound

This paper cites an unresolved cited work.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:29:24.182278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.553336Z digest=sha256:7eb21b4852085ed20181cb21abc4a10bf38ff04dd122653594907635d2034795

Observation 4be489d5-9272-4f91-a529-f1d35558c386 · outbound

This paper cites an unresolved cited work.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:29:24.160409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T15:29:23.559037Z digest=sha256:6f1a78137865cb0dcb4dcccb7a14947356af2f9d5431e6ce00cb187dd64fd2d6

Observation e4ff6ac4-3f87-425d-8de3-e6a9b6dc99b9 · outbound

This paper cites Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.564113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.564113Z digest=sha256:3f64b2efeefc33b3d69b3b91e91a74bce23f90d31019979df8a9d1606cd58b17

Observation 24787b38-ed46-4eff-b163-ec0f3a76cf44 · outbound

This paper cites Do Large Language Models Understand Logic or Just Mimick Context?.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Do Large Language Models Understand Logic or Just Mimick Context?

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.569643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.569643Z digest=sha256:2660824b615fdaef480d2325189754e414a2326974f32cef1ed0c1ff89154751

Pith citing papers

Observation f2239600-4e39-49c4-9d90-90fff3347a88 · inbound

The broader spectrum of in-context learning cites this paper.

The broader spectrum of in-context learning Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T22:10:22.783461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:10:22.783461Z digest=sha256:1fbe8b009b962e6208ffd1ebe02e7000917a1a425685e1d2f5470edc1eb53486

Observation e7d49004-4d1f-482e-b41d-28a37cf6b613 · inbound

Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning cites this paper.

Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:37:23.909879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:37:23.909879Z digest=sha256:efbbb07b31deaaa78be8ae99f7b02f0c9acac60d34d849188eb28c459428536f

Observation 86f4fb34-241e-452a-8ee5-9011c9417d3c · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:33.270067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:89fb6024826aff975be7fc4e44f8cc6907774fdd35e5bbb41237482d5d01fbf9

Observation 17d83803-132b-467d-96ad-d10b6fab8ca8 · inbound

Toward Reasonable Parrots: Why Large Language Models Should Argue with Us by Design cites this paper.

Toward Reasonable Parrots: Why Large Language Models Should Argue with Us by Design Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:09:59.546494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:09:59.546494Z digest=sha256:5b4a0c9359ccc28acd7c3949c917c7e2a1b4d3834bdafe4f6ef03264beb0527a

Observation 9a86c7f8-cdb4-41b4-8f8b-aa64fa20bf1f · inbound

RE-IMAGINE: Symbolic Benchmark Synthesis for Reasoning Evaluation cites this paper.

RE-IMAGINE: Symbolic Benchmark Synthesis for Reasoning Evaluation Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T19:40:08.810909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:40:08.810909Z digest=sha256:d8ea50772d449769d18b891889a5ffe9ca6232cd93a1b55297552e9774fd4474

Observation c1938ca5-e950-4a21-b155-715664967264 · inbound

Mechanistic Interpretability Needs Philosophy cites this paper.

Mechanistic Interpretability Needs Philosophy Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T23:50:47.367863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T23:49:19.683025Z digest=sha256:139d481034f1bd930cbcac2a70dd9a6b09c6d2b2677bca7f73f4fbe40732db15

Observation d87c7968-9b19-4564-bd8b-cbfae4473aab · inbound

Large Language Models Show Signs of Alignment with Human Neurocognition During Abstract Reasoning cites this paper.

Large Language Models Show Signs of Alignment with Human Neurocognition During Abstract Reasoning Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T21:10:32.572118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:10:32.572118Z digest=sha256:26cf90e41c2861690ff7473d3d217050f2b3cbb21095350771e465d4f7258244

Observation 25a96312-2b0b-475c-91b2-649f60b911e9 · inbound

On Robustness and Reliability of Benchmark-Based Evaluation of LLMs cites this paper.

On Robustness and Reliability of Benchmark-Based Evaluation of LLMs Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:31:02.177268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:31:02.177268Z digest=sha256:5e09ac8f8837df02d3ab73c8574b620874c6b3663158e72ae46225e951585dd3

Observation a0322c30-3942-409f-a309-44e1e89970ea · inbound

Can Large Language Models Generalize Procedures Across Representations? cites this paper.

Can Large Language Models Generalize Procedures Across Representations? Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T05:00:57.477973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:00:57.477973Z digest=sha256:48fa5bc8b289bae812ed7f925f8ad098115852e5c83ce3bfba4c549c9fb2e2de

Observation 838842db-da98-4aa3-9e63-08afcba0ecf0 · inbound

Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid cites this paper.

Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:56:06.210285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-09T14:33:11.033906Z digest=sha256:8aef24b9515cb3e054eefbe59831dcedc5b4d5b39aadacbfef9ec53ce4420b4b

Observation 1ca954c3-7fbb-4e6c-af01-05b0621262f0 · inbound

AGC-Bench: Measuring Artificial General Creativity cites this paper.

AGC-Bench: Measuring Artificial General Creativity Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:36:56.051567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-07-02T12:33:53.029578Z digest=sha256:762b9b879d7dadfaaec1ead65eec44d9f4db35c94671b2c34f5a134b88a3ef7d

Observation 08c71db3-2592-4c22-8b56-38fb14706716 · inbound

AGC-Bench: Measuring Artificial General Creativity cites this paper.

AGC-Bench: Measuring Artificial General Creativity Evaluating the Robustness of Analogical Reasoning in Large Language Models

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:28:58.095269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-07-03T21:25:14.030920Z digest=sha256:06f2c2ef12c8c2fc43a5b9dd4b971ce8338bce4487ab6469770b6efebde27ba8