Pith. sign in

Paper Citation Record · LEDGER

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

As of 8 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 6 inbound Pith citation observations for arXiv:2506.02126.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02126 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:35.587634Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:37:35.270920Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:59.009997Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 51c71118-a503-411f-8860-e4a5eddd26f9 · outbound

This paper cites https:// artofproblemsolving.com/wiki/index.php/AIME_Problems_and_Solutions, 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains https:// artofproblemsolving.com/wiki/index.php/AIME_Problems_and_Solutions, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:41.269022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:29.013985Z digest=sha256:9afe32757f06e4a09cd2d3b25f87ca25d1b542feaea1cac9619e7156b23ec714

Observation f07810c3-0ba2-414b-ad76-d559f4075738 · outbound

This paper cites https://www.vals.ai/benchmarks/ math500-03-24-2025, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains https://www.vals.ai/benchmarks/ math500-03-24-2025, 2025

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.961493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:29.108346Z digest=sha256:51403cc8dab166b487461392b04f5a071302474e41ae1879e20fa1858c59ed37

Observation 012733b1-6bf0-482e-bc68-8b9ce1f518c2 · outbound

This paper cites https://www.maa.org/math-competitions, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains https://www.maa.org/math-competitions, 2025

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.720342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:29.222146Z digest=sha256:d854e89a593d957f77520887e8c1ec08d4028c9ae0232e5299d9b09de9e866b1

Observation e033fb29-e84b-4ecb-aa90-7e63ba97413c · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.341067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.341067Z digest=sha256:c5f1a0ba44401c98d8cc62a415e9a71f2c5677048e9b4f3ba505801a39b2905f

Observation a381e39a-81c4-4b8d-b03a-f5c4c46c26f0 · outbound

This paper cites Large Language Models for Mathematical Reasoning: Progresses and Challenges.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.433062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.433062Z digest=sha256:4fb7f492ba6077f3acfed588678064dccbbdfd1df8318a6de60e1f7be8754be2

Observation d6c9ac2b-9ca5-49c2-bdce-fda723483429 · outbound

This paper cites Llama-nemotron: Efficient reasoning models.arXiv preprint arXiv:2505.00949, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Llama-nemotron: Efficient reasoning models.arXiv preprint arXiv:2505.00949, 2025

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.562739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.562739Z digest=sha256:c2636d9a2455b0bd0d28222de1940e054722127e17bc40c6552e465c09066b8d

Observation e55064ba-3086-4225-919a-7f1cec42a7f6 · outbound

This paper cites Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, andet al.Language models are few-shot learners.Advances in Neural Information Processing Systems, 33:1877–1901, 2020.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, andet al.Language models are few-shot learners.Advances in Neural Information Processing Systems, 33:1877–1901, 2020

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.377366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:29.713207Z digest=sha256:377c81d22772777da2b7ac3fe991e2ec290926cd2b9744041609863ca01ec405

Observation e7f0ef46-a258-4dfc-9f35-5d18dd21565a · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.797685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.797685Z digest=sha256:2834644dac46a5acad80ab84169e6ec8165b1cc9323fa1ad9b14ab2bbf49d1ae

Observation 8187428a-5f7f-447f-8618-a5a8e655858e · outbound

This paper cites HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.903235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.903235Z digest=sha256:8f5c58699024f67549ea3d749977abd47f575b0c19086ceafe2f4ce0118f6dca

Observation 12be5eea-c6d1-4762-8df6-7246e735a1c7 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.050544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.050544Z digest=sha256:9904e5dc2192c1a7f606a62230ad354122dd172d68337bf3f149bd81209172a3

Observation 20a7ff3a-fdcd-415c-bd86-c94e2f25f8a9 · outbound

This paper cites Deepseek -r1-distill-qwen-7b.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Deepseek -r1-distill-qwen-7b

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.134256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:30.146998Z digest=sha256:b863386f81da95a15ef7dd611e4562e31ec8cd7032f956fabc017d18314458b7

Observation f9f5d7ca-ffdd-40ea-90e8-1daa4de7e16f · outbound

This paper cites Rlhf workflow: From reward modeling to online rlhf.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Rlhf workflow: From reward modeling to online rlhf

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.823339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:30.259099Z digest=sha256:509b438f7571e4cdb2cd88463e1998561e65cb77522a2fcf8a0035247e85246b

Observation 587cfbef-66d5-486e-9b5b-5dd04b154ec3 · outbound

This paper cites Large language model influence on diagnostic reasoning: a randomized clinical trial.JAMA Network Open, 7(10):e2440969– e2440969, 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Large language model influence on diagnostic reasoning: a randomized clinical trial.JAMA Network Open, 7(10):e2440969– e2440969, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.604337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:30.368609Z digest=sha256:0bf2bca7313c35da382c4fa560d049cd3def531b3416a2c146560cd685d75cd8

Observation fe22c061-ad7b-4ba9-9e23-205a1610624f · outbound

This paper cites ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.453495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.453495Z digest=sha256:1032622f0791570e5c50835acad280dd5c57453eb1f4b522215996683d9bcf80

Observation b6ed8c87-55d0-47a0-bf9d-004784c97d3f · outbound

This paper cites Legalbench: A collaboratively built benchmark for measuring legal reasoning in large language models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Legalbench: A collaboratively built benchmark for measuring legal reasoning in large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.374226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:30.538974Z digest=sha256:478e1302d0916e259ac7fb043161988bb2f714ed8bd00051a580babc90d4ba9f

Observation 154a610e-0d18-41da-9180-20a211ddd787 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Measuring Mathematical Problem Solving With the MATH Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.670501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.670501Z digest=sha256:b78eae282c64ba48f68ef65e9af27d20c925e92ca68683eae3475d37af075e13

Observation 9d89ec7f-20f6-49b9-93d3-7b1a8ad52b32 · outbound

This paper cites m1: Unleash the potential of test-time scaling for medical reasoning with large language models.arXiv preprint arXiv:2504.00869, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains m1: Unleash the potential of test-time scaling for medical reasoning with large language models.arXiv preprint arXiv:2504.00869, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.723891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.723891Z digest=sha256:09d554969124587d0f251d6d51530a08b0bb118727b9a377c8aaaefb851ff195

Observation 4f91f5fa-e1cf-4bca-b590-e182d22fbc11 · outbound

This paper cites Am-thinking-v1: Advancing the frontier of reasoning at 32b scale.arXiv preprint, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Am-thinking-v1: Advancing the frontier of reasoning at 32b scale.arXiv preprint, 2025

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.129127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:30.837454Z digest=sha256:88c3b73e53b088813eb9ee805cd1c9356ad679687378ceacfa57ea54af9e6e23

Observation 3a4a5e86-3bf4-41c5-8a69-ffc6527d8224 · outbound

This paper cites What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.920949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.920949Z digest=sha256:bf6c3a4e3accffdd8e368231f0c57a0595af5d243918066020a20ebf6eb8708a

Observation e0c3bc71-577a-4196-be57-5c4171e4a8b4 · outbound

This paper cites Cohen, and Xinghua Lu.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Cohen, and Xinghua Lu

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.793900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:31.034678Z digest=sha256:8e4f4d06456d39256186b3fa4714b75bfe007a129ab754329e45a7c59d71670d

Observation cfd7e6fd-cb91-4301-a12b-406b3d170ba5 · outbound

This paper cites Is that your final answer? test-time scaling improves selective question answering.arXiv preprint, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Is that your final answer? test-time scaling improves selective question answering.arXiv preprint, 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.567344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:31.163237Z digest=sha256:f18ff638869a3b940178bf22f23549ce5c88337aca673d6af2b1fca87e0b045e

Observation c3d58d4a-3f9b-41c4-94ad-867c9c85e6f1 · outbound

This paper cites Solving Quantitative Reasoning Problems with Language Models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Solving Quantitative Reasoning Problems with Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.305077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.305077Z digest=sha256:d04e45d859127f3a71407f1aa76fcd38636bb13b5d267770586c266bb0e7afb6

Observation 973ddcb5-f52a-411d-9627-ab765f37e7a9 · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.arXiv preprint, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.arXiv preprint, 2025

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.346201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:31.405173Z digest=sha256:3cbb7b422e0ff6e10083666274f1df2020abab7f8ef7ce99b02e257d86148e45

Observation 8e67493e-4c61-4983-9bd8-a29398e010fc · outbound

This paper cites Decoupled weight decay regularization.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Decoupled weight decay regularization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.516131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.516131Z digest=sha256:2763c3e931ad838eec9703cb337f1426903e5faf6bb235fce2a19d78107a836c

Observation f22498d4-aee1-44fe-bf8b-9d3de22ba4c0 · outbound

This paper cites s1: Simple test-time scaling.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains s1: Simple test-time scaling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.664951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.664951Z digest=sha256:9d2e799108056f33adf0e0cf1b93954c3857b19fff67f13f96dee79ee48ff013

Observation 129ac18c-d3d8-45f7-80ac-398555204060 · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Capabilities of GPT-4 on Medical Challenge Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.762854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.762854Z digest=sha256:3aa8938e367659fe9f5e83d11e9a31cae19b9d4af7e43e3d6339df32ac1856de

Observation 2e38dda6-533d-4637-aca7-88abc95212e8 · outbound

This paper cites GPT-4o System Card.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.805616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.805616Z digest=sha256:3f8645aa77c7ed34d3ae73d1f94f9df5e1347768104dd85e0b4f6d3598502540

Observation 8a23c25b-a1a7-4ae4-bd90-c03ff8cddf42 · outbound

This paper cites Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.875875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.875875Z digest=sha256:1270c0fd4edf248492cd0409dab047a1dbf8b3725645c02ab68f456d878bd61a

Observation b064ae91-7bb8-49f5-a11d-061529955c26 · outbound

This paper cites Llm evaluators recognize and favor their own generations.Advances in Neural Information Processing Systems, 37:68772–68802, 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Llm evaluators recognize and favor their own generations.Advances in Neural Information Processing Systems, 37:68772–68802, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.972846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.972846Z digest=sha256:c2ce3b881303bd1d628d889a20926948a13ad2baba6e250a7719a4301d3278c7

Observation 4d15d7d0-eef1-42aa-a693-9fbaf3263aff · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.114341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.114341Z digest=sha256:ab7807b3d5d1ca60c64f28c2cde3f7c23e06e6bce8795707e49f481dab8d701b

Observation 7672ffa0-45d1-4b64-9b30-7aefa701cd49 · outbound

This paper cites ZeRO: Memory opti- mizations toward training trillion parameter models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ZeRO: Memory opti- mizations toward training trillion parameter models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.087445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:32.283572Z digest=sha256:7f419ce19de532e6c4f9627f8f426b11a92b35b4f1eb648cf3089ad92933ff83

Observation dca7b647-c670-402f-b4b5-45aa900b8a58 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.453949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.453949Z digest=sha256:f3c1f3c7641dea3c08940c4a5c4be0d28deee904371dabc9a8e7e8abee27b80e

Observation 53224f6e-7b4e-4ab6-9285-0f13495cd4a4 · outbound

This paper cites Seed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning.arXiv preprint arXiv:2504.13914, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Seed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning.arXiv preprint arXiv:2504.13914, 2025

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.588627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.588627Z digest=sha256:5fea1fd6bf0ebbb5bc029317bc2a107d49029c66ebe91c7fbefb566c21ec6288

Observation a4c77069-abd4-4c97-85a2-64de080ffde6 · outbound

This paper cites Chain of Logic: Rule-Based Reasoning with Large Language Models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Chain of Logic: Rule-Based Reasoning with Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.686544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.686544Z digest=sha256:8d16a367687357257d50c800d4704248d814081c423ce0d0d3a5f8bbe4234331

Observation 29164d6c-7811-4749-b999-aaf74366e062 · outbound

This paper cites Benchmarking Large Language Models for Math Reasoning Tasks.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Benchmarking Large Language Models for Math Reasoning Tasks

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:33:36.068998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:32.804825Z digest=sha256:ab999955df5f3d797862b61bb3b5cfd63e8655f677cacc3ec6958bf8c3016da0

Observation 68489465-3881-4a23-bcf7-86f6184de09e · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.941847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.941847Z digest=sha256:a88e00f508c159e3ed10fb80ea51b9e9d3e1e6a56cc35b652afc4e9bacbfb02f

Observation fe7f3c56-813c-4a8b-9537-068befa1d671 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Qwen2.5: A party of foundation models, September 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.082416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.082416Z digest=sha256:eb7bf83cf78d5f911bccee2d84cbc8ede980eb25d9dbd0c0d9bbde7db4bd020c

Observation e10cec3f-791d-475d-b1c1-ef57e33a37cd · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.207108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.207108Z digest=sha256:9e7f6604dfb2ae38e4e712eb2a6279db28e0b6560d3f832d2c9aaf7ffda45d25

Observation f7873fcd-0fcd-4537-a77d-6614009a898a · outbound

This paper cites Star-1: Safer alignment of reasoning llms with 1k data.arXiv preprint arXiv:2504.01903, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Star-1: Safer alignment of reasoning llms with 1k data.arXiv preprint arXiv:2504.01903, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.340389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.340389Z digest=sha256:78a097f7232b5901d0d9222b2ef0058f3b5e406c4d79aa655c0670343b3bc6b1

Observation d1177444-0d78-4953-ba15-e2b6c5b46be9 · outbound

This paper cites MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.461626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.461626Z digest=sha256:c5591f46af4bbd6856bd1b47750bfa471a164d8872e58a93cb30195c7129c012

Observation 3ceb2bf4-31f3-4871-9975-592c9fcb1d87 · outbound

This paper cites A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.628956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.628956Z digest=sha256:4e7ccd312262fbaa69bc984ab779c1197e23e3648850ab92550a812e629826cc

Observation c010e386-1bac-4814-9045-3421fe9f1e4c · outbound

This paper cites RCOT: Detecting and Rectifying Factual Inconsistency in Reasoning by Reversing Chain-of-Thought.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains RCOT: Detecting and Rectifying Factual Inconsistency in Reasoning by Reversing Chain-of-Thought

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.766737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.766737Z digest=sha256:d494e0a39b0a180c4ffc9742d2a9aaf4549a2643f61f97dfd0267db30c4eb2ae

Observation 75d3fe04-dd93-43c8-b428-6bd9287a36a2 · outbound

This paper cites Qwen3 technical report, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Qwen3 technical report, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.870910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.870910Z digest=sha256:548c2294675357524956ea9e1c74cbeeec339dafa4016ba386fa6fa64e7cf631

Observation 6a8ad560-8696-4773-9b3c-59f8bbad3830 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.050760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.050760Z digest=sha256:3ce50d53916bc4b97499a0fe3a4686fae8302842a180a8309013b765c34a28fd

Observation 1a7a26cc-7c39-4336-b220-606ad9a09fa3 · outbound

This paper cites LIMO: Less is More for Reasoning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains LIMO: Less is More for Reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.187408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.187408Z digest=sha256:c0f1b5469cd68ef9d15948d2d364b89acd455f10d562d1b2e056daf905841c25

Observation 2334dc10-f27f-4747-8d66-2b915d69ccee · outbound

This paper cites Temporal Data Meets LLM -- Explainable Financial Time Series Forecasting.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Temporal Data Meets LLM -- Explainable Financial Time Series Forecasting

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.324219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.324219Z digest=sha256:0d87504a14d2b5ae3cc9dd6830a7244221dae01e261a50a1137e8f85c87f1add

Observation 831e00fa-25c8-4172-9adb-342062953865 · outbound

This paper cites Online-dpo -r1: Unlocking effective reasoning without the ppo overhead.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Online-dpo -r1: Unlocking effective reasoning without the ppo overhead

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:37.920511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:34.457224Z digest=sha256:72dbd18d03a1d1d0bb041c0325f2d09f81d89943d831370bb7cbdd4f95da6ffa

Observation d864db78-78f2-47c2-8071-ff588464a83b · outbound

This paper cites Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.664997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.664997Z digest=sha256:0a4b265c236a85d3c3da6008583484ab98fbb9f5b6fee5fbd821cad2d16eac81

Observation b88cc0e5-d93c-4fef-8a9f-90ed663e0d5e · outbound

This paper cites MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.809777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.809777Z digest=sha256:344eb52f7f8ae9d0fcfca396e0c053f3c5c780d4bbbc4a2c213fb19f35b35d9d

Observation 1eb70941-d6c5-487c-83f2-b1617547ebf1 · outbound

This paper cites planning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains planning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:37.477016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:35.103026Z digest=sha256:a4450e7493a4ae4aea1b759571b12e2974833a486d2e9d71beb3fba7fcf70ab2

Observation 510d5c4e-42c2-4842-86a6-3be1b0e87f60 · outbound

This paper cites action" : a detailed description of the actions taken in this step.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains action" : a detailed description of the actions taken in this step

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:37.238152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:35.268957Z digest=sha256:c92e8e20f3c830725f9dcde45298982570af2b4dd9114d1c94479e2c32a3af82

Observation 8588112d-18d6-40ce-8058-5b88a17171bb · outbound

This paper cites step_text.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains step_text

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:36.969624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:35.376543Z digest=sha256:094e7aa2fb0f6dea56e32b46e69fdf955818b2d04ab1c0581b3b5b3a1bad5499

Observation 1b09d411-9ca0-46ce-895d-558d2821a2c7 · outbound

This paper cites an unresolved cited work.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:37.691645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:35.471267Z digest=sha256:5ce01639b5d4fd083461b0447fef7fc06885e2592f285a162a8ecba4651c6ea3

Observation a775b924-3f46-4e69-afa1-f5ee4d3a0e47 · outbound

This paper cites query" :.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains query" :

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:36.727622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:35.587634Z digest=sha256:7659f0c132ed0a6fb315f3ba75e942ce8d67effaecfd4a2d3d39c0dcb814593f

Pith citing papers

Observation ebf9708b-2714-4238-8cc7-40dfcc9b8769 · inbound

Kernel-Based Sparse Additive Nonlinear Model Structure Detection through a Linearization Approach cites this paper.

Kernel-Based Sparse Additive Nonlinear Model Structure Detection through a Linearization Approach Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T05:37:35.270920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:37:35.270920Z digest=sha256:dac5dd82c3bab68a1d54328c2a2a19a93e2dd17778a604e14ee5b32d1b443f41

Observation 71114f53-b738-47e8-967a-9c4fa12edf78 · inbound

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities cites this paper.

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T22:58:14.894642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:58:14.894642Z digest=sha256:80f717c433db45cf2e7d553b48021b20c814903a7d5b390daade3ea4e830755a

Observation 62c8e7d5-0d5b-49ca-8272-d8e6e32ac7ce · inbound

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models cites this paper.

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:21:08.792769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T17:20:19.586214Z digest=sha256:179391ecf6c408615ce505eb29b418b83c2953e1224a3637111b1721eccf6b8c

Observation f2fd2301-ed59-4121-89c4-612dc38c3841 · inbound

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning cites this paper.

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:23:03.823103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T05:19:35.222068Z digest=sha256:c2477f140bc0b915e1babf5513090bf47ffac3dc646a50fb26bfe4830a2e474b

Observation b108c978-aaa0-4e15-af7e-d170124bcbb1 · inbound

LLM Parameters for Math Across Languages: Shared or Separate? cites this paper.

LLM Parameters for Math Across Languages: Shared or Separate? Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:28:59.011429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T00:30:30.315423Z digest=sha256:50165e9b0c59eb686caa5e76a18f530ca741f6ed9e3bf3082acee03fdd16921a

Observation 23e2a58c-0da1-4563-b5bb-ca9237ff6cd3 · inbound

CausalMix: Data Mixture as Causal Inference for Language Model Training cites this paper.

CausalMix: Data Mixture as Causal Inference for Language Model Training Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:47:05.651702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-02T15:45:02.415577Z digest=sha256:d43f0076057063f5bfa076562e52935037da75ee877e5122f3a41076e672a053