Pith. sign in

Paper Citation Record · LEDGER

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

As of 17 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 6 inbound Pith citation observations for arXiv:2506.02126.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02126 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:35.587634Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:37:35.270920Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:59.009997Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 51c71118-a503-411f-8860-e4a5eddd26f9 · outbound

This paper cites https:// artofproblemsolving.com/wiki/index.php/AIME_Problems_and_Solutions, 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains https:// artofproblemsolving.com/wiki/index.php/AIME_Problems_and_Solutions, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:41.269022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:29.013985Z digest=sha256:5852a8462c1f302c4a16f9f30d38466f6d62aefb7ae9c8daf8dea9482d6bf743

Observation f07810c3-0ba2-414b-ad76-d559f4075738 · outbound

This paper cites https://www.vals.ai/benchmarks/ math500-03-24-2025, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains https://www.vals.ai/benchmarks/ math500-03-24-2025, 2025

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.961493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:29.108346Z digest=sha256:353bad7d5ef255f5eed71af7c403bd3ab1f32af2be25d0ee025ae5f58af43a81

Observation 012733b1-6bf0-482e-bc68-8b9ce1f518c2 · outbound

This paper cites https://www.maa.org/math-competitions, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains https://www.maa.org/math-competitions, 2025

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.720342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:29.222146Z digest=sha256:382a582dfbd85ad7e6e127e4dd099a2db15be59fe6cd7eb4e90550f09eb1da8e

Observation e033fb29-e84b-4ecb-aa90-7e63ba97413c · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.341067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.341067Z digest=sha256:2cd44d1750525af3b443f9653dceb0b7758954b3ec22621dd6595a076d33d227

Observation a381e39a-81c4-4b8d-b03a-f5c4c46c26f0 · outbound

This paper cites Large Language Models for Mathematical Reasoning: Progresses and Challenges.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.433062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.433062Z digest=sha256:81513333e0a33f1e50dc1acaa613dad0bd61facb491f025106563c25e025d4bf

Observation d6c9ac2b-9ca5-49c2-bdce-fda723483429 · outbound

This paper cites Llama-nemotron: Efficient reasoning models.arXiv preprint arXiv:2505.00949, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Llama-nemotron: Efficient reasoning models.arXiv preprint arXiv:2505.00949, 2025

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.562739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.562739Z digest=sha256:ae766efb9629513bbc6ed8385c25ec4a35c3cba93aeee7566cef6576a717972d

Observation e55064ba-3086-4225-919a-7f1cec42a7f6 · outbound

This paper cites Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, andet al.Language models are few-shot learners.Advances in Neural Information Processing Systems, 33:1877–1901, 2020.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, andet al.Language models are few-shot learners.Advances in Neural Information Processing Systems, 33:1877–1901, 2020

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.377366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:29.713207Z digest=sha256:5a8f57cd409ef4493c227f1662e6dd494ef6969f27b282a6eb52ff15c55d22b2

Observation e7f0ef46-a258-4dfc-9f35-5d18dd21565a · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.797685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.797685Z digest=sha256:6458acfb3d85cc9ded4376f530b77b1a0981ea52f5ab8da4aec856be112a2bcc

Observation 8187428a-5f7f-447f-8618-a5a8e655858e · outbound

This paper cites HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.903235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.903235Z digest=sha256:ebd5f82e04b53edc73c8e295833fef6a85d72d608e64aae0504252690c402384

Observation 12be5eea-c6d1-4762-8df6-7246e735a1c7 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.050544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.050544Z digest=sha256:98457d61e4638602cca7705053cc295cbed5e41b5ab4b027e8e57c76786d65e0

Observation 20a7ff3a-fdcd-415c-bd86-c94e2f25f8a9 · outbound

This paper cites Deepseek -r1-distill-qwen-7b.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Deepseek -r1-distill-qwen-7b

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:40.134256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:30.146998Z digest=sha256:cd0042ad80eef011747ddefa1843e923737aad259a207ce495bb4facb7d0b2d6

Observation f9f5d7ca-ffdd-40ea-90e8-1daa4de7e16f · outbound

This paper cites Rlhf workflow: From reward modeling to online rlhf.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Rlhf workflow: From reward modeling to online rlhf

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.823339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:30.259099Z digest=sha256:5fec3d915874a57e59e9ab88102289b62f55d3334cbf9d559a61f1587d498d44

Observation 587cfbef-66d5-486e-9b5b-5dd04b154ec3 · outbound

This paper cites Large language model influence on diagnostic reasoning: a randomized clinical trial.JAMA Network Open, 7(10):e2440969– e2440969, 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Large language model influence on diagnostic reasoning: a randomized clinical trial.JAMA Network Open, 7(10):e2440969– e2440969, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.604337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:30.368609Z digest=sha256:90743937bdfc81e2acee8bed3d60bd930111bebd6ac4a23c75ee70a9bcbdb366

Observation fe22c061-ad7b-4ba9-9e23-205a1610624f · outbound

This paper cites ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.453495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.453495Z digest=sha256:36595dd7fd75da66a0e8869499fe193a3dff50b622c265569429cb1b8a3ca898

Observation b6ed8c87-55d0-47a0-bf9d-004784c97d3f · outbound

This paper cites Legalbench: A collaboratively built benchmark for measuring legal reasoning in large language models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Legalbench: A collaboratively built benchmark for measuring legal reasoning in large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.374226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:30.538974Z digest=sha256:1e1e4ad7b1bfb9357317565b62190602c0a23a5132631a3cb05989448c51ec5f

Observation 154a610e-0d18-41da-9180-20a211ddd787 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Measuring Mathematical Problem Solving With the MATH Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.670501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.670501Z digest=sha256:0e6028c5d7f2b17eaea1157f020f5ef439e6a815352866a8ce05651ee130849e

Observation 9d89ec7f-20f6-49b9-93d3-7b1a8ad52b32 · outbound

This paper cites m1: Unleash the potential of test-time scaling for medical reasoning with large language models.arXiv preprint arXiv:2504.00869, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains m1: Unleash the potential of test-time scaling for medical reasoning with large language models.arXiv preprint arXiv:2504.00869, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.723891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.723891Z digest=sha256:03531bc7ee1ec277e67f852e35a11d35f148276e05c1232f2eb1be75cedd6b74

Observation 4f91f5fa-e1cf-4bca-b590-e182d22fbc11 · outbound

This paper cites Am-thinking-v1: Advancing the frontier of reasoning at 32b scale.arXiv preprint, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Am-thinking-v1: Advancing the frontier of reasoning at 32b scale.arXiv preprint, 2025

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:39.129127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:30.837454Z digest=sha256:15720c4afc53d1e5c16f727c95e8659629bd3d08f24ce48cf7f1a907d2cefcd7

Observation 3a4a5e86-3bf4-41c5-8a69-ffc6527d8224 · outbound

This paper cites What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:30.920949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:30.920949Z digest=sha256:a12ff37bdadcdef8ccea2f9909b0cb165af11af63f537260a871d99e91917a4f

Observation e0c3bc71-577a-4196-be57-5c4171e4a8b4 · outbound

This paper cites Cohen, and Xinghua Lu.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Cohen, and Xinghua Lu

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.793900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:31.034678Z digest=sha256:93fa55cea045ac86daac7ca38ea9e60ff9ec432c0be848ae8ba736e763a51fdd

Observation cfd7e6fd-cb91-4301-a12b-406b3d170ba5 · outbound

This paper cites Is that your final answer? test-time scaling improves selective question answering.arXiv preprint, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Is that your final answer? test-time scaling improves selective question answering.arXiv preprint, 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.567344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:31.163237Z digest=sha256:d86ba5c6bc5743a9ab8fd037bb99d7b9613d32c853641a8a3413d57ed7d9716c

Observation c3d58d4a-3f9b-41c4-94ad-867c9c85e6f1 · outbound

This paper cites Solving Quantitative Reasoning Problems with Language Models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Solving Quantitative Reasoning Problems with Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.305077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.305077Z digest=sha256:80f4b716ad77cadc48ebf8db4d2886c873b8523cd1cea976efdadf8074edc204

Observation 973ddcb5-f52a-411d-9627-ab765f37e7a9 · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.arXiv preprint, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.arXiv preprint, 2025

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.346201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:31.405173Z digest=sha256:78ed10887540b41f68524d3e1f10cea9b1dae1b2d9a9d5749c77022f52cf40cf

Observation 8e67493e-4c61-4983-9bd8-a29398e010fc · outbound

This paper cites Decoupled weight decay regularization.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Decoupled weight decay regularization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.516131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.516131Z digest=sha256:65aa46ddeb974b0d5978f8b83f295ae221934efdc00de6df1626e55850243da6

Observation f22498d4-aee1-44fe-bf8b-9d3de22ba4c0 · outbound

This paper cites s1: Simple test-time scaling.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains s1: Simple test-time scaling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.664951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.664951Z digest=sha256:505e26e15f91dcf5362f657f7879b2b0d436e441571d5eebaa09ef27c55b035e

Observation 129ac18c-d3d8-45f7-80ac-398555204060 · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Capabilities of GPT-4 on Medical Challenge Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.762854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.762854Z digest=sha256:620aa5e1d3cf81bd5af584d6297475ae60714847417001f690038a6193c27823

Observation 2e38dda6-533d-4637-aca7-88abc95212e8 · outbound

This paper cites GPT-4o System Card.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.805616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.805616Z digest=sha256:3e49023513345e60571bb9645bf4c35e61940763b71294cb5233c26a77d634af

Observation 8a23c25b-a1a7-4ae4-bd90-c03ff8cddf42 · outbound

This paper cites Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.875875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.875875Z digest=sha256:baa6266da3485d46787fc2e035edd227d472d3ed02a8350501a1afc44c51f4af

Observation b064ae91-7bb8-49f5-a11d-061529955c26 · outbound

This paper cites Llm evaluators recognize and favor their own generations.Advances in Neural Information Processing Systems, 37:68772–68802, 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Llm evaluators recognize and favor their own generations.Advances in Neural Information Processing Systems, 37:68772–68802, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:31.972846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:31.972846Z digest=sha256:c2c8f827cab3ca6e8b8fdc60a1fbc03eac2be4e6be1c136873df2b344bb879d3

Observation 4d15d7d0-eef1-42aa-a693-9fbaf3263aff · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.114341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.114341Z digest=sha256:59f3ad6dfe8d459453524f3290e26ca9d2e5640bfcae581959af6aef445d5a82

Observation 7672ffa0-45d1-4b64-9b30-7aefa701cd49 · outbound

This paper cites ZeRO: Memory opti- mizations toward training trillion parameter models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ZeRO: Memory opti- mizations toward training trillion parameter models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:38.087445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:32.283572Z digest=sha256:693bb98522620d6555a09b5747a25af9de6196bd3e9e4cef9ed7e47c6bb589e2

Observation dca7b647-c670-402f-b4b5-45aa900b8a58 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.453949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.453949Z digest=sha256:a13bb175936d33aecf1f371277e51af155290415bc565f6c9fda8e4be9309794

Observation 53224f6e-7b4e-4ab6-9285-0f13495cd4a4 · outbound

This paper cites Seed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning.arXiv preprint arXiv:2504.13914, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Seed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning.arXiv preprint arXiv:2504.13914, 2025

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.588627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.588627Z digest=sha256:66dc0634a9e164b04fa64a6bdeca29b9e67655ede38a4e8978ce8a1e432b3740

Observation a4c77069-abd4-4c97-85a2-64de080ffde6 · outbound

This paper cites Chain of Logic: Rule-Based Reasoning with Large Language Models.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Chain of Logic: Rule-Based Reasoning with Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.686544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.686544Z digest=sha256:55d746f455b7b5110d25eae9c32bbf036dd89914210143f8b1a7cea3b7a7a835

Observation 29164d6c-7811-4749-b999-aaf74366e062 · outbound

This paper cites Benchmarking Large Language Models for Math Reasoning Tasks.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Benchmarking Large Language Models for Math Reasoning Tasks

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:33:36.068998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:32.804825Z digest=sha256:d3f19b6ed6ddcb464723cd15f492fd079bc2d771ff2c6306287c9b8ae1246160

Observation 68489465-3881-4a23-bcf7-86f6184de09e · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.941847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.941847Z digest=sha256:bed28dc46b2c1e31bb0740feb60b55d8fb8d10f0d3de1fff7cd4aa5a1516c7e9

Observation fe7f3c56-813c-4a8b-9537-068befa1d671 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Qwen2.5: A party of foundation models, September 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.082416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.082416Z digest=sha256:fa12b84d2891d13aba39c93d189dea522cefdf4575cf24c1689be95237593bf8

Observation e10cec3f-791d-475d-b1c1-ef57e33a37cd · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.207108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.207108Z digest=sha256:f52d0221b5d107baf906de41e4fbe2f8204b8a71e87323af46995278bde1cd72

Observation f7873fcd-0fcd-4537-a77d-6614009a898a · outbound

This paper cites Star-1: Safer alignment of reasoning llms with 1k data.arXiv preprint arXiv:2504.01903, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Star-1: Safer alignment of reasoning llms with 1k data.arXiv preprint arXiv:2504.01903, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.340389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.340389Z digest=sha256:b51420e884efc63fa3b631a063fe93d0c887a726fc2200f3622ee5ac810e5bb5

Observation d1177444-0d78-4953-ba15-e2b6c5b46be9 · outbound

This paper cites MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.461626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.461626Z digest=sha256:77fa305ab76f3cbea542e95e439e3c5d9133798cffe8daf98155ea4901b26497

Observation 3ceb2bf4-31f3-4871-9975-592c9fcb1d87 · outbound

This paper cites A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.628956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.628956Z digest=sha256:4229a64a14465da5e9bc334ae48250f9c5b16973cf1a6ad432b8577f966e27cf

Observation c010e386-1bac-4814-9045-3421fe9f1e4c · outbound

This paper cites RCOT: Detecting and Rectifying Factual Inconsistency in Reasoning by Reversing Chain-of-Thought.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains RCOT: Detecting and Rectifying Factual Inconsistency in Reasoning by Reversing Chain-of-Thought

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.766737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.766737Z digest=sha256:d95fd724a5b0b3e6805ec2aef6b5bcb1b5b7922fbb3904a85bc05bb5c5eae7b9

Observation 75d3fe04-dd93-43c8-b428-6bd9287a36a2 · outbound

This paper cites Qwen3 technical report, 2025.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Qwen3 technical report, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:33.870910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:33.870910Z digest=sha256:8d6f898073a4ff175efb629b52d9cbbfb34e9a213fcf067ecfa89c970a1bd75c

Observation 6a8ad560-8696-4773-9b3c-59f8bbad3830 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.050760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.050760Z digest=sha256:742296628a5d3791911eba9365046385aad2751bcc7c9a9c55d4c9a48c9bc020

Observation 1a7a26cc-7c39-4336-b220-606ad9a09fa3 · outbound

This paper cites LIMO: Less is More for Reasoning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains LIMO: Less is More for Reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.187408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.187408Z digest=sha256:b44c72a598df6c9319955db6089c18a9639f2b6fdfaba31e8825a9125f043082

Observation 2334dc10-f27f-4747-8d66-2b915d69ccee · outbound

This paper cites Temporal Data Meets LLM -- Explainable Financial Time Series Forecasting.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Temporal Data Meets LLM -- Explainable Financial Time Series Forecasting

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.324219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.324219Z digest=sha256:d396ccb6790f5c5ef7d07269119726166fba5b4a6f61dd8d9bb8b7ce46fa3dfc

Observation 831e00fa-25c8-4172-9adb-342062953865 · outbound

This paper cites Online-dpo -r1: Unlocking effective reasoning without the ppo overhead.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Online-dpo -r1: Unlocking effective reasoning without the ppo overhead

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:37.920511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:34.457224Z digest=sha256:6b31a960ae847872f0720db16379d12c76840386aab52e0f53725c0f27f6515f

Observation d864db78-78f2-47c2-8071-ff588464a83b · outbound

This paper cites Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.664997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.664997Z digest=sha256:4fcfd881c145a194aae3f84119ab32814d43bd46b366d781ef0aef27dd4e4117

Observation b88cc0e5-d93c-4fef-8a9f-90ed663e0d5e · outbound

This paper cites MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:34.809777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:34.809777Z digest=sha256:b6a7cad7ae5c898f74a0e830b36989daae08bc0d257dfae65dcb818118c88742

Observation 1eb70941-d6c5-487c-83f2-b1617547ebf1 · outbound

This paper cites planning.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains planning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:37.477016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:35.103026Z digest=sha256:34facd538afc36c91d5a47d803633595616e66de39d49c3228554cf6b0bc2a45

Observation 510d5c4e-42c2-4842-86a6-3be1b0e87f60 · outbound

This paper cites action" : a detailed description of the actions taken in this step.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains action" : a detailed description of the actions taken in this step

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:37.238152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:35.268957Z digest=sha256:0ae7239501bd9e44daf124f3e1b7c839c3d4adcd829b516b3b0b7376f703a38d

Observation 8588112d-18d6-40ce-8058-5b88a17171bb · outbound

This paper cites step_text.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains step_text

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:36.969624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:35.376543Z digest=sha256:d6d0f5257c279fac9debab65d59c48fd6fbbeaca19742031f0e34dd07fcf5bcf

Observation 1b09d411-9ca0-46ce-895d-558d2821a2c7 · outbound

This paper cites an unresolved cited work.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:37.691645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:35.471267Z digest=sha256:a81bd111409a2fac927e6777432328aea34573b5e157f0a2f51304ce9a440ee0

Observation a775b924-3f46-4e69-afa1-f5ee4d3a0e47 · outbound

This paper cites query" :.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains query" :

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:36.727622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T11:33:35.587634Z digest=sha256:cee5659cba0e8879f027089c1a999295cca7fa39f791c2b1ea50d3785e48cd5b

Pith citing papers

Observation ebf9708b-2714-4238-8cc7-40dfcc9b8769 · inbound

Kernel-Based Sparse Additive Nonlinear Model Structure Detection through a Linearization Approach cites this paper.

Kernel-Based Sparse Additive Nonlinear Model Structure Detection through a Linearization Approach Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T05:37:35.270920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:37:35.270920Z digest=sha256:a201538bc5f10f9c3ac89ab759673d221f3d73d38750f6b5be9707d5cd76b845

Observation 71114f53-b738-47e8-967a-9c4fa12edf78 · inbound

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities cites this paper.

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T22:58:14.894642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:58:14.894642Z digest=sha256:d2422b34d244b856030c3c8bc70162d3e0ed73f248077cf30922896e8e40bdf8

Observation 62c8e7d5-0d5b-49ca-8272-d8e6e32ac7ce · inbound

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models cites this paper.

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:21:08.792769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T17:20:19.586214Z digest=sha256:d7be505dacf0fb6b761d6668c6477f18cdfc8cc28ebf655aa6fdc493ef95fd4c

Observation f2fd2301-ed59-4121-89c4-612dc38c3841 · inbound

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning cites this paper.

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:23:03.823103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T05:19:35.222068Z digest=sha256:76684e7ef371f72b0d00964f0293e2279b2a86208096b6b84e3e16b5e2f8a52a

Observation b108c978-aaa0-4e15-af7e-d170124bcbb1 · inbound

LLM Parameters for Math Across Languages: Shared or Separate? cites this paper.

LLM Parameters for Math Across Languages: Shared or Separate? Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:28:59.011429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T00:30:30.315423Z digest=sha256:35cae5f9ef0805af427b3632aac8c1454cec3675c128882ddf70cf5a9579759f

Observation 23e2a58c-0da1-4563-b5bb-ca9237ff6cd3 · inbound

CausalMix: Data Mixture as Causal Inference for Language Model Training cites this paper.

CausalMix: Data Mixture as Causal Inference for Language Model Training Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:47:05.651702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-07-02T15:45:02.415577Z digest=sha256:221294834f114cc0c52093e3db303da76885ffc018c96b844d0eaef3e427a87a