Pith. sign in

Paper Citation Record · LEDGER

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

As of 18 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 3 inbound Pith citation observations for arXiv:2504.16414.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16414 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:08:27.681538Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T07:36:16.071678Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 5c5e6c69-650d-4882-aaa5-87e6b52bbe80 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Chain-of-thought prompting elicits reasoning in large language models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.033380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.033380Z digest=sha256:24731e8dcd0f81885d58965f2df1221e97b7bf5bc85f2482f0e84cfb36eba581

Observation 786957e3-11ee-4bd9-8cf6-b17ab453974b · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.041280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.041280Z digest=sha256:ecc84b8eb30ae63525a812f68f177152c31a33b0589411e595e0e83fb42f9b40

Observation 944effd2-5832-4a98-a42d-1168409c548c · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Tree of thoughts: Deliberate problem solving with large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.047544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.047544Z digest=sha256:ecbe954d668c1d533343ba2015f9934b508badd8adfcb77f1bf9aac704a64ad8

Observation 32f29f2b-c9a0-4b7c-921d-153931ba86ab · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Graph of thoughts: Solving elaborate problems with large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.694755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.055284Z digest=sha256:5612f3b2bc499dc7df483a44b169e8ec7f8035261c580ef94043fa1ea4655a02

Observation 13d59eb0-efc0-4df1-b3e2-b9aa560dd2f3 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.062988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.062988Z digest=sha256:37f3d17fe90a0f709c1e61b2ea8d1836c75c12c3ebec9f065d2fc6da0df4064e

Observation a7b19a6c-a234-4bdc-ae53-5b648bc62361 · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Retrieval-augmented generation for knowledge-intensive nlp tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.078801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.078801Z digest=sha256:3a80d5da6629c6017836d5da9e97e6564b6a6f899fd308e986f35cc1abaca2b3

Observation e6a74663-37b9-4855-9319-3a2c3aa1ec46 · outbound

This paper cites Neurosymbolic ai: the 3rd wave.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Neurosymbolic ai: the 3rd wave

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.453558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.086596Z digest=sha256:7aa3e3343d820f660423e5e5c5239171f4cf9c1c216144fc248319790313720f

Observation 9d08a5c9-624c-448b-99f9-32434f31f4a2 · outbound

This paper cites A simple neural network module for relational reasoning.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study A simple neural network module for relational reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.093328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.093328Z digest=sha256:03be56d8caf5c436e41da993a0b627e60f1feb345dfa1a3c8b2f311a0d4250c2

Observation 97572626-9b5c-4b01-8063-0926d91a566f · outbound

This paper cites Openai o1 system card, 2024.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Openai o1 system card, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.352691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.101824Z digest=sha256:741c7dd16926226d0b61f69ea2a5d1014d41334a31a93329f7348605bb2825fe

Observation 6cc56a61-3388-4dce-a4a6-828a2c935406 · outbound

This paper cites Openai o3 mini system card, 2024.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Openai o3 mini system card, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.256718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.111018Z digest=sha256:b88423ad4aba28965798f2eee544fe3ff45de774bccaf8cab6950dd629426ad6

Observation 66bfd8a0-6bde-405b-a898-8e781124fff6 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Star: Bootstrapping reasoning with reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.118353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.118353Z digest=sha256:20647f94c2509f8d3921c2cc8c18054fd9184c6b749f303f0331b4fe67304872

Observation 3fa5a002-6426-4f1f-831d-0a6aaf77bba8 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.125584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.125584Z digest=sha256:4d3addee05f1b1eff787ebfa7d36ed4c86b748dc8c140d98a56813447c4d10c5

Observation 29e29159-c387-401a-ac6e-a105ab72a4c6 · outbound

This paper cites Advancing Reasoning in Large Language Models: Promising Methods and Approaches.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Advancing Reasoning in Large Language Models: Promising Methods and Approaches

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.135194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.135194Z digest=sha256:88fc59691ce2cb4b20ab98e6eb7ec05eceb61d9bc2012e6dfa71254210e883be

Observation b230e105-f40e-4ff2-841c-8136172dd967 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.145988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.145988Z digest=sha256:4285164f2d8881b4ed916bef32160d091a4a6c2d4d5705cf4d0a6dad4445b3bc

Observation 24e21779-dfdd-4895-b1b4-90719cd3759b · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Measuring Mathematical Problem Solving With the MATH Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.152376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.152376Z digest=sha256:69294d4460c1660d572a80cf205cdce2ceffb5d343821abaadd78bdf7eda3f28

Observation 82536658-847e-4cb7-b819-fb84870c44b8 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Evaluating Large Language Models Trained on Code

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.160516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.160516Z digest=sha256:f7cff76eaacf8010ed8407288c088e60f253751d0c58dde71a139022624af0d7

Observation 9204ca91-0a5e-4a88-8c70-8adb6cb214de · outbound

This paper cites Program Synthesis with Large Language Models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Program Synthesis with Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.165982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.165982Z digest=sha256:c1e4dc682d7f061e38a5ffd964d3dc8194539a5214fb64f9ed97791312f031eb

Observation bf725a04-e15f-44d4-a024-f57cc3f71d37 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.176304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.176304Z digest=sha256:770467f619ada39281f3f2fb6d9190304b6a47c47fad0265ebedc91bea1297fc

Observation 740f52cd-73f3-4f59-a006-42bdbc301f89 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.183147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.183147Z digest=sha256:438b33d574c38e550a534d2ae8a09d3d7c7e641da831e06a72ef775126967269

Observation 3a88ce35-5293-4571-ab1c-c47953128198 · outbound

This paper cites Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.198025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.198025Z digest=sha256:a928275cdb9c9327af31c5171e6322837b2c28d525a4cdee0e39599fc92d35fa

Observation 06b0d660-6999-4e5d-bc64-5b22630fae49 · outbound

This paper cites Chemlit-qa: A human evaluated dataset for chemistry rag tasks.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Chemlit-qa: A human evaluated dataset for chemistry rag tasks

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.156864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.205317Z digest=sha256:e61f4ac31e174e68f6d3ca07ab910cfa6b5b4045a3876f0e3ec61c050b94c0c2

Observation e907d867-c733-4885-97e2-288143e5a7a0 · outbound

This paper cites OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far?.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.211322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.211322Z digest=sha256:623f6b58c1ff105e3df067c319596201d37286c40a1e08923055dfd5d144cdd2

Observation d7f7b64c-6296-4cbc-bb26-b4ab59c56f40 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Gpqa: A graduate-level google-proof q&a benchmark

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.228224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.228224Z digest=sha256:9b48d4975491139c132a1c434a2b960b7af2d343ffc2d1d2ab8024f22f2b0532

Observation e8236d33-6d50-4e29-9cec-efee71ba2911 · outbound

This paper cites Large language models for reticular chemistry.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Large language models for reticular chemistry

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.002936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.234683Z digest=sha256:8aa6f95b693d959c93c03b0bfc7d1a17b15c042a23f9b3bb5450a775030efa68

Observation 36a26e47-c694-4748-9e8d-ca759c55071e · outbound

This paper cites Multi-hop question answering.Foundations and Trends® in Information Retrieval, 17(5):457–586, 2024.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Multi-hop question answering.Foundations and Trends® in Information Retrieval, 17(5):457–586, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.907012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.241942Z digest=sha256:5a636b47510349ffb7246a2428138b11edad87eac0c1e4d0dc36732232f988a7

Observation b5031814-288e-4a46-a47c-b1f0504b416d · outbound

This paper cites Constructing datasets for multi-hop reading comprehension across documents.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Constructing datasets for multi-hop reading comprehension across documents

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.862076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.249274Z digest=sha256:44cb98568249872dfcb39b8c3651d3520a6e9f0df556805df7cd32bf11f96e77

Observation 9d0f70dd-d8b9-4988-9160-e4680425be07 · outbound

This paper cites Musique: Multihop questions via single-hop question composition.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Musique: Multihop questions via single-hop question composition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.255231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.255231Z digest=sha256:a13a4b835cab1a9af47b7e83710a8990d5cf6c62d1fb35d698cd965f4fb42707

Observation dbc7d224-0086-49fc-94a0-00d046a57c52 · outbound

This paper cites MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.261607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.261607Z digest=sha256:d4b34717d3845ec940c61953df5f8c836b12d339c6c9499efb2c4d7068dbf052

Observation 3e491458-4e77-48f0-840e-19e2a3c7e46a · outbound

This paper cites A framework for evaluating the chemical knowledge and reasoning abilities of large language models against the expertise of chemists.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study A framework for evaluating the chemical knowledge and reasoning abilities of large language models against the expertise of chemists

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.702468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.269213Z digest=sha256:d25d73249e38311528ba44c52588ac7817bae86ebeeaec173c703f3301ee6fca

Observation 8ccf9fa0-2b60-46a4-ae87-028672acace6 · outbound

This paper cites Knowledge Graph Generation From Text.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Knowledge Graph Generation From Text

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.275286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.275286Z digest=sha256:5b1ce5a5168ab806c6c9e7e0559fbca662122ecaedcb80f28c0eb6fd02f29ad7

Observation 52d9b364-0a8c-4870-afbd-052a4e0e89e6 · outbound

This paper cites Extract, Define, Canonicalize: An LLM-based Framework for Knowledge Graph Construction.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Extract, Define, Canonicalize: An LLM-based Framework for Knowledge Graph Construction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.280841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.280841Z digest=sha256:097ca54083c5902e38787e83fe71741b44511c5131c5ec92e9709cf53bb8cbb8

Observation 0c10439b-070b-4430-866b-dc528a1dd822 · outbound

This paper cites Building Dynamic Knowledge Graphs from Text using Machine Reading Comprehension.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Building Dynamic Knowledge Graphs from Text using Machine Reading Comprehension

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.287652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.287652Z digest=sha256:8193cd69de07047269526461d9d0dbf2d977be5e268f34de0c78757ae9ae35bd

Observation ba8d481d-d409-4b24-8b71-6407539fcf87 · outbound

This paper cites CEAR: Automatic construction of a knowledge graph of chemical entities and roles from scientific literature.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study CEAR: Automatic construction of a knowledge graph of chemical entities and roles from scientific literature

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.295596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.295596Z digest=sha256:0781914806be64ac9077152630fc32dba51cc22ed6ff32ae275cde5b5700be3b

Observation a1ac075f-247c-4ed7-bd18-ba773281e3b5 · outbound

This paper cites Coarse-to-fine knowledge graph domain adaptation based on distantly-supervised iterative training.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Coarse-to-fine knowledge graph domain adaptation based on distantly-supervised iterative training

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.614194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.301310Z digest=sha256:57ca774f19933c984468286a45b84669532e2eb5c7824fb1e640c45925f8fdd7

Observation 5576a2ae-05d5-4cc1-b051-9616aaf09c2f · outbound

This paper cites Nilinker: attention-based approach to nil entity linking.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Nilinker: attention-based approach to nil entity linking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.537106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.307601Z digest=sha256:bf8d5d192614e4c5df597d1fba0e07b64ecd7b61840823205b6868601d1053e6

Observation da72f682-ae47-4430-aec0-92ac7fb792b6 · outbound

This paper cites Domain-specific language model pretraining for biomedical natural language processing, 2020.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Domain-specific language model pretraining for biomedical natural language processing, 2020

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.313344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.313344Z digest=sha256:167bc9923106ecc2618e74da443b6f0b39feac691826150698b4a8116652a066

Observation 7fb675f9-8468-4f99-8af4-6e5a5aa47793 · outbound

This paper cites Pubchem in 2021: new data content and improved web interfaces.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Pubchem in 2021: new data content and improved web interfaces

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.372348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.318415Z digest=sha256:bd530553b2cfedfbdee190622802db6383203691657676a937cae78b5b79200b

Observation 2436015d-32bf-4163-bba8-422f32711f77 · outbound

This paper cites Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.323673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.323673Z digest=sha256:5a2214c938203e94d057252a44fb5a3bd4209c1ab743d4d407328bde78f8315d

Observation 32c19761-3314-44f5-9941-174a5c16880e · outbound

This paper cites Hierarchical Retrieval-Augmented Generation Model with Rethink for Multi-hop Question Answering.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Hierarchical Retrieval-Augmented Generation Model with Rethink for Multi-hop Question Answering

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.341341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.341341Z digest=sha256:bf5cd8873ddb86c21e1e9b6dd8d805fe8a72f23846529799da05a244f54c17db

Observation 421bf404-e7fa-49fc-851e-293c94ef9bd2 · outbound

This paper cites HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.361060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.361060Z digest=sha256:3a8f72527e6aa17a4837e52617069d86aa326d9d4e6eef4c20e8faeb5513fc27

Observation a1a2e9e7-84bb-4fd1-993c-d390a5c8f5d1 · outbound

This paper cites If an entity appears in the text but has no meaningful chemical relationship with another entity in the set, ignore it.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study If an entity appears in the text but has no meaningful chemical relationship with another entity in the set, ignore it

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.292668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.379975Z digest=sha256:c78bd2dd0cf9199182991f40c25cb0d39ccd766831cce7dc44c0f2748fc0a8ac

Observation f1c16554-379c-4678-a088-4ba9c7626a5b · outbound

This paper cites reacts with,.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study reacts with,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.177310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.385205Z digest=sha256:0113e946bb0f36e3360fe39dc79272050859ef0fbfd484f65c01b31722768de9

Observation 798ee359-8162-46fb-bcad-67cfede7a0ae · outbound

This paper cites Avoid observations, opinions, and findings.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Avoid observations, opinions, and findings

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.133824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.391936Z digest=sha256:6fe70259fdef9884bc4cfec962b57cee9f04b4bf5f84722ebdbd50cf2fdf2c2c

Observation 43c2180b-0fd6-48b7-b3cf-fe615e51a8b6 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:29.013438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.413495Z digest=sha256:5a6695921b0f0731a898a2ff44f07089fd0fc5c22f13bdba4e0ffb19a0240f2d

Observation ce6a1cbe-346f-4bc0-b870-e256db6e8ae0 · outbound

This paper cites is," "are,.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study is," "are,

Reference 45

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T11:08:28.936126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.457183Z digest=sha256:dae38b2c78ed3bdddecaa27aeed5fc152bce273ab05f2851fdb0bd71cd87c8dc

Observation 92e7b5b1-7d25-4811-bb54-ede837e6e937 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.830681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.464119Z digest=sha256:cbd5fdc30e55caa8b790b0df0f52541def9487d3953ea36a74f63c0d04a09daa

Observation 2399d8b1-ad29-420e-9761-98510f168c3a · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.737345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.469176Z digest=sha256:5e3b81c44857fdbd722a8543f86742717d3a3b854f448c829c3b436f796825b8

Observation cc2f68e7-ea64-42c1-a70d-759ebd42b168 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.625815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.474935Z digest=sha256:21904b3aa1c5107173be5638a7f31c9f0c2005ddc6f2613ab8b8cbd2eedbe436

Observation 553dea26-92c7-4f1a-b04b-e8b0689bdf64 · outbound

This paper cites Path (multi-hop chain of reasoning): carbon dioxide → formic acid → carbonylation reactions * Source 1 and source 2 are coming from different documents.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Path (multi-hop chain of reasoning): carbon dioxide → formic acid → carbonylation reactions * Source 1 and source 2 are coming from different documents

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:28.560799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.480439Z digest=sha256:effb12cc96d99f383689f53564863fcea13d77119e21c392a66138fd97b7d72a

Observation 243384e8-1122-4e3f-a7ae-2c24014fd39b · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.465748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.565772Z digest=sha256:6f25c340d9798db5c5273a3336564b4629b28f3982bb8deaf1739917c299092d

Observation 034087ee-1c86-4b52-985c-2fd3ea838638 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.390841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.645073Z digest=sha256:f5cac5039adcba84eb1e9d762270711f67d6bb14da2b5de2985f32abc1e2f65b

Observation 178c6e3d-e52b-4415-a339-25f1f22d0379 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.340057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.676452Z digest=sha256:8f6f7be669700dd1c2108950f3d3f4c3ec5a931b33065003044c46017afdc99f

Observation 6a04d0a5-cf6c-430a-ae55-36855dfef0b2 · outbound

This paper cites Path (multi-hop chain of reasoning): solution → graphene → membranes→ nitrogen→ Cr3(Cr4Cl)3(BTT)82 *Sources 1–4 are extracted from four different documents.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Path (multi-hop chain of reasoning): solution → graphene → membranes→ nitrogen→ Cr3(Cr4Cl)3(BTT)82 *Sources 1–4 are extracted from four different documents

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:28.270081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T11:08:27.681538Z digest=sha256:2bfaf71877d9f3f9ed116dc0bf8479d6429b4b0e76ab999a7e2122cdc190783c

Pith citing papers

Observation 21c4f34c-9e8a-4df6-9b7f-669d779d9684 · inbound

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering cites this paper.

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:32:44.995439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T10:31:04.987672Z digest=sha256:621be38ce1c0efbcfb23c901ed8b3a7f4bcd01be4b1466b6ab868aa1d1035ada

Observation 1693f0ed-7475-441d-9649-716484bb71d1 · inbound

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering cites this paper.

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T07:36:16.071678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:36:16.071678Z digest=sha256:ffdd6a474d7ce806652ff0d6f9ef5bf62e0800d6a8c473fa551b05fc9fe79a26

Observation 03480b6e-f309-4b1a-ab25-7a48a7342950 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-06-27T13:00:55.949775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:3ef63aa93eb0d9326e00e9e95071e149de2aaae23c4025ecb9ec143aa585fd61