Pith. sign in

Paper Citation Record · LEDGER

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

As of 16 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 3 inbound Pith citation observations for arXiv:2504.16414.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16414 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:08:27.681538Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T07:36:16.071678Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 5c5e6c69-650d-4882-aaa5-87e6b52bbe80 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Chain-of-thought prompting elicits reasoning in large language models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.033380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.033380Z digest=sha256:d29059b00e6d2db144248dc512471cec5d2d6a2474e73cb834646435b6dc55b0

Observation 786957e3-11ee-4bd9-8cf6-b17ab453974b · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.041280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.041280Z digest=sha256:10d6a1a472d8249214b53ba3368db82b5a568f365bd96bea36d030a05263c96c

Observation 944effd2-5832-4a98-a42d-1168409c548c · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Tree of thoughts: Deliberate problem solving with large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.047544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.047544Z digest=sha256:9eab9811f9bd95113385180ad0dcc33786baa9fa8a8e2d34a18ca8f0f90d97d6

Observation 32f29f2b-c9a0-4b7c-921d-153931ba86ab · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Graph of thoughts: Solving elaborate problems with large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.694755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.055284Z digest=sha256:03d91f0c456a138caa414f2907e321270d911e6a65629a80e556f9dcb0dea3ac

Observation 13d59eb0-efc0-4df1-b3e2-b9aa560dd2f3 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.062988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.062988Z digest=sha256:d0badcd7c3a18deb421b17c44a38bf80c97c1028a7b10ea9e396fae04923929f

Observation a7b19a6c-a234-4bdc-ae53-5b648bc62361 · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Retrieval-augmented generation for knowledge-intensive nlp tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.078801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.078801Z digest=sha256:5b04ff34a9139932c072dcc7339e1af16ce328215134d22c8641fa923b137494

Observation e6a74663-37b9-4855-9319-3a2c3aa1ec46 · outbound

This paper cites Neurosymbolic ai: the 3rd wave.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Neurosymbolic ai: the 3rd wave

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.453558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.086596Z digest=sha256:7a88959d7d172830f3ab33eb2134834f088f982e4b9b758599b00d4703e26db1

Observation 9d08a5c9-624c-448b-99f9-32434f31f4a2 · outbound

This paper cites A simple neural network module for relational reasoning.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study A simple neural network module for relational reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.093328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.093328Z digest=sha256:d35220e494dd4c5cd35ac8826d7e119fa9709ff0f1c75764b24350a598ad666f

Observation 97572626-9b5c-4b01-8063-0926d91a566f · outbound

This paper cites Openai o1 system card, 2024.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Openai o1 system card, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.352691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.101824Z digest=sha256:39f2e7d71b52138d3781d37f5add0c30875d33cf75f60d32c5f78481cc2f6381

Observation 6cc56a61-3388-4dce-a4a6-828a2c935406 · outbound

This paper cites Openai o3 mini system card, 2024.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Openai o3 mini system card, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.256718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.111018Z digest=sha256:95a8290aa8fd32e0f2a935ec28e38df3f2d796ffdbba5c45f8340607930157ff

Observation 66bfd8a0-6bde-405b-a898-8e781124fff6 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Star: Bootstrapping reasoning with reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.118353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.118353Z digest=sha256:32c312dae1fe8c44625ffedc56fb7401569139a081584b109481b0ef76ed2ed9

Observation 3fa5a002-6426-4f1f-831d-0a6aaf77bba8 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.125584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.125584Z digest=sha256:1ef7856fa07902f26dacab6fa6f4a1e9215c5947e35464b540eb6eb3ef0521e0

Observation 29e29159-c387-401a-ac6e-a105ab72a4c6 · outbound

This paper cites Advancing Reasoning in Large Language Models: Promising Methods and Approaches.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Advancing Reasoning in Large Language Models: Promising Methods and Approaches

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.135194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.135194Z digest=sha256:5ca05017b0856baec705b2bc5aa6daf58425d7eced3eb997bdbf25b1ee23a673

Observation b230e105-f40e-4ff2-841c-8136172dd967 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.145988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.145988Z digest=sha256:d301001cdb048f3dfe9052e482d1dd2b072e0fc3f7571f2ef96da0feec274b02

Observation 24e21779-dfdd-4895-b1b4-90719cd3759b · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Measuring Mathematical Problem Solving With the MATH Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.152376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.152376Z digest=sha256:aad17dbf1dfded2acf2d7e1538250b0877256a7db8624b751d1f38d1d7e86502

Observation 82536658-847e-4cb7-b819-fb84870c44b8 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Evaluating Large Language Models Trained on Code

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.160516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.160516Z digest=sha256:1d67bee1c38a15f5a87b37a54ca3b3d4b8c8ed583cfd6205a83ae1d8a4f6bd70

Observation 9204ca91-0a5e-4a88-8c70-8adb6cb214de · outbound

This paper cites Program Synthesis with Large Language Models.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Program Synthesis with Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.165982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.165982Z digest=sha256:2505da2cc41fac16f847c98a97cf601589fe7df5c0ca0b69f4425f557039db9e

Observation bf725a04-e15f-44d4-a024-f57cc3f71d37 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.176304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.176304Z digest=sha256:937bc1990ee877c26c7011968a12c3c1916a58705b334e51a8a747bd4ba41327

Observation 740f52cd-73f3-4f59-a006-42bdbc301f89 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.183147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.183147Z digest=sha256:a1a0eed6e42e225fbd4ab6f7d72552647a1134bf21bdcaca7a93f2e819690d06

Observation 3a88ce35-5293-4571-ab1c-c47953128198 · outbound

This paper cites Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.198025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.198025Z digest=sha256:f1bc69dafa941b67dd63fd7bf47bee72c70604e6bc21e520fef43f35411e290a

Observation 06b0d660-6999-4e5d-bc64-5b22630fae49 · outbound

This paper cites Chemlit-qa: A human evaluated dataset for chemistry rag tasks.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Chemlit-qa: A human evaluated dataset for chemistry rag tasks

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.156864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.205317Z digest=sha256:6c8013b8a28ec058f258fbc42c551b9b5115cc202ba22a7870a8962e01fb6776

Observation e907d867-c733-4885-97e2-288143e5a7a0 · outbound

This paper cites OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far?.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.211322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.211322Z digest=sha256:9a2cd9e90ab9e0e7aa70f279061b0ab50e15a579dc5ae0d5a009e8512838e60f

Observation d7f7b64c-6296-4cbc-bb26-b4ab59c56f40 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Gpqa: A graduate-level google-proof q&a benchmark

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.228224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.228224Z digest=sha256:1572c4fc17e4d41e841ce7c6a13817b61d48024896dbabfc09977e067d62bfa3

Observation e8236d33-6d50-4e29-9cec-efee71ba2911 · outbound

This paper cites Large language models for reticular chemistry.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Large language models for reticular chemistry

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:30.002936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.234683Z digest=sha256:7f70859e4c749e3b51e4b0d027fb9109a894c93a48cfce2d833df87806012756

Observation 36a26e47-c694-4748-9e8d-ca759c55071e · outbound

This paper cites Multi-hop question answering.Foundations and Trends® in Information Retrieval, 17(5):457–586, 2024.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Multi-hop question answering.Foundations and Trends® in Information Retrieval, 17(5):457–586, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.907012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.241942Z digest=sha256:947864d3b203175053f41a1b90175c5b459c2f0889194f68b7e352e123805b16

Observation b5031814-288e-4a46-a47c-b1f0504b416d · outbound

This paper cites Constructing datasets for multi-hop reading comprehension across documents.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Constructing datasets for multi-hop reading comprehension across documents

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.862076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.249274Z digest=sha256:81f4ca82d9d8ec3001c34a276734e04967d3c59a073c2e31c0b96ab4e82f5172

Observation 9d0f70dd-d8b9-4988-9160-e4680425be07 · outbound

This paper cites Musique: Multihop questions via single-hop question composition.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Musique: Multihop questions via single-hop question composition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.255231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.255231Z digest=sha256:0bc01956f1907e8778fb40d48af43442a5b3c57c76341896340002fb62e9b216

Observation dbc7d224-0086-49fc-94a0-00d046a57c52 · outbound

This paper cites MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.261607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.261607Z digest=sha256:e0af6dfedc963ac69aff676e48133c79e00f7ab5fbef940fab06c7c2c2d48208

Observation 3e491458-4e77-48f0-840e-19e2a3c7e46a · outbound

This paper cites A framework for evaluating the chemical knowledge and reasoning abilities of large language models against the expertise of chemists.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study A framework for evaluating the chemical knowledge and reasoning abilities of large language models against the expertise of chemists

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.702468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.269213Z digest=sha256:6314c09ac66bc6d0e9751b39aa5e1452cd1d5447efb425a79c41d2815093cd78

Observation 8ccf9fa0-2b60-46a4-ae87-028672acace6 · outbound

This paper cites Knowledge Graph Generation From Text.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Knowledge Graph Generation From Text

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.275286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.275286Z digest=sha256:9f60e3e3b7e832f7765777240051e14cdfce9d14ecf1046987db938a925c37c2

Observation 52d9b364-0a8c-4870-afbd-052a4e0e89e6 · outbound

This paper cites Extract, Define, Canonicalize: An LLM-based Framework for Knowledge Graph Construction.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Extract, Define, Canonicalize: An LLM-based Framework for Knowledge Graph Construction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.280841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.280841Z digest=sha256:3c31cc71c04a51663b4f3018472e4abec3f46dd4a0eb79f00f74f8567d9344e1

Observation 0c10439b-070b-4430-866b-dc528a1dd822 · outbound

This paper cites Building Dynamic Knowledge Graphs from Text using Machine Reading Comprehension.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Building Dynamic Knowledge Graphs from Text using Machine Reading Comprehension

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.287652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.287652Z digest=sha256:1592f277bb49d4b9f182385e959069a1ebb747fc0524d2e5048584abaf1636f6

Observation ba8d481d-d409-4b24-8b71-6407539fcf87 · outbound

This paper cites CEAR: Automatic construction of a knowledge graph of chemical entities and roles from scientific literature.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study CEAR: Automatic construction of a knowledge graph of chemical entities and roles from scientific literature

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.295596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.295596Z digest=sha256:33204d858071c3d762648329459b4842e71bae6341c640264be2ae29526251cb

Observation a1ac075f-247c-4ed7-bd18-ba773281e3b5 · outbound

This paper cites Coarse-to-fine knowledge graph domain adaptation based on distantly-supervised iterative training.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Coarse-to-fine knowledge graph domain adaptation based on distantly-supervised iterative training

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.614194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.301310Z digest=sha256:2d84454e5ef2025a602c60ce36926f34ecc99ae9187597398d12d0c093614bb9

Observation 5576a2ae-05d5-4cc1-b051-9616aaf09c2f · outbound

This paper cites Nilinker: attention-based approach to nil entity linking.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Nilinker: attention-based approach to nil entity linking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.537106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.307601Z digest=sha256:cd2a61fe5c9ad6289eee5fb735ef032c8490641dba40cfdff7130450f7a6e150

Observation da72f682-ae47-4430-aec0-92ac7fb792b6 · outbound

This paper cites Domain-specific language model pretraining for biomedical natural language processing, 2020.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Domain-specific language model pretraining for biomedical natural language processing, 2020

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.313344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.313344Z digest=sha256:821bac45d37dde46fb58fca3184fa73bde1cabc750c2252bced5a48209f5b8b2

Observation 7fb675f9-8468-4f99-8af4-6e5a5aa47793 · outbound

This paper cites Pubchem in 2021: new data content and improved web interfaces.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Pubchem in 2021: new data content and improved web interfaces

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.372348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.318415Z digest=sha256:991f30a558638ea0093ff8b98a72a29c6b00ff317860e8b09f8a8c0b216ffa13

Observation 2436015d-32bf-4163-bba8-422f32711f77 · outbound

This paper cites Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.323673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.323673Z digest=sha256:b1f339669810d2aa58854615a186f556edf391c9d31b526303ed2fe0bee314f1

Observation 32c19761-3314-44f5-9941-174a5c16880e · outbound

This paper cites Hierarchical Retrieval-Augmented Generation Model with Rethink for Multi-hop Question Answering.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Hierarchical Retrieval-Augmented Generation Model with Rethink for Multi-hop Question Answering

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.341341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.341341Z digest=sha256:0141392d0ed507d9b0fae51aa2aa2b50618cebe9e9c9b8b7a69596b31df16efa

Observation 421bf404-e7fa-49fc-851e-293c94ef9bd2 · outbound

This paper cites HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:08:27.361060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:08:27.361060Z digest=sha256:fb452a45c9d5e94cca65bb23498462eaf77e0c44e3020ca212b6cfcce6bb7549

Observation a1a2e9e7-84bb-4fd1-993c-d390a5c8f5d1 · outbound

This paper cites If an entity appears in the text but has no meaningful chemical relationship with another entity in the set, ignore it.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study If an entity appears in the text but has no meaningful chemical relationship with another entity in the set, ignore it

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.292668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.379975Z digest=sha256:57bab153c8f3f0ed0d9b588de8cace5f75a776b6180a03f0ffba23db1b60384b

Observation f1c16554-379c-4678-a088-4ba9c7626a5b · outbound

This paper cites reacts with,.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study reacts with,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.177310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.385205Z digest=sha256:f321131ba33d317ce4dd3fcd4bd031695e652d1c0f17a192b6f3f9724fc180ed

Observation 798ee359-8162-46fb-bcad-67cfede7a0ae · outbound

This paper cites Avoid observations, opinions, and findings.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Avoid observations, opinions, and findings

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:29.133824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.391936Z digest=sha256:964727bf34791077376b0b3191427d5c2d29cabdb7e62084a140d05dc608e059

Observation 43c2180b-0fd6-48b7-b3cf-fe615e51a8b6 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:29.013438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.413495Z digest=sha256:e752007029df708e76ba997060c35dca850035ff305cd7d967d994a8ec16c67a

Observation ce6a1cbe-346f-4bc0-b870-e256db6e8ae0 · outbound

This paper cites is," "are,.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study is," "are,

Reference 45

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T11:08:28.936126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.457183Z digest=sha256:86f57467f5b867c92ea3d4c88017f8ff128a42f2aecf371ca64dca60ebb7d09a

Observation 92e7b5b1-7d25-4811-bb54-ede837e6e937 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.830681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.464119Z digest=sha256:63c3e50b42570244d7a98fa25c19e3a197fcd8dfe02389bc156907567f54cc54

Observation 2399d8b1-ad29-420e-9761-98510f168c3a · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.737345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.469176Z digest=sha256:a4510ed44fc79bfd9b7957aad94a82c2b05c62bd43949588245dbad9ff982cac

Observation cc2f68e7-ea64-42c1-a70d-759ebd42b168 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.625815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.474935Z digest=sha256:a27b1128d8877847560a4ebe27ff675231734cc480e0a5438ad0f0f2a96d0ce2

Observation 553dea26-92c7-4f1a-b04b-e8b0689bdf64 · outbound

This paper cites Path (multi-hop chain of reasoning): carbon dioxide → formic acid → carbonylation reactions * Source 1 and source 2 are coming from different documents.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Path (multi-hop chain of reasoning): carbon dioxide → formic acid → carbonylation reactions * Source 1 and source 2 are coming from different documents

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:28.560799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.480439Z digest=sha256:4cdb41851db6971d94b5a4e7413147adc79247c35d8ffa5719638d3c196f188c

Observation 243384e8-1122-4e3f-a7ae-2c24014fd39b · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.465748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.565772Z digest=sha256:340bc7a853d191d6b2e4891070aeb575394079876592a898eb21ce77cf2a9293

Observation 034087ee-1c86-4b52-985c-2fd3ea838638 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.390841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.645073Z digest=sha256:4965e86df826b8821387785dcc4526054efd6050823106c8e9cc44a8a0e42057

Observation 178c6e3d-e52b-4415-a339-25f1f22d0379 · outbound

This paper cites an unresolved cited work.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:08:28.340057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.676452Z digest=sha256:1c332c6f78c146ea67ede838b757300f2d529810916c2f78cb9d3f1f351586b5

Observation 6a04d0a5-cf6c-430a-ae55-36855dfef0b2 · outbound

This paper cites Path (multi-hop chain of reasoning): solution → graphene → membranes→ nitrogen→ Cr3(Cr4Cl)3(BTT)82 *Sources 1–4 are extracted from four different documents.

Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study Path (multi-hop chain of reasoning): solution → graphene → membranes→ nitrogen→ Cr3(Cr4Cl)3(BTT)82 *Sources 1–4 are extracted from four different documents

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:08:28.270081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:08:27.681538Z digest=sha256:cd783240ab32260b60f459c866b1ed6aad683fe738572764d46c3889d9b26081

Pith citing papers

Observation 21c4f34c-9e8a-4df6-9b7f-669d779d9684 · inbound

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering cites this paper.

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:32:44.995439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-16T10:31:04.987672Z digest=sha256:68265ba45c8f68ba8a0fea625d2341fe47aa51fab2b3f0c4bd7212eead528dd9

Observation 1693f0ed-7475-441d-9649-716484bb71d1 · inbound

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering cites this paper.

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T07:36:16.071678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:36:16.071678Z digest=sha256:509aa0cec1ac8c11cc536fa07922d439359fa7960d32fa92a47ffdb23c065c6e

Observation 03480b6e-f309-4b1a-ab25-7a48a7342950 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Evaluating Multi-Hop Reasoning in Large Language Models: A Chemistry-Centric Case Study

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-06-27T13:00:55.949775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:56fdec52baa4b1a2b6d98f85dcadd01c0c98556a1a49283699f43f1d93660fbd