Pith. sign in

Paper Citation Record · LEDGER

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts

As of 17 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.02000.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02000 v2

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:31:45.191937Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bdf90adb-491d-4380-8de7-725ff8812b0c · outbound

This paper cites L-Eval: Instituting Standardized Evaluation for Long Context Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.196733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.196733Z digest=sha256:8443ae920f8c7739ed9aa8c9f0389b5780b926580784edf05513ccca15d747a2

Observation 4bfa2222-e7cb-441d-981c-dda2d9e3b3a4 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.981200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.259177Z digest=sha256:beb302658052277f4c7b993414e9859301abf94b786c65d6a992725d0c1a0b0e

Observation 6a246a3d-f5e8-40bc-a3c9-31178427557d · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.323817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.323817Z digest=sha256:ce2e5942f660de00e7b9f15f9f44892a14a0b562698a70bcef15a2e5ddbb9197

Observation 9cca8458-b8f9-41d0-8186-0a909cd0acc8 · outbound

This paper cites Longformer: The Long-Document Transformer.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Longformer: The Long-Document Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.391811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.391811Z digest=sha256:e1f1fc6c43b484a90c8dbd274ec130b8677822505bc585bfb0ec1404669b1f3f

Observation cd6a0f39-8641-449a-b3e6-64e5d9eb0c55 · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.483574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.483574Z digest=sha256:31d6f9484386cc4111a91fa32d19b52b5038358ce59c4cc0613d34642fe0db5e

Observation f582631c-bdcd-4b32-91b4-cfe7c5b2df20 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.792392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.578742Z digest=sha256:6e70fd756f0ce2df8a5d7e6d4f385f9315b42fe808fa49e2f9042600cf086d8c

Observation 4ca6d699-4ec4-4609-8dd3-3d697496675a · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.484142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.681448Z digest=sha256:c37227da2a992921611ca6208c3e5a47ddea0f1165de1e042d5a21a73533bc05

Observation 071381ea-b736-4724-b2e0-c4d4e06a518d · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.756912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.756912Z digest=sha256:6cc7df920828faf45d2b47dd487a80145982602e46e11b9332ba4b9a2d587c5e

Observation 75c8e2a6-0592-48b2-b12b-00e8ac1a680e · outbound

This paper cites EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.825940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.825940Z digest=sha256:f9c89f8291d6848027c39c5d02a6353fdc845d22eb0363fe10c667b462dcd458

Observation f1fee9ab-7475-48f8-a9c8-70c0d1062e74 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.900418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.900418Z digest=sha256:ddbf2e0c8a855137c307efc3409b051a31ea87bff08be428184447fe08efae64

Observation 06748e38-5412-4ddc-bcb9-28dadad64678 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.292313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.958223Z digest=sha256:d3956587323401b6842d162b745a67cd7e42ed25e758a7f5403a07380c06ca3c

Observation 37e3e0ab-b28e-4f8d-ab60-48725b56b11a · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.018867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.018867Z digest=sha256:d368892470a55c587f2364da3b3ec31c37c77171d21f071edf58118280f9f205

Observation 041d8801-d5db-41c0-acae-fa4c3265b030 · outbound

This paper cites One Thousand and One Pairs: A "novel" challenge for long-context language models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts One Thousand and One Pairs: A "novel" challenge for long-context language models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.077816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.077816Z digest=sha256:edbe8036bec542b03e144dc58c82ba095a29dd8d8762696f4dd304ab83efbb72

Observation e0e4f0b7-fa54-45f8-98ff-be9aacc1dd34 · outbound

This paper cites The NarrativeQA Reading Comprehension Challenge.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts The NarrativeQA Reading Comprehension Challenge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.137091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.137091Z digest=sha256:96106d06bf575e8781210662ba1128f9add3e95d9d087f991107faf8eabd313b

Observation 37534588-8791-4d18-a9c2-29caca12fa5e · outbound

This paper cites BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.210471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.210471Z digest=sha256:bf24a60678ef0e5773b22f06f2d8c6d2cf967d190012459c00273fcd76ed614a

Observation fbefefc8-41f3-446a-9377-e2ec961d736b · outbound

This paper cites Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.267912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.267912Z digest=sha256:80375be5d0675bd6755a2c24197138d3c8f8b4f7e0cf858d3db97df993516cb6

Observation 15cd5cb7-6c85-4d8e-998e-4bf9a0f4ecde · outbound

This paper cites LooGLE: Can Long-Context Language Models Understand Long Contexts?.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts LooGLE: Can Long-Context Language Models Understand Long Contexts?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.324780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.324780Z digest=sha256:209bd7cea131e10eb7efc9d8301cc93b5fbb16e013ce94da6e74fcd42627dbf0

Observation 82393d29-4155-4ee5-88d5-97d148794423 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.379678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.379678Z digest=sha256:40abceb5071ae09eb91ef2a06998897d2dcfdd1939b289403db01c535014f341

Observation 48ac52bb-5647-451b-98ea-0052fb0bfbf5 · outbound

This paper cites Making Long-Context Language Models Better Multi-Hop Reasoners.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Making Long-Context Language Models Better Multi-Hop Reasoners

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:31:46.057841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.438051Z digest=sha256:522a080afc15fd535393d86b699aafdf62dedcea1b5a967efddc5c230fc9f8a8

Observation 83efaa79-557c-479d-8e13-c36d6c8758f4 · outbound

This paper cites Lost in the Middle: How Language Models Use Long Contexts.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Lost in the Middle: How Language Models Use Long Contexts

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.593760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.593760Z digest=sha256:07daccb4a46e763f7dd894d0dc377891461a16128dc240c81e399dfd0653cc7d

Observation dd451ad6-6fed-415e-bc01-b57e1d81c82e · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:47.972259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.665498Z digest=sha256:6fa51bfd67b931911ca9943f99d4cf3a652d98e50e3ccf1c57fc093aef2ad337

Observation f01e73c1-716e-4e4c-a9cd-3dde364aaf0d · outbound

This paper cites GPT-4 Technical Report.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts GPT-4 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.724505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.724505Z digest=sha256:09a4923981fceab5af4d38241d1e7f4bc652f0fbe18db78acd560fa1c8974a34

Observation 67c01229-56bb-4924-a0dc-e91351698c36 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:47.557596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.810014Z digest=sha256:257da5687f824528bfa48ef2dbc50fbe33d3b0a37febd6d9fe0cb7fda85744ae

Observation de9afda2-7828-471d-a59d-7f08b679dea1 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:47.160960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.875143Z digest=sha256:03f131bb8c814a526e199621f945e7aaa73d1e3ac19e3697ec0cf2af80b78894

Observation 7ac20032-db84-4839-b3d8-6f59f3fbc621 · outbound

This paper cites QuALITY: Question Answering with Long Input Texts, Yes!.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts QuALITY: Question Answering with Long Input Texts, Yes!

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.967050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.967050Z digest=sha256:b99e32bc3a18d3f5bc9b44f9d20ff6c939874e9910b29179d328f53f723da0e8

Observation 25f096e5-05cd-4146-bdfb-a0e5a630a30c · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:46.779048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:44.038710Z digest=sha256:74220d4cad98f5442a7418d1f2b8cfd1032307fb886bf4ecdd471a80fee01e37

Observation 2ef4bcdd-f7fc-43c1-bf0f-47719216ef2b · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.114430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.114430Z digest=sha256:4e217719a266142461ee5fcfac6ad3a63eec55920871771cbce3421eae9ef86e

Observation cbd023b7-282d-4cc8-8988-8356a1ee77bf · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.178221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.178221Z digest=sha256:50e04aad68fe6ad8dc064920ea63558f9f3f1151a1284bf825911f4a4e356844

Observation 4b1bcb33-ba5a-42d5-a058-b33a85c2f86c · outbound

This paper cites Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.194237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.194237Z digest=sha256:af85dd812a0a422dad078c88ca33591b14a723769252fc31939e908d1e087842

Observation b9581f1e-2f10-4153-8766-480f809f0ba6 · outbound

This paper cites Constructing Datasets for Multi-hop Reading Comprehension Across Documents.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Constructing Datasets for Multi-hop Reading Comprehension Across Documents

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:31:45.621660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T15:31:44.246400Z digest=sha256:3ec503b00d8ddb7d432a7611525fa66c7d995e57eed026115c6e93a760e07695

Observation 5ff3d9e1-8d88-46ad-9254-180582f74010 · outbound

This paper cites Memorizing Transformers.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Memorizing Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.312581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.312581Z digest=sha256:6effcdd2bc6c855420d58151c0f17e6eabaa155fd80dc4a225ac9207a73458fb

Observation 56785ef9-37ba-433d-8bfc-7cc9266b1d39 · outbound

This paper cites Retrieval meets Long Context Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Retrieval meets Long Context Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.407153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.407153Z digest=sha256:99259073265d1d5ccea23fd2d005ba98c12a6a43605e1d97b85d3240cd8b3260

Observation a4e40009-68b7-40be-bbd4-977788e6046e · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.488966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.488966Z digest=sha256:8fb31e16e021060c8ec774a9dba2828659a15822d0716820de862800ca9c61d1

Observation e8e23118-a6f0-417a-ba76-045f36f4d1ea · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.581369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.581369Z digest=sha256:2f56d1c3267d9ce2b941b72df0c6643ccac9dbc69f37331c3a0ea17d7e0c4451

Observation 6dd3275c-d11d-4ba1-a9c9-f1f20bed4d6c · outbound

This paper cites Big Bird: Transformers for Longer Sequences.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Big Bird: Transformers for Longer Sequences

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.720004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.720004Z digest=sha256:fba4ba1ff7d1c74c52500cc0b0ce012c8e5e8d4097b38af75a57a26289b9a89e

Observation 72a004f4-ea60-4886-ac44-e919bf40462d · outbound

This paper cites Marathon: A Race Through the Realm of Long Context with Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Marathon: A Race Through the Realm of Long Context with Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.836384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.836384Z digest=sha256:273f4af7a5ea60e02d57cdfae0094439c93195fb695a1206db32ac77ed03fc64

Observation 9019518f-0e40-4776-96d6-5b58b9a00648 · outbound

This paper cites FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.948205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.948205Z digest=sha256:3f67c30353eb71f6cbfa61adeff6d70cf322351eeeac06c06ed55ea4d97d5409

Observation 6401e0da-9b36-47da-ab4a-61d95f037865 · outbound

This paper cites online" 'onlinestring :=.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts online" 'onlinestring :=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:45.066367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:45.066367Z digest=sha256:023f92151a7389efa2ea99ef4824a150011f3f202b3582fd878e589441c8b50d

Observation 5adbf85b-d3d8-4f0e-9ad2-0e3d1f172064 · outbound

This paper cites write newline.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts write newline

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:45.191937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:45.191937Z digest=sha256:bb11f4ec75907be812968b8d224cbafa8debace80148fd3812bac39167fdfcf5

Pith citing papers

No inbound Pith citation observations are available.