Pith. sign in

Paper Citation Record · LEDGER

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning

As of 20 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2607.09328.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.09328 v2

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T07:40:58.826285Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4165941-492b-4e1c-b4a2-6a4d33f0e1ac · outbound

This paper cites Raw-source evaluation.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Raw-source evaluation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.826285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.826285Z digest=sha256:a425b6a7c31d11c4242216ed80fd3bcba2444e3795d10f01aa3dc8f69452e156

Observation 3a4b5c20-6e72-4afc-8839-fd879b986223 · outbound

This paper cites URL https://arxiv.org/ abs/2606.19348.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning URL https://arxiv.org/ abs/2606.19348

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.860969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.860969Z digest=sha256:83d9e85f20be97b1eb24f4d2c16b7545c95cff06384d0f55d494db836c726128

Observation 4622702c-c9ff-45af-95ab-d7f71a5845ed · outbound

This paper cites DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.937791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.937791Z digest=sha256:9aaf9d9d5a33d52f153e67030f4f251f5565b5380acad6526db72a8ae8548fb6

Observation 197b182d-0c74-49dd-adfd-2f2f0e5128d3 · outbound

This paper cites DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.029050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.029050Z digest=sha256:b145fc5969725e63473b3440495a5ae2097e4280181f7fa34d328d20dded02c3

Observation 7ca26883-e3ce-42d8-9d41-76d8e40b4db9 · outbound

This paper cites Gemma 3 Technical Report.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Gemma 3 Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.168353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.168353Z digest=sha256:b760fb9ad64618d5f796daa888f22fa8eb4d045cf124d78f433243e1ece9e428

Observation 8af18281-6573-4540-b62e-5ed686f74739 · outbound

This paper cites Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.445058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.445058Z digest=sha256:17ec5b9cb7d1a9eba00130760dabe056f87c5a6d2f1007ea12acbd5987ab1533

Observation 4fa34d82-c498-4dd2-b4a5-5e6b5208bdea · outbound

This paper cites an unresolved cited work.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.611667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.611667Z digest=sha256:5e89569da97b1a38261fbf5c168cd3ff6146f91b1a8218fdf26eb28ecc478fb4

Observation 2f2da1a5-1f07-41fa-8233-ec0c4fa9635c · outbound

This paper cites Nelson F.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Nelson F

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.670841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.670841Z digest=sha256:9921e12bb2a7321e4b59bb6c31a08f024db20fbcb71eeb959b62d2af6705a923

Observation 7260b785-016b-44bb-a2ff-a3db3dd3e8e7 · outbound

This paper cites Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.900586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.900586Z digest=sha256:3108133dbcd90e16cbf178f5f43e2f24ea7eebda3af8478a37095d43cc1da675

Observation 0d42286d-b6b7-4820-98a6-6a6c136b73a7 · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.024753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.024753Z digest=sha256:03928d2c30417af8e0ca1eb59fb347fa623c3f5f5408222e1846d7ebf7a72e6b

Observation d22003ee-8a96-45b7-926b-793b11998dfc · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.166701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.166701Z digest=sha256:ae3fec9b8722c1c5eb2dcc09116cb23b16857179a34cffdb1a4e21ba7bec0140

Observation 3296dad3-f356-467d-b8b3-19ec7a5bf16f · outbound

This paper cites Humanity's Last Exam.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Humanity's Last Exam

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.554756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.554756Z digest=sha256:6d62e9c6e6eb4a996c71125ff365ce7c16d13c8af47d2a59adc9d426c475b484

Observation 120ccf54-e084-4efd-8057-16fdb228a777 · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.650142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.650142Z digest=sha256:11d0563204c436946e724b971240723a4c9d88284c71865440caef2bd9ee8034

Observation 9a7f0693-f639-49ef-8d61-87bfdc6691d9 · outbound

This paper cites BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.984751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.984751Z digest=sha256:572544cdf1407456e349c8a83bfc0a0d4f9fb9685308947e0f46c253e251c0e8

Observation fe763d1b-0d53-4554-bd91-4ce47b9a6495 · outbound

This paper cites Kai Yan, Zhan Ling, Kang Liu, Yifan Yang, Ting-Han Fan, Lingfeng Shen, Zhengyin Du, and Jiecao Chen.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Kai Yan, Zhan Ling, Kang Liu, Yifan Yang, Ting-Han Fan, Lingfeng Shen, Zhengyin Du, and Jiecao Chen

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.114779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.114779Z digest=sha256:2e71c728618a480814411597e5b2d551b3704f1f3678a354bdbd1dbe407387fa

Observation b118c4f5-32ae-4d1a-aa9d-fbb97ea9b043 · outbound

This paper cites 100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning 100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.250915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.250915Z digest=sha256:179d8f03dc420bf5891ecbc0ac141825fd46291ef1905678b931094531e47d44

Observation 7b9d11d5-bfea-4a30-a41a-fdc19305862f · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.349130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.349130Z digest=sha256:231b2850fa56e8f9402abc8fcc2970324e1b8e26ff0c299b95868dc4bb2e4991

Observation ef65d078-8b46-4f19-b04d-d11b6c463472 · outbound

This paper cites Academiceval: Live long-context llm benchmark.arXiv preprint arXiv:2510.17725,.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Academiceval: Live long-context llm benchmark.arXiv preprint arXiv:2510.17725,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.484748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.484748Z digest=sha256:02a2ae357ca31609ec8cca7db950ca94f3ab3d02784a18fc78dc3f2678a400d3

Observation b8133e90-1526-4ae2-bf84-6b0bf42b83dd · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.784842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.784842Z digest=sha256:3354e265c731a3923ea27c14edd8c8e8888a44436ce8a02e97a1f65c1b41dcd8

Observation 5b0fcf00-ff01-4511-b418-633c4506c10f · outbound

This paper cites BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.540931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.540931Z digest=sha256:69515bd0496ecb46ff1d651dcb86eec62326fcb544f4e5b11cb8a44d05b441d7

Observation f9655137-add8-40a1-8f72-c849a75f5c15 · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.244749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.244749Z digest=sha256:59af5499c598c2a412de02c2c7ed6e545a17eef717b4663b40ec088a66dfad11

Observation b1c97332-b68b-4b22-9915-9ee9be2efefc · outbound

This paper cites doi: 10.18653/v1/2021.naacl-main.365.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning doi: 10.18653/v1/2021.naacl-main.365

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.772989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.772989Z digest=sha256:d4360e94e68295761224f9fbcc87fdd7f1e112c1430f29421545785935122f7d

Observation a7c6c680-f4cd-4972-a310-bef86399acc0 · outbound

This paper cites Gemma 3 Technical Report.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Gemma 3 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.105396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.105396Z digest=sha256:78e0a5d6df7a4a4cc57e474df718eb6da814ca555a7b13fdb9e26a2b77178493

Observation ceac9ee5-faf5-4ba3-9d24-290cb87df2b8 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.337896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.337896Z digest=sha256:328e4258e392d901660099b14371d0a84ff5625bd11e04cd3b301c50900ed0a3

Observation 61016f6d-48ac-4a0b-bd86-1562f88466a4 · outbound

This paper cites acl-long.183.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning acl-long.183

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.418414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.418414Z digest=sha256:936fe23dc2f2387558e0da7c6f991f8f1196b0828383468450e612c3f64320c7

Observation f2e7e000-6061-4f37-8cfe-1dcaf8c8331e · outbound

This paper cites Pradeep Dasigi, Kyle Lo, Iz Beltagy, Arman Cohan, Noah A.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Pradeep Dasigi, Kyle Lo, Iz Beltagy, Arman Cohan, Noah A

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.620497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.620497Z digest=sha256:9a8f1f2883db0bfd744b54762adbeb17de32a2b9331953d1d72ed89bc9a2bf6d

Pith citing papers

No inbound Pith citation observations are available.