Pith. sign in

Paper Citation Record · LEDGER

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments

As of 8 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2607.17291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.17291 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T18:30:54.704835Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a302ecf0-7aab-4957-86e5-b71a8b4ceaa6 · outbound

This paper cites Self-RAG : Learning to retrieve, generate, and critique through self-reflection.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Self-RAG : Learning to retrieve, generate, and critique through self-reflection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.362916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.362916Z digest=sha256:7d571de72ed845fdd172dad9e055227549e86d0b17253037875867cee4087ffc

Observation 3fd44d94-8aa0-4cce-b1b0-7abd5885881f · outbound

This paper cites Benchmarking large language models in retrieval-augmented generation.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Benchmarking large language models in retrieval-augmented generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.407939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.407939Z digest=sha256:0bf6b5e17c0cac134a254cdc3e437e64ff0660282c1a552922eb51ca611e2051

Observation 087a70de-e1d6-43c4-8c3e-a932de9f66b9 · outbound

This paper cites BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.495212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.495212Z digest=sha256:a6e3ca7443b867c581a447a72806d6e787c8844643662c691be1c8d47a5cd73b

Observation 1a102dd7-26a8-4f18-87f4-d76c736dd404 · outbound

This paper cites The power of noise: Redefining retrieval for RAG systems.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments The power of noise: Redefining retrieval for RAG systems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.576961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.576961Z digest=sha256:0cfb61f43c78d925ba1b869e110dd3978bfb36ee0e49fca3bd1fd38fc332bf6e

Observation d3fce4c0-e569-48d4-b4d2-3d5836a235ea · outbound

This paper cites InteractComp: Evaluating Search Agents With Ambiguous Queries.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments InteractComp: Evaluating Search Agents With Ambiguous Queries

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.639364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.639364Z digest=sha256:016436b4f55ab6498e734442be92863050f623be7c0429a9150e171c0a81d4eb

Observation 8a21d8bf-4f7c-4c0b-a1ac-e2b9325ba385 · outbound

This paper cites Mind2Web : Towards a generalist agent for the web.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Mind2Web : Towards a generalist agent for the web

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.696841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.696841Z digest=sha256:a5d950399462cdfbfe4fd7a9f6bd1f906b26ce6007e48fea1cc8b297e2c7dab2

Observation 31a707f2-89e4-442c-8cce-3d0aa584e0d6 · outbound

This paper cites Enabling large language models to generate text with citations.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Enabling large language models to generate text with citations

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.790279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.790279Z digest=sha256:315d031c7b473c920e4b2e050c5429dd486d81b89df61fbf669eb39a60eecc8e

Observation 47de9def-c6ef-4b8c-bbe6-307e3b9653cd · outbound

This paper cites Large language models cannot self-correct reasoning yet.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Large language models cannot self-correct reasoning yet

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.871204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.871204Z digest=sha256:ce88b4f22dc682791d344741cf7ffe2e8bf232a3499c8064a0a08ecccb15d57f

Observation e4819322-b758-436c-be22-9509f0de1a39 · outbound

This paper cites Survey of hallucination in natural language generation.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Survey of hallucination in natural language generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.946310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.946310Z digest=sha256:28966f1e4019568d65351efca0c56b78c0a04abaa154ef20e4aee45a21d9b2d8

Observation fc3eae84-44d1-416b-b83f-9c6812745797 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.047071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.047071Z digest=sha256:b3e6f966a9edd4fca983bd9f695b244aa30815f5bb64021b457ab6b588219718

Observation 4ebc4be2-8a1e-4196-a517-ad28fa53ec7e · outbound

This paper cites Dense passage retrieval for open-domain question answering.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Dense passage retrieval for open-domain question answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.098388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.098388Z digest=sha256:c2ea491202a515d721275f029c913aa7157a65b4d7692aab9c13362ff89c4339

Observation ccc8f459-a55a-4a6f-a86a-6967a6d413d2 · outbound

This paper cites u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.138565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.138565Z digest=sha256:e02fbce46c4ee927b5663e1b8070b5a2c4de08dbce56b36d13d5b4fa661dd297

Observation 5027daa3-2cb5-47ee-a96f-15307575cdc8 · outbound

This paper cites Entity-based knowledge conflicts in question answering.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Entity-based knowledge conflicts in question answering

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.195723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.195723Z digest=sha256:d356586ff722e51a4eec540dae218bb1fb62dc910a6ef88d04c6c4b1976d0ffd

Observation e00d7f77-85f8-49b9-945a-30f86d347030 · outbound

This paper cites Self-Refine : Iterative refinement with self-feedback.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Self-Refine : Iterative refinement with self-feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.276903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.276903Z digest=sha256:5b243235c9e3dd5ba4001bb078adc05e43d8c67a6903ace477f645c343130263

Observation a784d6c6-357c-4f6c-8b73-e41b9b9789b6 · outbound

This paper cites GAIA : a benchmark for general AI assistants.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments GAIA : a benchmark for general AI assistants

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.366500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.366500Z digest=sha256:d36b0213ffb58a8dc824a12760b8334f1bf2bc6edd581517f41695c5159c1b64

Observation 27e21187-0083-4d61-904c-192dac824c12 · outbound

This paper cites FActScore : Fine-grained atomic evaluation of factual precision in long form text generation.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments FActScore : Fine-grained atomic evaluation of factual precision in long form text generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.441867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.441867Z digest=sha256:8aa8f59e6cbc3d19eee2da94511c02e8ceaabf7cec4ed3fdf6975d1b491765f7

Observation b3026955-af08-4d11-bd8e-c87f2e9fd5bd · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments WebGPT: Browser-assisted question-answering with human feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.526685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.526685Z digest=sha256:8657537e96aaee2ca1550843fdde07d045304c1def2f3ea74fd65434b90e785f

Observation 5c6604d0-5cbb-4039-b99e-476e3d3a52ab · outbound

This paper cites On the risk of misinformation pollution with large language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments On the risk of misinformation pollution with large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.617131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.617131Z digest=sha256:861725144dccc9f1ada849df6ff21f382ccb035a238edb37b668dcd1485e5b07

Observation 9f9430e7-baa4-4a53-bc03-722d8c1dd1ad · outbound

This paper cites Tell me more! towards implicit user intention understanding of language model driven agents.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Tell me more! towards implicit user intention understanding of language model driven agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.694401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.694401Z digest=sha256:3b6c5fe23ab17aac84d21dd543f931ab4adc3fd40cce2895fe0f35b82b35e5e7

Observation cc042a88-9acf-4e27-8c21-6e5899611d24 · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Toolformer: Language models can teach themselves to use tools

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.757617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.757617Z digest=sha256:756914ef652344a2a2ae44d488d465d7b4fa46e0f39b11f7db53d77b956ffe88

Observation 7e865149-129d-42d3-8c6f-7ef6bc5e22c3 · outbound

This paper cites Towards understanding sycophancy in language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Towards understanding sycophancy in language models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.859890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.859890Z digest=sha256:b19b837fe7a08dfc08d0fe374dba054e71292b3fc7870c9dd6ee06e2210138ab

Observation 46d162ac-b9c6-4384-8ed9-d93c46e01de2 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Reflexion: Language agents with verbal reinforcement learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.980983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.980983Z digest=sha256:82c8a0a163daef1422ed0a42890a931e9a031e9c549fbb32cba1be4751e36a76

Observation 8abf6510-6d49-450f-920f-5033fb5b907c · outbound

This paper cites FEVER : a large-scale dataset for fact extraction and VER ification.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments FEVER : a large-scale dataset for fact extraction and VER ification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.162196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.162196Z digest=sha256:81c9c79f40ecea592e9cef20e4dbb0f0814518d2d9b423afde834deccf56c1f6

Observation fcc34e7a-1047-4684-a31a-a9b8d2d34290 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Chain-of-thought prompting elicits reasoning in large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.315268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.315268Z digest=sha256:62d0844608c0274ec407bc6c157b3118f4141bef7741cd263d7a244e2c999f2a

Observation af24c92a-46cd-4183-a95a-4a974bf07c1f · outbound

This paper cites BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.419774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.419774Z digest=sha256:f3d5a9d2e1d8cf9f942b6531f33f3b6b6a988110090c0e78b5289ca601500fb2

Observation a1ae14c5-955d-4c88-8cac-20e311514e66 · outbound

This paper cites WebWalker : Benchmarking LLMs in web traversal.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments WebWalker : Benchmarking LLMs in web traversal

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.539216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.539216Z digest=sha256:e82d90b3140448762504023facb0008bbfa993baf0527c38f9ca8897fa43e47c

Observation 118c5f78-2fb3-470c-912b-082c0aaaca96 · outbound

This paper cites Knowledge conflicts for LLMs : A survey.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Knowledge conflicts for LLMs : A survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.662093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.662093Z digest=sha256:b383434e66060a504ed52cd2aa0e46286de305094cb1fbc75b50fb3b383b1f8b

Observation e9b249ed-0017-4c05-82c0-216a1b371d4b · outbound

This paper cites Narasimhan, and Yuan Cao.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Narasimhan, and Yuan Cao

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.786556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.786556Z digest=sha256:6af74b50864a271a5d474f36549e47de9938a816b64ed3948ebbe6c36d01ac1e

Observation 5fbf085d-a3d0-4284-a062-5189c386e136 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.883343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.883343Z digest=sha256:1917facbd37a74eb2c9a5066c9130732373458cb3425d231078e9e107532e8da

Observation 14afcf14-edf6-4119-92ce-93c9e0096223 · outbound

This paper cites WebArena : A realistic web environment for building autonomous agents.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments WebArena : A realistic web environment for building autonomous agents

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.036837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.036837Z digest=sha256:0000655caccd926daa1fbe3f522a3ab16235df8d0beeb3d9ba4b44f8f95e1d6c

Observation 2626c290-ef7a-4cad-af9b-d8d16cad648d · outbound

This paper cites PoisonedRAG : Knowledge corruption attacks to Retrieval-Augmented generation of large language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments PoisonedRAG : Knowledge corruption attacks to Retrieval-Augmented generation of large language models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.133085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.133085Z digest=sha256:a8fc16aa967842facbb19b530aaf1523efafe22f95b366221e4b21639ed2d29b

Observation d8f26253-572f-4248-9ecc-71e2ff060956 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.240545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.240545Z digest=sha256:a1af1c94a4128ee06a0d58c9bd0a0e82cabf2d4f570cf3827b8fbc68b0a7f6bb

Observation 24386697-1ff7-4160-97f4-4663c22a6651 · outbound

This paper cites International Conference on Learning Representations,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments International Conference on Learning Representations,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.387792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.387792Z digest=sha256:7de31cbf4b99ffa6866ffcb26e8e92c92d71e16722ccd00742262eb23e38a20b

Observation 94c3416a-d612-4545-addc-461a0ad37a96 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.578812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.578812Z digest=sha256:0ae799fa865af5ade3e7a173dfde17ad8cca2dba627ce8370f5c1908a30cad35

Observation 68c8cb30-307d-403e-be00-026087f6a092 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.715225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.715225Z digest=sha256:f75d23c80a719eeeeca0332ed9f8e777d3542670bca8f1f93f8ac6ff47a9a7a0

Observation ca9f41df-acd7-492f-97c2-308b90aa48de · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Retrieval-augmented generation for knowledge-intensive

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.905269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.905269Z digest=sha256:a4855ad96f8d09035e10189bc52251c2fc3d3ba64b405cc86b90be1d18ffe714

Observation 3a7d9e27-96f2-46f5-8352-3a5782ff7789 · outbound

This paper cites Proceedings of the 2020 conference on empirical methods in natural language processing,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 2020 conference on empirical methods in natural language processing,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.984448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.984448Z digest=sha256:c58328e88592826fe3fcbe531eedcc42ee94f86d3d10ec1f5e5c29a9d4eb748e

Observation f6386fb2-2cc8-46f7-a8a3-7db6374717fb · outbound

This paper cites Advances in neural information processing systems,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Advances in neural information processing systems,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.161291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.161291Z digest=sha256:746b143cc269c2d58fd91a3495e71660eb8d5ba8fa203a9fde5f2cf9d09d1943

Observation cc2ca09c-7579-4a99-8af9-9b2fb4a0fa83 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.212681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.212681Z digest=sha256:24f458d6b83c951abd32c2cbe89b1134814138d55529235e363c18cb82a7e9dc

Observation 9a2fe180-8d98-4edc-ae97-5f024eb42f54 · outbound

This paper cites Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , year=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.300520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.300520Z digest=sha256:508e19a6c92edeebdaca18680d07e13021fd25c7c9760f61ac80fcfd430656fe

Observation 20e2b92d-b6fd-4522-a213-4b8c5f6dbc2d · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.373073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.373073Z digest=sha256:1ded7a86b0b66437d45d2c5b158844396adaf2495c8287cad85821bb6c7d6e78

Observation a9550cff-415e-4372-9de5-a084da42943d · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.456817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.456817Z digest=sha256:b4872f503cb99e8e5871ba10066e8fa3b8475400a74d405ecdd359d4d2790829

Observation 63d6eb31-f9c5-4546-9956-69cc6974e712 · outbound

This paper cites Narasimhan and Yuan Cao , title =.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Narasimhan and Yuan Cao , title =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.517577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.517577Z digest=sha256:727207bc4175b5b6abec28884d09c984704e01310ac05630a2f08d706cbb1453

Observation 04e0d7b4-a457-4e34-8ed1-d6e1da21f615 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.623173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.623173Z digest=sha256:7119224474d2ee0153ba55987fbb1c82170287817060afeae672675092100efa

Observation 98bb099f-f976-4953-b67d-8cf757c8787b · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.703708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.703708Z digest=sha256:5b6f88f67f6be3ffc5eea9724519867f4aa29b4fff94e3f09a45acaeaf73e2db

Observation b88c4ed1-6cc5-4d60-b120-4767088bf492 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.804675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.804675Z digest=sha256:a8a2d72f05a05da8119776bb0f4df5530b396619969258feedfa1b7d97c000c8

Observation cb3e5076-8324-4f9a-b3df-077ce9150a9e · outbound

This paper cites International Conference on Learning Representations,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments International Conference on Learning Representations,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.875427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.875427Z digest=sha256:4f48d09a5230495c5f1297be7ff1a69a9eab11ef6687640bce737e81f14a5e96

Observation 7a4291d9-4e4e-4afc-bd18-a5dab6a63273 · outbound

This paper cites Proceedings of the 2021 conference on empirical methods in natural language processing , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 2021 conference on empirical methods in natural language processing , year=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.932355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.932355Z digest=sha256:a9836f5002f48ce219c66e9bac5e0f9da96cbf7e5fa159b32ea1154c07c6053f

Observation ac8c6a14-45ca-4811-9709-0b2ae0c71c48 · outbound

This paper cites Knowledge conflicts for.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Knowledge conflicts for

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.035692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.035692Z digest=sha256:5b8ab7bbb25976466f2b7a8f87b078b07c59118d82d46d76fa5176b5907b86fd

Observation 3404e851-8364-4f91-b112-042520531c5b · outbound

This paper cites Findings of the association for computational linguistics: EMNLP 2023 , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Findings of the association for computational linguistics: EMNLP 2023 , year=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.153674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.153674Z digest=sha256:eec7396af1aa69ab667565cff3885c89124a2b33aafd38831d168a2caf3bf813

Observation 75d32e88-8c6e-44d6-9d46-3326682625df · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.240303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.240303Z digest=sha256:cefe588753a734b1f781bfdf3cd8b85f5e5d208a78b533ca3bc3c4682668b17c

Observation a0864b6d-45bd-4ac5-93bd-0ea9cc4af7c3 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the AAAI Conference on Artificial Intelligence , year=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.339219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.339219Z digest=sha256:bb2c100b2bfafdf1ad0d77ab3a1c169fe4556949a163a31ca62579dcc82954e4

Observation 7177693f-58f2-4ffc-8829-d85a4fb47a2b · outbound

This paper cites The power of noise: Redefining retrieval for.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments The power of noise: Redefining retrieval for

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.476428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.476428Z digest=sha256:c4fe58390b34b51662a9f9c250b2ca32716e3600b5f635bc47c4bd6d5155a266

Observation cd5f14d6-be07-42d5-976e-efd91c1c88ea · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.593282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.593282Z digest=sha256:9271c1f27fc492130ddabd95fe6759e234fcc502615a57120435ac400aeff1f3

Observation dbd07658-8d80-488e-96a2-de381090796f · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , year=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.755816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.755816Z digest=sha256:bd25028e2bdedb376db8211ffc23b8a3dd1dd3bbe69ae3a284daeddd4d4d8125

Observation 6f227416-7b62-4e30-a457-5d7bb19384c1 · outbound

This paper cites International conference on learning representations,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments International conference on learning representations,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.888810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.888810Z digest=sha256:71b1dc92f059bf08e6672f903923a4cbdd295d024f5c4fc1fd359030aabd49fa

Observation 7c687296-3b47-4cfb-b3f3-0b165ab31ad9 · outbound

This paper cites Advances in neural information processing systems,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Advances in neural information processing systems,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.071638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.071638Z digest=sha256:15d1fa028eac164bd5ee43b53b3070d76c2ff0844cded4046b176ea18ccc8cfa

Observation 0623ac8f-3949-412d-a6c7-c9aa8aa58b40 · outbound

This paper cites Advances in neural information processing systems,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Advances in neural information processing systems,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.191828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.191828Z digest=sha256:f2534039937f4c3b32fecc3cee523fcf5dcdcbe703ca01e8c3c1182715150958

Observation 8095dc41-b30b-49b3-933c-4cafb2caa247 · outbound

This paper cites ACM computing surveys , volume=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments ACM computing surveys , volume=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.349976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.349976Z digest=sha256:2877af67caf1961aee230ccf39052cb8a5aa3ae07c44d2006f723ae3319e82b3

Observation 2c080c04-1acf-480a-a183-5fe5f80c52a2 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.431740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.431740Z digest=sha256:cc7544cfea4a5f15e42c2e4cbf73e8865ca90730ddd6570d7b9d9a1b6a64cca2

Observation b8e53f1d-9d21-43ff-b06f-ac972076c0f0 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.596742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.596742Z digest=sha256:784655d8abd834825e3d09bc80ddbc26e23176a35e72c7792073359b1c5b0b63

Observation 3aa868db-6bbe-43e0-ad86-48085ab17bd7 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.704835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.704835Z digest=sha256:4e6cf269ca9e0a0c8b808b524e08929bb413694541e6aa93b62b49cce93b4710

Pith citing papers

No inbound Pith citation observations are available.