Pith. sign in

Paper Citation Record · LEDGER

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments

As of 17 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2607.17291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.17291 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T18:30:54.704835Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a302ecf0-7aab-4957-86e5-b71a8b4ceaa6 · outbound

This paper cites Self-RAG : Learning to retrieve, generate, and critique through self-reflection.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Self-RAG : Learning to retrieve, generate, and critique through self-reflection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.362916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.362916Z digest=sha256:5b588791a5ec51101562d428d2e24659f0cd3faaa52093de2f868f137f9b3257

Observation 3fd44d94-8aa0-4cce-b1b0-7abd5885881f · outbound

This paper cites Benchmarking large language models in retrieval-augmented generation.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Benchmarking large language models in retrieval-augmented generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.407939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.407939Z digest=sha256:9537449b343dcdcca6a72797e6bf2ceb9a282ede8968f556e3a3536dd4a4c1c3

Observation 087a70de-e1d6-43c4-8c3e-a932de9f66b9 · outbound

This paper cites BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.495212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.495212Z digest=sha256:ba25d7038979c7a0e650040fc2379ae68a6889e858527cd3c0f3a57c6a546a0d

Observation 1a102dd7-26a8-4f18-87f4-d76c736dd404 · outbound

This paper cites The power of noise: Redefining retrieval for RAG systems.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments The power of noise: Redefining retrieval for RAG systems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.576961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.576961Z digest=sha256:636b9b02e6c19d20c3390fd93d7a38f2d7a9dd226b74ca9ecdbc8b56615252e8

Observation d3fce4c0-e569-48d4-b4d2-3d5836a235ea · outbound

This paper cites InteractComp: Evaluating Search Agents With Ambiguous Queries.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments InteractComp: Evaluating Search Agents With Ambiguous Queries

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.639364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.639364Z digest=sha256:da32938e7dd546bf30d79591cdc5dc5859d0b5a9d8d2bce6077d1a8c55d8c0dc

Observation 8a21d8bf-4f7c-4c0b-a1ac-e2b9325ba385 · outbound

This paper cites Mind2Web : Towards a generalist agent for the web.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Mind2Web : Towards a generalist agent for the web

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.696841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.696841Z digest=sha256:8add94459b01320428883b6d0aa3d00b96b2bf4518331bdcaf5a521ab27dd379

Observation 31a707f2-89e4-442c-8cce-3d0aa584e0d6 · outbound

This paper cites Enabling large language models to generate text with citations.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Enabling large language models to generate text with citations

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.790279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.790279Z digest=sha256:af84ce8fadd71f45e141c88cb1a88611c486f5eb040d13ef0888483d2dc27c84

Observation 47de9def-c6ef-4b8c-bbe6-307e3b9653cd · outbound

This paper cites Large language models cannot self-correct reasoning yet.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Large language models cannot self-correct reasoning yet

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.871204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.871204Z digest=sha256:73a7318efc4d0bcf3a7906caa79154bd74aaa9d858ea68d0e6a63b4a1f1558e8

Observation e4819322-b758-436c-be22-9509f0de1a39 · outbound

This paper cites Survey of hallucination in natural language generation.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Survey of hallucination in natural language generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:48.946310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:48.946310Z digest=sha256:5dfb2eacb1ba10d167c627a9039b74b045f1b4f3755d0887918cfe738744bae2

Observation fc3eae84-44d1-416b-b83f-9c6812745797 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.047071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.047071Z digest=sha256:624e067fd6cf479a7bca422a7de0de3633850dba87c7a1e0e25b83e1c06b6eec

Observation 4ebc4be2-8a1e-4196-a517-ad28fa53ec7e · outbound

This paper cites Dense passage retrieval for open-domain question answering.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Dense passage retrieval for open-domain question answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.098388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.098388Z digest=sha256:16522ad0250df6e95162bb8135dcf470f25b46f26c4c32855292bd5c51fde664

Observation ccc8f459-a55a-4a6f-a86a-6967a6d413d2 · outbound

This paper cites u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.138565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.138565Z digest=sha256:8177a42fd97aed35b95ce0a48039f3a283d2d16053c8ad3922a48a6161d6ab87

Observation 5027daa3-2cb5-47ee-a96f-15307575cdc8 · outbound

This paper cites Entity-based knowledge conflicts in question answering.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Entity-based knowledge conflicts in question answering

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.195723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.195723Z digest=sha256:aca13307d8ca83241ea889f5f27b4ff254171a9a9bda22b2e4249bbe13cbb9b5

Observation e00d7f77-85f8-49b9-945a-30f86d347030 · outbound

This paper cites Self-Refine : Iterative refinement with self-feedback.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Self-Refine : Iterative refinement with self-feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.276903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.276903Z digest=sha256:c3ac1dc2b9c3e451c13cac9387a56ad676e64ba9a42b1e1573b080df040a5ccc

Observation a784d6c6-357c-4f6c-8b73-e41b9b9789b6 · outbound

This paper cites GAIA : a benchmark for general AI assistants.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments GAIA : a benchmark for general AI assistants

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.366500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.366500Z digest=sha256:88803a68df1dcc2c31ffccbcbf0764bbb3871c82ee2dd90233e685754b1c5a56

Observation 27e21187-0083-4d61-904c-192dac824c12 · outbound

This paper cites FActScore : Fine-grained atomic evaluation of factual precision in long form text generation.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments FActScore : Fine-grained atomic evaluation of factual precision in long form text generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.441867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.441867Z digest=sha256:09310063e2c134e4d126b298a319ba736c270b33a8214091b28a3582cbf8240f

Observation b3026955-af08-4d11-bd8e-c87f2e9fd5bd · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments WebGPT: Browser-assisted question-answering with human feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.526685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.526685Z digest=sha256:ed396505ff7e4b10f7f60f64a7db98bd5571be5a9489500c720f08d1b935c802

Observation 5c6604d0-5cbb-4039-b99e-476e3d3a52ab · outbound

This paper cites On the risk of misinformation pollution with large language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments On the risk of misinformation pollution with large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.617131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.617131Z digest=sha256:62b7fa8802032516c0e908a548b39579b30c51a28d536bfebbf404f8bf668eee

Observation 9f9430e7-baa4-4a53-bc03-722d8c1dd1ad · outbound

This paper cites Tell me more! towards implicit user intention understanding of language model driven agents.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Tell me more! towards implicit user intention understanding of language model driven agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.694401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.694401Z digest=sha256:0ec3620ad22ee9a0fc2d5c7ff39a65c3f7832dd390f893915af322a0dd6b590a

Observation cc042a88-9acf-4e27-8c21-6e5899611d24 · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Toolformer: Language models can teach themselves to use tools

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.757617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.757617Z digest=sha256:b37484ad444f1ccae67637a985609c9308e3912054bdd328bf6a108103984488

Observation 7e865149-129d-42d3-8c6f-7ef6bc5e22c3 · outbound

This paper cites Towards understanding sycophancy in language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Towards understanding sycophancy in language models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.859890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.859890Z digest=sha256:ca337002660fd117a85a99b787c029ab0ea4e61459211286ab9d0f080aa64b9e

Observation 46d162ac-b9c6-4384-8ed9-d93c46e01de2 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Reflexion: Language agents with verbal reinforcement learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:49.980983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:49.980983Z digest=sha256:e19e94b16e0e4e3f70ecb9aca617b2c6dd356cf77eea1ed3f874b17ac48f1f44

Observation 8abf6510-6d49-450f-920f-5033fb5b907c · outbound

This paper cites FEVER : a large-scale dataset for fact extraction and VER ification.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments FEVER : a large-scale dataset for fact extraction and VER ification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.162196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.162196Z digest=sha256:ce47a522c52684544fc3bfa6ca7151c1c96ec5215b0517794cf6d08572c1b734

Observation fcc34e7a-1047-4684-a31a-a9b8d2d34290 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Chain-of-thought prompting elicits reasoning in large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.315268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.315268Z digest=sha256:c482b7ea84cf675ae082418285ac1ba28fb6d5437bf0ff510a68d8a36807c179

Observation af24c92a-46cd-4183-a95a-4a974bf07c1f · outbound

This paper cites BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.419774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.419774Z digest=sha256:83dff24e31609b49e29ca726dcb05c9c98fe433c7005933d829d256a86242883

Observation a1ae14c5-955d-4c88-8cac-20e311514e66 · outbound

This paper cites WebWalker : Benchmarking LLMs in web traversal.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments WebWalker : Benchmarking LLMs in web traversal

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.539216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.539216Z digest=sha256:6b55e57d1b30d689c5af513569b6cca95c27975c8f7c801132727fff60c637e1

Observation 118c5f78-2fb3-470c-912b-082c0aaaca96 · outbound

This paper cites Knowledge conflicts for LLMs : A survey.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Knowledge conflicts for LLMs : A survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.662093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.662093Z digest=sha256:4c1e31a8fee08e0f3cae692527f1c1c3627c9f1c4d744738b2de3b4b4dc78e65

Observation e9b249ed-0017-4c05-82c0-216a1b371d4b · outbound

This paper cites Narasimhan, and Yuan Cao.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Narasimhan, and Yuan Cao

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.786556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.786556Z digest=sha256:c605cd78c302221acf92471b0ff1bbeb3a5fc2f121488071e151494377bf56a0

Observation 5fbf085d-a3d0-4284-a062-5189c386e136 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:50.883343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:50.883343Z digest=sha256:ee543c47d9450ac7ebd16321890a37587b25b8e3166adb8240d732d4ae524746

Observation 14afcf14-edf6-4119-92ce-93c9e0096223 · outbound

This paper cites WebArena : A realistic web environment for building autonomous agents.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments WebArena : A realistic web environment for building autonomous agents

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.036837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.036837Z digest=sha256:cf547dbb67a7b1f32df86a13ee4b2154fcacf4b4c7d59be509881ae79ec5e241

Observation 2626c290-ef7a-4cad-af9b-d8d16cad648d · outbound

This paper cites PoisonedRAG : Knowledge corruption attacks to Retrieval-Augmented generation of large language models.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments PoisonedRAG : Knowledge corruption attacks to Retrieval-Augmented generation of large language models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.133085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.133085Z digest=sha256:ff0c4ecfc541c713f8d19cb1c8792cff74390fba9bb6270085b45d47ef6f6316

Observation d8f26253-572f-4248-9ecc-71e2ff060956 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.240545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.240545Z digest=sha256:23e63c71504c3db3cd5bc2db2c2008c99551afc44c2d0ac55e561c6fd6dd2e9b

Observation 24386697-1ff7-4160-97f4-4663c22a6651 · outbound

This paper cites International Conference on Learning Representations,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments International Conference on Learning Representations,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.387792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.387792Z digest=sha256:1bea077dc3f76833a62a5f49ca542e40ef807a9ca115415922f501422edd8767

Observation 94c3416a-d612-4545-addc-461a0ad37a96 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.578812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.578812Z digest=sha256:2c532298f39b9329aa0aaa46f91b42dd31a0bffbd514badd3b53537e6789a121

Observation 68c8cb30-307d-403e-be00-026087f6a092 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.715225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.715225Z digest=sha256:0c0d971ca0860a69d22e80de27d3beb5a18d0807e769da714943fd8b0c03fbea

Observation ca9f41df-acd7-492f-97c2-308b90aa48de · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Retrieval-augmented generation for knowledge-intensive

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.905269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.905269Z digest=sha256:1c7985b9cda75f579e0f501d2ee059ee6f332876db91936381f9b86269212c22

Observation 3a7d9e27-96f2-46f5-8352-3a5782ff7789 · outbound

This paper cites Proceedings of the 2020 conference on empirical methods in natural language processing,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 2020 conference on empirical methods in natural language processing,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:51.984448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:51.984448Z digest=sha256:4409108ea4e1c57523451a933437e81057e286e3bc0525a646176432d9abcb9a

Observation f6386fb2-2cc8-46f7-a8a3-7db6374717fb · outbound

This paper cites Advances in neural information processing systems,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Advances in neural information processing systems,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.161291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.161291Z digest=sha256:1e4f72d0f78c27c50d8d4a2ce0ef94663d5efa102fa0bc09c7c2d1b234083ae7

Observation cc2ca09c-7579-4a99-8af9-9b2fb4a0fa83 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.212681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.212681Z digest=sha256:7f55b4ef7aef6c2acfc5576b963d7d3bc62480cdcf257457c0c00e44c2b4326e

Observation 9a2fe180-8d98-4edc-ae97-5f024eb42f54 · outbound

This paper cites Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , year=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.300520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.300520Z digest=sha256:4917fd9b2c8b84a413a1dd922f23f40b92392d7d371d8b304fa79b35c9050f82

Observation 20e2b92d-b6fd-4522-a213-4b8c5f6dbc2d · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.373073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.373073Z digest=sha256:5ff65aa05e773a1a6bb2e43dac97759ab25341270a79c30eb606025c521ba722

Observation a9550cff-415e-4372-9de5-a084da42943d · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.456817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.456817Z digest=sha256:c2f50cb673982f5f596702da2bced70b248cb7e7b68e468623fb8518dde4580e

Observation 63d6eb31-f9c5-4546-9956-69cc6974e712 · outbound

This paper cites Narasimhan and Yuan Cao , title =.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Narasimhan and Yuan Cao , title =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.517577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.517577Z digest=sha256:5b811ce958fa96177947bf530310a1f3323519f9925b85c2588c25e90188c8b5

Observation 04e0d7b4-a457-4e34-8ed1-d6e1da21f615 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.623173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.623173Z digest=sha256:92c97a00d75a2e2251525df33c6e62375980003de70e6f85c578804548c2449f

Observation 98bb099f-f976-4953-b67d-8cf757c8787b · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.703708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.703708Z digest=sha256:19c57631bdd27c36af34c3a4a3a86fc4defe34b654cae32d1028d0d89b167be2

Observation b88c4ed1-6cc5-4d60-b120-4767088bf492 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.804675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.804675Z digest=sha256:4b0a7e95a19d2929a66d778eb8d3f4d8853694cba545454780718ef43e29b266

Observation cb3e5076-8324-4f9a-b3df-077ce9150a9e · outbound

This paper cites International Conference on Learning Representations,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments International Conference on Learning Representations,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.875427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.875427Z digest=sha256:4a0182e8d36d3e6483d64e943e08078b528f0db8d76d5ddc3e852140f1b18813

Observation 7a4291d9-4e4e-4afc-bd18-a5dab6a63273 · outbound

This paper cites Proceedings of the 2021 conference on empirical methods in natural language processing , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 2021 conference on empirical methods in natural language processing , year=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:52.932355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:52.932355Z digest=sha256:b1730b203093c7d0e2f1165f167a95dae4dab79d47a640c498b48d59a86bb019

Observation ac8c6a14-45ca-4811-9709-0b2ae0c71c48 · outbound

This paper cites Knowledge conflicts for.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Knowledge conflicts for

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.035692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.035692Z digest=sha256:24081fa5bd75d2a92d3b4e45ca8d8b8dabeda50fc9aaf19de249c030bc17e33d

Observation 3404e851-8364-4f91-b112-042520531c5b · outbound

This paper cites Findings of the association for computational linguistics: EMNLP 2023 , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Findings of the association for computational linguistics: EMNLP 2023 , year=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.153674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.153674Z digest=sha256:9c70971fc198337755d2451bb895e813c035fe73e084662f5d6e762e62e6d19d

Observation 75d32e88-8c6e-44d6-9d46-3326682625df · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.240303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.240303Z digest=sha256:1cc6b1b7e980521a04a31e613358241ef8b32a6be90a091c5f795d04a71c5d07

Observation a0864b6d-45bd-4ac5-93bd-0ea9cc4af7c3 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the AAAI Conference on Artificial Intelligence , year=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.339219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.339219Z digest=sha256:884b25ea8f9eb3dd71aff85555222f96f5c0e48e98760d72ba6c430d9a5519c2

Observation 7177693f-58f2-4ffc-8829-d85a4fb47a2b · outbound

This paper cites The power of noise: Redefining retrieval for.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments The power of noise: Redefining retrieval for

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.476428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.476428Z digest=sha256:df3427cf8427c0d57acd040dd3306ae9f877263e8da0d6d6ab13e072e40246a4

Observation cd5f14d6-be07-42d5-976e-efd91c1c88ea · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.593282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.593282Z digest=sha256:8f43517d73131d938305484a675a563977216d5876c696bb7b8f029d092283df

Observation dbd07658-8d80-488e-96a2-de381090796f · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , year=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , year=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.755816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.755816Z digest=sha256:94122e14f09f8891842c6a04291293e3e2f83ecbdb570ea9eada1467fe16ae7f

Observation 6f227416-7b62-4e30-a457-5d7bb19384c1 · outbound

This paper cites International conference on learning representations,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments International conference on learning representations,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:53.888810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:53.888810Z digest=sha256:7b43324b68cf5cfc90efe4a5e0a3f39aff561f1071d55ef1a869dd1ef4d56e8d

Observation 7c687296-3b47-4cfb-b3f3-0b165ab31ad9 · outbound

This paper cites Advances in neural information processing systems,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Advances in neural information processing systems,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.071638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.071638Z digest=sha256:0900e708a6864aa54af83a51cbeee9d5b455adb6f5efd7412602ab07f5a1a1bc

Observation 0623ac8f-3949-412d-a6c7-c9aa8aa58b40 · outbound

This paper cites Advances in neural information processing systems,.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Advances in neural information processing systems,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.191828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.191828Z digest=sha256:dc87b91dc1a3085813edfa4a626f1b658b978761ce21b523e43883d5c185f52b

Observation 8095dc41-b30b-49b3-933c-4cafb2caa247 · outbound

This paper cites ACM computing surveys , volume=.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments ACM computing surveys , volume=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.349976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.349976Z digest=sha256:9600bdd3469e7098b46651f43bb7a4d6bee7ce21140852c32fc0f9a14090f276

Observation 2c080c04-1acf-480a-a183-5fe5f80c52a2 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.431740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.431740Z digest=sha256:ceaaa46ed4e40b66891e672fd0e8b8380132f3ae7c7c42ac9fbe2a394b869465

Observation b8e53f1d-9d21-43ff-b06f-ac972076c0f0 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.596742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.596742Z digest=sha256:dd6afcb513291a9a39c90b050b68e122da682000e011845c430ffc7e3b30c26c

Observation 3aa868db-6bbe-43e0-ad86-48085ab17bd7 · outbound

This paper cites an unresolved cited work.

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T18:30:54.704835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:30:54.704835Z digest=sha256:d63c6bf2ca470e062c9e0e0d51a3ed558e0390b58502661a14624b6741af28c1

Pith citing papers

No inbound Pith citation observations are available.