Pith. sign in

Paper Citation Record · LEDGER

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts

As of 7 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.02000.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02000 v2

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:31:45.191937Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bdf90adb-491d-4380-8de7-725ff8812b0c · outbound

This paper cites L-Eval: Instituting Standardized Evaluation for Long Context Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.196733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.196733Z digest=sha256:e722f7c8c512869b6fe2dba77573f1ed3fa6b742629b2c769419b0075d53a62b

Observation 4bfa2222-e7cb-441d-981c-dda2d9e3b3a4 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.981200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.259177Z digest=sha256:ec55ce73c84a3613d72d3a0984e682990d772f8660b1a39e8bddf117e0b1a390

Observation 6a246a3d-f5e8-40bc-a3c9-31178427557d · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.323817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.323817Z digest=sha256:22fb142e1409746b2855ed059e50b6c354697ef5e57652322359d15a446c3c48

Observation 9cca8458-b8f9-41d0-8186-0a909cd0acc8 · outbound

This paper cites Longformer: The Long-Document Transformer.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Longformer: The Long-Document Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.391811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.391811Z digest=sha256:42963cbee55c47daaad3cf461263ff415cfd15d80fce91ed7b2bdde567947ccd

Observation cd6a0f39-8641-449a-b3e6-64e5d9eb0c55 · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.483574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.483574Z digest=sha256:f0d1f4d8006e55a750fbe7eaa24b1cde29d80cde96ca0126f5d5614d1dab456d

Observation f582631c-bdcd-4b32-91b4-cfe7c5b2df20 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.792392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.578742Z digest=sha256:af2706d950bec5113415820a24e805d0ed94e854d7377cd0d01b6d908e513940

Observation 4ca6d699-4ec4-4609-8dd3-3d697496675a · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.484142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.681448Z digest=sha256:986acc37550a4a0455466618b05256d644de17dc21d38794641ed15544664610

Observation 071381ea-b736-4724-b2e0-c4d4e06a518d · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.756912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.756912Z digest=sha256:4bf1125eea96a11f10446a14e85fa128d7082ed18b5ad0aaa07a8cfb6c5ce255

Observation 75c8e2a6-0592-48b2-b12b-00e8ac1a680e · outbound

This paper cites EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.825940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.825940Z digest=sha256:571a148baa575f4b4dbce41a9cede7a2f89cb7b0685fdbf82f33c7b6f7966c00

Observation f1fee9ab-7475-48f8-a9c8-70c0d1062e74 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:42.900418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:42.900418Z digest=sha256:d9735fe13a04ba52353307c1891dfdb0daccaca1acfd6e4cdef98ce2675f2ff7

Observation 06748e38-5412-4ddc-bcb9-28dadad64678 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:48.292313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:42.958223Z digest=sha256:080bc1bff382d63db376701f813123c240a291ac752de4b3121326afdaaddbcb

Observation 37e3e0ab-b28e-4f8d-ab60-48725b56b11a · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.018867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.018867Z digest=sha256:738b983dceb161ac2a9f6d3ed6e46a1d90ac2b7b037d69fee467f0fcbed2874a

Observation 041d8801-d5db-41c0-acae-fa4c3265b030 · outbound

This paper cites One Thousand and One Pairs: A "novel" challenge for long-context language models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts One Thousand and One Pairs: A "novel" challenge for long-context language models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.077816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.077816Z digest=sha256:3ff96aaacdb2d46aca3eede0ba10c7457bc5855ee14f360cac99c9cb3473febe

Observation e0e4f0b7-fa54-45f8-98ff-be9aacc1dd34 · outbound

This paper cites The NarrativeQA Reading Comprehension Challenge.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts The NarrativeQA Reading Comprehension Challenge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.137091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.137091Z digest=sha256:e0c5ce0ebe067168d021cd348c08094b1ddba7de9da7d47d1cb61ca37a8a838d

Observation 37534588-8791-4d18-a9c2-29caca12fa5e · outbound

This paper cites BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.210471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.210471Z digest=sha256:f429759e8da473557dd156803daa6dba80dce0d98be4d0ec3b94c803cad49103

Observation fbefefc8-41f3-446a-9377-e2ec961d736b · outbound

This paper cites Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.267912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.267912Z digest=sha256:1cc23032920ac98c30daadc91f5b0428f6a63303538ee62f5d9d03c3827543c3

Observation 15cd5cb7-6c85-4d8e-998e-4bf9a0f4ecde · outbound

This paper cites LooGLE: Can Long-Context Language Models Understand Long Contexts?.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts LooGLE: Can Long-Context Language Models Understand Long Contexts?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.324780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.324780Z digest=sha256:4aff2a31c7a1b80d628c8a6dcc7dbeb34aff24e965fa3e1a240f70c1fd343031

Observation 82393d29-4155-4ee5-88d5-97d148794423 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.379678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.379678Z digest=sha256:377a22196b7f455221be3cc9f22cf528dc868f22c2b66f940949b602b8016169

Observation 48ac52bb-5647-451b-98ea-0052fb0bfbf5 · outbound

This paper cites Making Long-Context Language Models Better Multi-Hop Reasoners.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Making Long-Context Language Models Better Multi-Hop Reasoners

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:31:46.057841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.438051Z digest=sha256:f9fdd8c64617423c25b10734b058056b040f15c163c85dacb817ef91978b1ad7

Observation 83efaa79-557c-479d-8e13-c36d6c8758f4 · outbound

This paper cites Lost in the Middle: How Language Models Use Long Contexts.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Lost in the Middle: How Language Models Use Long Contexts

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.593760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.593760Z digest=sha256:9d747a34bb2b05423ef4502318e9d54beb36f7b947aff4b29a1b0bc998a26b39

Observation dd451ad6-6fed-415e-bc01-b57e1d81c82e · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:47.972259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.665498Z digest=sha256:fa609ee9929e0a804ce524cbac37e166f32dc71d690ff5e4f47395b4808874e6

Observation f01e73c1-716e-4e4c-a9cd-3dde364aaf0d · outbound

This paper cites GPT-4 Technical Report.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts GPT-4 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.724505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.724505Z digest=sha256:952df95c1fbea97860f83e69de15cc98ecf627b604783b31e77ca6b02eef7365

Observation 67c01229-56bb-4924-a0dc-e91351698c36 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:47.557596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.810014Z digest=sha256:8d54c73f8a3a8142f67a5a7e3ea6492d64cbf6fc329c3344f7a4ee8a5e8e790b

Observation de9afda2-7828-471d-a59d-7f08b679dea1 · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:47.160960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:43.875143Z digest=sha256:17baaf925e99fdda251c7d06d6fecb85576b71810874b56e9298a8d4955e59c1

Observation 7ac20032-db84-4839-b3d8-6f59f3fbc621 · outbound

This paper cites QuALITY: Question Answering with Long Input Texts, Yes!.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts QuALITY: Question Answering with Long Input Texts, Yes!

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:43.967050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:43.967050Z digest=sha256:f4fd03a100ba89a18cdd3199f771a978e3206ef4ffd28db23457e5e57a4245a4

Observation 25f096e5-05cd-4146-bdfb-a0e5a630a30c · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:31:46.779048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:44.038710Z digest=sha256:818f138cc263ab782c740ddff46f5bfa3627c0450c5c1af9599aa235849707a1

Observation 2ef4bcdd-f7fc-43c1-bf0f-47719216ef2b · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.114430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.114430Z digest=sha256:49759bfd7a31817b7492ff6b0ccdb8e0381145e9827852e4e2e63ae03ee903ff

Observation cbd023b7-282d-4cc8-8988-8356a1ee77bf · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.178221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.178221Z digest=sha256:84bef4bf106fd1cbdb6ec628d1f3ed570b28943871920008641ff14532bd6220

Observation 4b1bcb33-ba5a-42d5-a058-b33a85c2f86c · outbound

This paper cites Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.194237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.194237Z digest=sha256:5bc55f4e84fad29033b5fc3127cc189b515645a9c58afc9e1b28a3a2ddaecd9f

Observation b9581f1e-2f10-4153-8766-480f809f0ba6 · outbound

This paper cites Constructing Datasets for Multi-hop Reading Comprehension Across Documents.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Constructing Datasets for Multi-hop Reading Comprehension Across Documents

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:31:45.621660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:31:44.246400Z digest=sha256:bdbc2221236cb2ac2cad9d85e9a85aaea0a63f06ff1faddd4842457777c958e5

Observation 5ff3d9e1-8d88-46ad-9254-180582f74010 · outbound

This paper cites Memorizing Transformers.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Memorizing Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.312581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.312581Z digest=sha256:b5b1aef7dbfcceeeded9cc5c3eaa70be4cd888024a01c46192ce71dd42822149

Observation 56785ef9-37ba-433d-8bfc-7cc9266b1d39 · outbound

This paper cites Retrieval meets Long Context Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Retrieval meets Long Context Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.407153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.407153Z digest=sha256:372e85cc35b18225260512e9c28f193cfc954f3abcf79d3e3437ff731fe2956d

Observation a4e40009-68b7-40be-bbd4-977788e6046e · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.488966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.488966Z digest=sha256:32cb07634d583a0436242a16f627978908bedd2e2530d96530745534905198e3

Observation e8e23118-a6f0-417a-ba76-045f36f4d1ea · outbound

This paper cites an unresolved cited work.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.581369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.581369Z digest=sha256:22b1bbdc6aeb88b078d649ba0de1d7c3df52f10c2d065e7d53040253bb856546

Observation 6dd3275c-d11d-4ba1-a9c9-f1f20bed4d6c · outbound

This paper cites Big Bird: Transformers for Longer Sequences.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Big Bird: Transformers for Longer Sequences

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.720004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.720004Z digest=sha256:e959c09e7e7ebe4bd42de6af8e2de44d508e88e3adc517c040928de0814cfa9f

Observation 72a004f4-ea60-4886-ac44-e919bf40462d · outbound

This paper cites Marathon: A Race Through the Realm of Long Context with Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts Marathon: A Race Through the Realm of Long Context with Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.836384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.836384Z digest=sha256:29a599651135ed471487e7d1d14fc9fa17cbf170c70ace4d353841669f52b45a

Observation 9019518f-0e40-4776-96d6-5b58b9a00648 · outbound

This paper cites FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:44.948205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:44.948205Z digest=sha256:cc982a4fb3264cf7ba9f77677b92a06de76f3f41521be65fb051277beb7f46ee

Observation 6401e0da-9b36-47da-ab4a-61d95f037865 · outbound

This paper cites online" 'onlinestring :=.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts online" 'onlinestring :=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:45.066367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:45.066367Z digest=sha256:623052381558ee41f78fd5232e287a68864978055c537489b5b8922a3230e508

Observation 5adbf85b-d3d8-4f0e-9ad2-0e3d1f172064 · outbound

This paper cites write newline.

NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts write newline

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:31:45.191937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:31:45.191937Z digest=sha256:0fe406126b084e951d6803fe7bf5b1e6faf3b132ca7d58ec542814c0524cc721

Pith citing papers

No inbound Pith citation observations are available.