Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:55:14.902112Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 4 inbound Pith citation observations for arXiv:2412.17032.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:55:14.902112Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:19:14.992316Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
57 of 57 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation c3fa33c9-9fff-4ece-937b-ff73f8c67a48 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ccf4b38-807d-498b-ae17-62afec84d6b8 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Learning to Recover Reasoning Chains for Multi-Hop Question Answering via Cooperative Games
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 15774753-2e12-4a10-9856-cf60599fbd4a · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge The Llama 3 Herd of Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04c5b463-0714-4264-a42a-5631638fdc3a · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 500c559b-11b4-4c74-8f0e-2d9a29db9ef0 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06437c8e-b0eb-469b-93d5-ccc896513f73 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Qwen2.5-Coder Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b02736f-ad64-4f45-9448-9412ecd0ec97 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Open-RAG: Enhanced Retrieval-Augmented Reasoning with Open-Source Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd17f135-a0c9-4e03-b441-1586e96e3c1f · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c031ec2-67df-42d5-852b-7bb713bb8a1a · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fbad84af-98c5-481e-bdbb-9f151e4b04d1 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Mistral 7B
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a00bcdd-b499-4755-a9fd-2bb4ff41f467 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Evaluating Open-Domain Question Answering in the Era of Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8851b154-f772-4be2-92ea-668b71633222 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0955d653-7425-44d3-94e7-ab48c022ca2f · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d43c95b4-fa1f-4398-8c3a-b04ee565d55b · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Gonzalez, Haotong Zhang, and Ion Stoica
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb878223-5a60-4e33-b24d-03755f0d0a05 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02b048db-2b15-4816-8818-97996dcbb148 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Large Language Models with Controllable Working Memory
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55862471-7f16-4987-a417-96e3dfdfbf31 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c8441089-dcbf-443e-ace2-e2863a413b03 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6e0d918c-de0a-447f-9e7d-c2d6a31eef0f · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Retrieval Helps or Hurts? A Deeper Dive into the Efficacy of Retrieval Augmentation to Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ef7d9c5f-36a5-43c4-8394-17cb944a296f · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0fa58127-4eac-4b8f-ac01-1566fc47bb09 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2c7b9b02-5790-4afd-8d9b-fe30b38ab018 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Large Dual Encoders Are Generalizable Retrievers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec105f1-b80f-4935-bc89-fca35185e6d7 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Guo, and Xueqi Cheng
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a05b6d6b-a9dc-4058-8d9b-a2fb603eff54 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e030baa-9c5a-4ac3-a45e-0d9f5340b4e0 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Robertson and Hugo Zaragoza
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1afe389f-61dd-4943-85fd-8433934924f2 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Mintaka: A Complex, Natural, and Multilingual Dataset for End-to-End Question Answering
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 219b7b2a-d79e-4d49-a115-73da98a3c1d0 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 968fa1e3-3d56-43a0-af76-4db2c8f9dc94 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5ad18645-bd1e-461f-8481-c266fa38f5a0 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2632d840-75c1-46a7-80a1-555e5e21e3dd · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge One Embedder, Any Task: Instruction-Finetuned Text Embeddings
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 550c42f2-7140-49fe-bea9-de6241d3e40c · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Head-to-Tail: How Knowledgeable are Large Language Models (LLMs)? A.K.A. Will LLMs Replace Knowledge Graphs?
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c5cc6f1-7512-4a28-93d6-cc3f8b5542ca · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd6bcbcd-51f6-4416-ba5d-e533e5e6260f · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Gemma 2: Improving Open Language Models at a Practical Size
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b3bcba-d562-4e67-904f-6eae0c253f65 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8f51e6fb-cd0f-4d0a-8aa2-06f12269e604 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 034e0e7b-aedc-43f0-9c26-b3fce9532a38 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac56400-d0a1-4af8-9c5c-c8eca28261bd · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Self-prompted Chain-of-Thought on Large Language Models for Open-domain Multi-hop Reasoning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06dc3a3d-0616-4137-9848-230da6870e7c · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c120a452-40e0-4e3c-b2a9-211633a44dbf · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d376a72f-5f67-4854-83dc-c55de4e41a2f · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96b9467a-0c67-42a6-bbeb-f1efb39208df · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0cfbe42-b200-4ed5-a25b-255c5e4a5cf4 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b86cd3b5-149c-4dd1-aa52-7d208e5707f3 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge LM-Cocktail: Resilient Tuning of Language Models via Model Merging
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e64a726-be0e-4ea9-8590-bacd53e343a1 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 59b3c53c-3d12-4743-acb5-a38b61c4ffb7 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Answering Complex Open-Domain Questions with Multi-Hop Dense Retrieval
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79a03d29-3262-41a0-b30c-5d87dfe99e0d · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 69b744c9-ad5f-4f5b-ae1d-60f4bf7c28c2 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Cohen, Ruslan Salakhutdinov, and Christopher D
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 07a26606-07bf-4f5e-a3d6-63e8dd4602e7 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cb29f5ca-c30b-4377-8b3e-2a55bd819688 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Yu, Wenhao Yu, Chenguang Zhu, Zaitang Li, Zhiting Hu, Qingyun Wang, Heng Ji, and Meng Jiang
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0004aca3-7805-4f9e-966e-2ea550cf9e33 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge RetrievalQA: Assessing Adaptive Retrieval-Augmented Generation for Short-form Open-Domain Question Answering
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7161233-33ae-4eeb-af2f-c2a0691163f1 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45b8792c-1047-4b93-9b26-99ed99da4b89 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4954c2a4-3254-45a3-9c22-9d1d0b6196c0 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3217fc6-19cf-4db7-b636-3b5574086b6c · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c182206e-d2c1-45f7-a7a2-daa558bbfe16 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge EfficientRAG: Efficient Retriever for Multi-Hop Question Answering
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b4d1837-57d9-4514-a222-7d94ae615d56 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge online" 'onlinestring :=
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4af56952-683f-4c5e-a41d-42f68a081186 · outbound
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge write newline
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eee85ad-94c3-4cb4-99d2-2ce6b8a33120 · inbound
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24d508a4-30e9-4d4d-8781-b77d9e3ff792 · inbound
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c9cb728-9e24-481d-9046-4501a99e1a5d · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e47dcf84-bf68-4a0b-ac11-8d4ce0bb4789 · inbound
FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.