Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2406.10149.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:31:43.210471Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
6
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation dcbe3a3d-3b5a-46fa-b0fa-600ee6d9fdf8 · inbound
RULER: What's the Real Context Size of Your Long-Context Language Models? BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54cc1b36-120f-4524-a2db-2901e5781a05 · inbound
MiniLongBench: The Low-cost Long Context Understanding Benchmark for Large Language Models BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37534588-8791-4d18-a9c2-29caca12fa5e · inbound
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e03c3cd-abd0-49d9-a16a-3c4563e4d3a2 · inbound
IntPhys 2: Benchmarking Intuitive Physics Understanding In Complex Synthetic Environments BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39fbe1fd-1b36-448d-8992-9e86b99e4ae2 · inbound
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 156
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97b852f-4873-4ae1-adbe-de8d3a8c5f21 · inbound
LOOM-Scope: a comprehensive and efficient LOng-cOntext Model evaluation framework BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 902bc66f-9254-4443-bada-97d472af23bf · inbound
Positional Biases Shift as Inputs Approach Context Window Limits BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e42fbbc-7002-4020-9ab2-847602c1be3e · inbound
BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e56728d6-6922-42a7-8896-6aef5a8a41dd · inbound
Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc35383c-c636-4fbf-954f-9f50c968b197 · inbound
Retrieval and Multi-Hop Reasoning in 1M-Token Context Windows: Evaluating LLMs on Classical Chinese Text BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b95822b3-b6f5-48a9-811f-eb0545578443 · inbound
Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fdcd53c-50f1-43f1-a1ff-93b82fc0110a · inbound
Diagnosing Evidence Utilization in Long-Context and Retrieval-Augmented Language Models under Matched Evidence Conditions BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e02aa06-8a04-4e02-9488-c52188e65884 · inbound
Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d286266-e25c-4358-8884-435a42ecc16e · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 123
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 445558d0-253e-443a-afe8-19916fc7558a · inbound
Randomized YaRN Improves Length Generalization for Long-Context Reasoning BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 133c98d7-fa10-43da-8001-1d13b617990c · inbound
What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 550637b4-7d3e-4350-a0bd-dd5be3140afc · inbound
WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0fcf00-ff01-4511-b418-633c4506c10f · inbound
WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1498b5e4-839d-4439-9266-7ebf3bdaad0a · inbound
UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3174dc67-2214-4bee-ac8f-24deb82b378a · inbound
Dropping the Anchor: Statistical Context Summarization for Distributed Systems via Pulsar Attention BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.