Pith. sign in

Paper Citation Record · LEDGER

$\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2402.13718.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.13718 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 35 of 35 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:25:16.984092Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 131b3c3b-b1de-4099-a819-baee0149ff7e · inbound

MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies cites this paper.

MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T18:00:53.511381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T18:00:53.389420Z digest=sha256:232a166248e9b67103800abf52dddbd1b461010aa0dc011f2e1b13cbe777af5a

Observation 61d27b91-65d6-485d-908a-ca05b83339af · inbound

MLVU: Benchmarking Multi-task Long Video Understanding cites this paper.

MLVU: Benchmarking Multi-task Long Video Understanding $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:55:26.502053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:55:26.333923Z digest=sha256:cf16d0c2b0f1ad1283ad5fe32c9c60cf42bd42c9aaaa4a887d9a3546f3f2d6f3

Observation 3e30bb0f-e511-45ba-b73e-943899986e8b · inbound

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression cites this paper.

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:17:31.180588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T04:15:36.906263Z digest=sha256:e3ba6fb6a66b0b4f63b50aef78050f1c31fd52fb0c48cf0b85e541efa9cced50

Observation 613c0194-d071-4128-8a77-b9f1fd881478 · inbound

NoLiMa: Long-Context Evaluation Beyond Literal Matching cites this paper.

NoLiMa: Long-Context Evaluation Beyond Literal Matching $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:33.797188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:33.797188Z digest=sha256:5b4f487fa2dbcb27e572b4a913c636954f7fdc4a65bd2edd6c5f65a1e35517f4

Observation c4e73476-9160-44b4-97d4-1b5ad729b0da · inbound

GSM-Infinite: How Do Your LLMs Behave over Infinitely Increasing Context Length and Reasoning Complexity? cites this paper.

GSM-Infinite: How Do Your LLMs Behave over Infinitely Increasing Context Length and Reasoning Complexity? $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T20:25:16.984092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:25:16.984092Z digest=sha256:d4d81ad22c9ffb94043705fc970afd2108eddd3b2881a778988cccbb57a38ebb

Observation 7df593e2-e375-404e-be22-b30d877a80c0 · inbound

InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU cites this paper.

InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T23:19:37.949529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:19:37.949529Z digest=sha256:e07633697be3f29a6311e60ba283eaed3c8c2251d5d067262bc194f27096d699

Observation d2df9c47-4011-41fa-ab68-c3d47672659b · inbound

Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs cites this paper.

Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-07T20:56:53.962989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:56:53.962989Z digest=sha256:32a1c0b9574816901a7298de06cf4f8667ae550da12266a2594ccd9bf0ff1fce

Observation 743f2d16-f512-428a-a506-87516fafeef0 · inbound

QwenLong-CPRS: Towards $\infty$-LLMs with Dynamic Context Optimization cites this paper.

QwenLong-CPRS: Towards $\infty$-LLMs with Dynamic Context Optimization $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:40:56.019774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:40:56.019774Z digest=sha256:a8e0d5c6c5282c924883a445543e2ed2d4142053ec02aaa3bbc0cc87ef1f66ff

Observation f936a41a-b2a4-4d44-b710-f6820d013140 · inbound

AbsenceBench: Language Models Can't Tell What's Missing cites this paper.

AbsenceBench: Language Models Can't Tell What's Missing $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:13:17.794692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:13:17.794692Z digest=sha256:95b25a16a35a5830ea3c3deb8d9c17518e6aa4f5a0381fbc2e8fc8fe5c213b05

Observation b02d03ba-048e-436f-8d19-bc15d32c8d47 · inbound

StateX: Enhancing RNN Recall via Post-training State Expansion cites this paper.

StateX: Enhancing RNN Recall via Post-training State Expansion $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T12:31:22.135209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T12:27:06.616920Z digest=sha256:58ab6f7def71168c52d63e1db81f7ebb2d442f9fbf6daf5872bede85b00ab22b

Observation 2b6bd92f-2b7f-4389-a7ef-76de11f0a04f · inbound

ChiKhaPo: A Large-Scale Multilingual Benchmark for Evaluating Lexical Comprehension and Generation in Large Language Models cites this paper.

ChiKhaPo: A Large-Scale Multilingual Benchmark for Evaluating Lexical Comprehension and Generation in Large Language Models $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T09:11:58.610538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:11:58.610538Z digest=sha256:1e531c8caa7ce0318e19ea3d31d9e301bc4f7ebd227e393855460ddd3bdba9b7

Observation 0b58d72a-6099-4f09-991f-47c904079b56 · inbound

SnapStream: Efficient Long Sequence Decoding on Dataflow Accelerators cites this paper.

SnapStream: Efficient Long Sequence Decoding on Dataflow Accelerators $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T02:00:39.299754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T01:59:30.583027Z digest=sha256:30fb581f79eb667b952cd7e2faf59d7522480bcd7eaa3931ddf882c5e8ee963b

Observation 1b681f5a-3dfe-4d30-964b-97071f816c23 · inbound

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs cites this paper.

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T12:42:28.024235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:42:28.024235Z digest=sha256:7d22d08e7e2b252b7f2c87bc5b48c3e89ad00a5cf1a7e7f4bdfdea7a6f6e8558

Observation 71ede16f-f364-4be3-9839-37c55f85d1e6 · inbound

MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens cites this paper.

MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:00:10.028746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:00:03.578685Z digest=sha256:5b43f80d142a7e262caec682790c33f5e6cc65efb26612ab574f7d04937a38f1

Observation 475b7349-4c93-4772-b742-c52e4dc60b7c · inbound

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments cites this paper.

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:49.417343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:07:46.077831Z digest=sha256:c6f15f6bbdd75f10e4dcbbfff41229f0b4cf8a7755ca6e337aaec57ef0108031

Observation fb42417b-c1f9-4f90-a1dd-deea1b670594 · inbound

MemSearch-o1: Empowering Large Language Models with Reasoning-Aligned Memory Growth in Agentic Search cites this paper.

MemSearch-o1: Empowering Large Language Models with Reasoning-Aligned Memory Growth in Agentic Search $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:16:20.793604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T06:15:11.432788Z digest=sha256:ac92af033a9915ff5e90325ac5d6286e97da59e91d1097bfb4981f32617f9af1

Observation c1a03221-8d01-4372-86c2-847e83d522a6 · inbound

CL-bench Life: Can Language Models Learn from Real-Life Context? cites this paper.

CL-bench Life: Can Language Models Learn from Real-Life Context? $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:41:27.012435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T09:42:33.635866Z digest=sha256:c6208fc4fcedb9d42a3d54b1d1a964825784de7425983291d50ee4b1379e4c31

Observation ab77f047-e1dd-4cd6-8dd4-c034c6b0419b · inbound

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference cites this paper.

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:10:53.538724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:37:52.545943Z digest=sha256:c22ffbcd28b862bfb4f0916f7bb5b70185f3e2948f16c67262782fc2a1fd4f8b

Observation 4153ef6a-b7ff-485d-960b-2b169f97546e · inbound

LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues cites this paper.

LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:57:12.681055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T03:52:58.177144Z digest=sha256:7f282dae1dd9d510e26a90b03d498b848a8f9fb67d7b8b6d8073237548be785f

Observation 544d91ea-4fa1-46e2-9506-bdf78bd4a8f3 · inbound

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks cites this paper.

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.882013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-25T04:58:15.184063Z digest=sha256:65fda4e903d9a9ab7929a1b0e8c6d838e0236cb0da5eeba817e8685759b60931

Observation af6999a9-f568-4fb8-9ce1-de878c1585fb · inbound

Dense Contexts Are Hard Contexts: Lexical Density Limits Effective Context in LLMs cites this paper.

Dense Contexts Are Hard Contexts: Lexical Density Limits Effective Context in LLMs $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:36:56.798880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T02:00:37.167896Z digest=sha256:c859ebe27f063867608a30030c8213557f9354423c74e252fe54d5d9dd11c72f

Observation b9e625b0-cc9c-47ea-8a77-7a8d71aa27a8 · inbound

MemTrace: Probing What Final Accuracy Misses in Long-Term Memory cites this paper.

MemTrace: Probing What Final Accuracy Misses in Long-Term Memory $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T18:18:49.817849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T03:13:52.804489Z digest=sha256:5015b19e758b75422480ce7b7d312c46927ba73ee4993f57e891eff7f89daa75

Observation 5119249a-f988-4267-8315-4fd4ebd9bd1d · inbound

NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama cites this paper.

NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-03T19:28:52.480993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T01:48:08.794210Z digest=sha256:01bff227a370dbb1cc0d05e1ad298904b46fe37dc55bfdd56a5a1cebfd0fe64f

Observation 740ed88a-fa5d-44be-976f-28f5d03c97ed · inbound

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning cites this paper.

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 190

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.667928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T14:19:53.450263Z digest=sha256:edc59e97bd3288fb3577a79464715f2b78039bd71b2946d2924b9146a4b04bd4

Observation 5b2d1083-a868-4c2c-b9b3-72eed0357df7 · inbound

The Verbose Context Problem in Medical Records cites this paper.

The Verbose Context Problem in Medical Records $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:34:21.563320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:29:42.068362Z digest=sha256:2a137060a65d74c5ba343f5d9b2492d4672397ec4d07e613d46b3f9f36d1a066

Observation cbad6bff-9346-4594-988c-675e45938fee · inbound

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving cites this paper.

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:45.298052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T05:37:13.211613Z digest=sha256:9f5e134b1eb27d38b3b8f4064d7053861e250b7f2820d1b121e2e9c99083bdd6

Observation 2b478166-6003-4193-bd26-c3cbaaa49d21 · inbound

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving cites this paper.

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:55:35.112954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T07:06:53.318182Z digest=sha256:d06ff82f7c741d09e8adb0e4a48480d679616330494b48d7667c55b5fc020bc0

Observation 3c810e5b-db20-4172-a8f9-efac3c8dc984 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 149

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.321053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:c082ed9a5acba77012763c817644fe7aef3239454db46267cc3a2eb362e58f59

Observation 95a74e48-2564-44cb-8a06-3dccc99dfdee · inbound

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning cites this paper.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-13T03:52:24.872919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:52:24.872919Z digest=sha256:8631d5996cec416b330ee69cc813d8932dfd1ee6158c7ce611745b7da1425096

Observation b8133e90-1526-4ae2-bf84-6b0bf42b83dd · inbound

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning cites this paper.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.784842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.784842Z digest=sha256:4bb61d631246217c960dd084cab197181a334ea6cbfa3e7c1edf0a56d7d48a4b

Observation 62a9804a-a428-4246-988f-f67628a57fed · inbound

Where Facts Go Missing: A Layerwise Taxonomy and Per-Layer Attribution of Information Omission in Air-Gapped LLMAgent Pipelines cites this paper.

Where Facts Go Missing: A Layerwise Taxonomy and Per-Layer Attribution of Information Omission in Air-Gapped LLMAgent Pipelines $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T04:48:29.418599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:48:29.418599Z digest=sha256:ea9173396b618e0504c890ab03c85b32f21865a57601b820364c48e806909ac4

Observation 77af0014-4020-42b7-8390-5fb5d7ca707a · inbound

E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios cites this paper.

E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-30T15:06:50.750693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T15:06:50.750693Z digest=sha256:a56f06dc9331381b028ae84e54b1d1d6da10e4db8418597db228ba528f8ebd1f

Observation 47cdab76-80e3-4aed-b820-02c3d39a103a · inbound

Memory for Large Language Models cites this paper.

Memory for Large Language Models $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T02:37:54.442812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:37:54.442812Z digest=sha256:0de103abfcdd9201476fd02f9dbf59821a2ee64f9f2c1b254c885d1cadc0445e

Observation 034893be-4464-47b1-933e-19a3fabebe93 · inbound

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing cites this paper.

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:00.919480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:00.919480Z digest=sha256:b1a4d71ab7187e9f7e8488650c732b6fa0d49a3a7f06adf5f470abb05147cc06

Observation b76b7076-7ac9-4a03-9408-1683c1071703 · inbound

Training-Free Hashing-Based Attention via Binary Principal Components cites this paper.

Training-Free Hashing-Based Attention via Binary Principal Components $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.547366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.547366Z digest=sha256:c9316e9f4a611d9b9e84161e78adf32930456fc7cf8c347cd4cbe68b65784f00