Pith. sign in

Paper Citation Record · LEDGER

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

As of 6 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 54 inbound Pith citation observations for arXiv:2412.15204.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15204 v2

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T20:38:00.338445Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 54 of 54 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:54:05.362827Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 9bd087a9-6480-4184-ab5d-b9181a858396 · outbound

This paper cites Many-Shot In-Context Learning.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Many-Shot In-Context Learning

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:38:00.357547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:506b1a129ac9f2b811723725c2fa8e035f4a1974943f78fc98095105534e2f46

Observation eadd456f-d5d7-4377-8c15-64190b4cb5d9 · outbound

This paper cites The Llama 3 Herd of Models.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks The Llama 3 Herd of Models

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:38:00.362970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:2a6ba886a001141e8e412f024aaa1a792e316e201ed3f1e626212f085fa13085

Observation 407ae64c-aa4b-490a-9106-401b6a192e4a · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:38:00.367453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:fd8c48bf13bf47391d97998656c2f7fd11497e509f2d5f0451aa4a09243e1d55

Observation fd7e5352-4d2a-4c98-9c29-af78f90840a5 · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:38:00.371198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:05f11029c1060054a98fa0b951137fe40a921cb7618cda4bfc4212aacf315c21

Observation bf697d1e-154f-4238-ba11-f81bf75126d0 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.374853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:630f5b6408a8e16761f487586458ea3b163eed21b100081ae6ffdd1b033e837e

Observation b0638deb-39c7-4d48-8b3d-1cd49a12a48f · outbound

This paper cites It is recommended to change such questions to listing all elements.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks It is recommended to change such questions to listing all elements

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.377828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:0cdd03200b34693e9ed0de1fad4ecf9681d28ff97037834a6bb0bf8b4c574137

Observation eeebc7e1-cc64-48f3-a8d9-20985075de93 · outbound

This paper cites an unresolved cited work.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-16T20:38:00.380433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:df625feeed2cccd2e621cd56ec69e3c2896e58ade446718d7b390749df03402b

Observation d634e646-7ad8-416a-a317-795a983784ca · outbound

This paper cites It is acceptable if it only requires common sense or a small amount of professional knowledge.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks It is acceptable if it only requires common sense or a small amount of professional knowledge

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.383013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:4e06f8b9d2f6ffa9a25555408c340a9e5ed87f6e112cecf4613ae0cd8f2332dd

Observation 5631ed34-2658-4ea5-9dcf-cee3ed417a10 · outbound

This paper cites Questions should be more natural, try to be close to the real needs of users’ questions, and should not be deliberately set to unreasonable challenges just to increase difficulty.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Questions should be more natural, try to be close to the real needs of users’ questions, and should not be deliberately set to unreasonable challenges just to increase difficulty

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.385927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:299af87f32ce0d1cd12601efc60da02b52f4dedefc67c43fdd66048dc892e3aa

Observation 542ae6d4-e3f4-4922-b382-a725897d17fe · outbound

This paper cites an unresolved cited work.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-16T20:38:00.388335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:e78b6d5c52d53ec761ea2a23f702c84cf89af6db71b7d5ff82a46a37f3846c39

Observation 3204d68e-2adb-4dd4-9da3-3c3e6a0e3d4d · outbound

This paper cites an unresolved cited work.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-16T20:38:00.390827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:554571a2b980dde82e6021508e2cf4cffa9599e5ceec85505e7f6d799459221a

Observation 38c1b1e3-895c-46ef-a347-972a3274895c · outbound

This paper cites If all models answer correctly, the data will be disqualified; after passing the model’s automatic test, we will have human reviewers answer the questions.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks If all models answer correctly, the data will be disqualified; after passing the model’s automatic test, we will have human reviewers answer the questions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.393357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:7b7bbb5d2712967d30115b00f7d60256c77e5c635d6a4d06da1f1f4e17cfecb7

Observation 1c75ac3c-5a86-494b-a18e-36cc319a208e · outbound

This paper cites Data Annotation.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Data Annotation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.396201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:fd6899723175a3c2d3598a9959e5deee3edbd06a13d1ab49681157b72cfe5621

Observation 5e9113f0-888a-4cf5-a605-5ea9758d3e6f · outbound

This paper cites Data Annotation.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Data Annotation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.399002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:3f8d4949cff3d8e4936c909a509aded2f35cbfe0ef1934783ac35b22dbd09f28

Observation 0ba3804c-c2ac-4fba-a022-5c95746657f3 · outbound

This paper cites Upload Files.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Upload Files

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.401546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:61a5ad368401c90211013ce9d945eec2c99c205421f5cab0e1db6cf5433a4667

Observation 55c9578d-7a38-4893-9677-9871209179c9 · outbound

This paper cites Evidence.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Evidence

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.405707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:71db20bb02173e16b7d19dcd0e1ab667f6fd0839227793d4c7c4bb93a11902e0

Observation 1c67684f-48a4-47cd-9d27-d267cac806df · outbound

This paper cites Submit” (you cannot submit if there are blanks), and you will see the status of your submitted annotated data in the “main.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Submit” (you cannot submit if there are blanks), and you will see the status of your submitted annotated data in the “main

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.408755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:efea1785f50154f41f5eeed289bc61e3d877f3e7cd8785fde76823095e09a9ae

Observation c4c87656-bf69-4693-a79d-d8c80c1bcdf3 · outbound

This paper cites _id” of the original data in the “Modify My Annotation.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks _id” of the original data in the “Modify My Annotation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.411518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:e8d50094d0eb7ec6a63d1c1bc85ff96f28aee007fa647167c548555b2dfd728a

Observation 42a1de28-a1f2-43b5-8db3-9ffcb8c0e430 · outbound

This paper cites Guidelines for the reviewers, displayed on the data verification page.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Guidelines for the reviewers, displayed on the data verification page

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.414246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:a6d7c22a0292a1f81923d5697ba69acd4de1a30768d5d3a48ba1a39326eeb3ed

Observation 2371de70-773b-4d13-9fca-79bf01aed4a8 · outbound

This paper cites Data Verification.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Data Verification

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.416903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:e3b7c9e677c2ccda376993be3fad37954b043c7f408a92f0ed6424029894297a

Observation 879f0d55-06b6-42c0-8d29-a2e271c61856 · outbound

This paper cites Start Verification.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Start Verification

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.419410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:431b4a31616e13ddd1cbff1298790b9f6c9b3573fce716037da1f5ea04b6ee5e

Observation 49f8f08e-d397-463d-ba0c-45b60a8dce9b · outbound

This paper cites Submit Verification Result.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks Submit Verification Result

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.421683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:86147b4b039282f8706b4c6937395a78d471c8cce08c89cc61aa7dde8875375e

Observation a86342a5-25be-4bb3-8096-b627ebe48525 · outbound

This paper cites I don’t know the answer.

LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks I don’t know the answer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:38:00.424196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:38:00.338445Z digest=sha256:9c1d282799d5d83cc7724d5c84fa2a72236e59f7c1fc9df10d14c288aaf2b212

Pith citing papers

Observation 87065359-f7c8-4d12-b1bd-c756b2f47da0 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-23T17:35:44.233236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:e6ad08ae2234fab78cb6f1d2ef1265c875fb6b8a8aff9a4d3cb3b2fdf4bf1810

Observation fc7823f8-f28e-4b38-997c-1d039f92ef9f · inbound

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression cites this paper.

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-23T04:17:31.206056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T04:15:36.906263Z digest=sha256:dca9334db2c51fe065b640b6b9a0891162eb1f543a54bbe59b26777b62e4bc94

Observation 5b3e7e35-31d4-4657-b486-35d5ac92efa7 · inbound

MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention cites this paper.

MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T09:28:16.189617Z digest=sha256:169f009e9f26c14db216fb589747cc5a8dcc36b2539abe04fbf154e354299ec4

Observation b8b52979-3ccf-4acf-89ae-ff4253ffb335 · inbound

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions cites this paper.

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-16T21:20:22.214060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T21:20:22.146417Z digest=sha256:26ceb3c1e227848894a212e234279fe2e3684718d5a80a22349ab5dc099c9da6

Observation a6e6f269-a5f5-471b-bf80-2261e7cdf64e · inbound

Kimi K2: Open Agentic Intelligence cites this paper.

Kimi K2: Open Agentic Intelligence LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:49:27.926646Z digest=sha256:fe68e86d2c6fc96653413b628a21d9a0b3357a81dfe1baeb25c8b9e2a8b5ace5

Observation 82e1f335-9913-4b0a-b091-c3dc8bf9e61f · inbound

SEAL: Structure and Element Aware Learning to Improve Long Structured Document Retrieval cites this paper.

SEAL: Structure and Element Aware Learning to Improve Long Structured Document Retrieval LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T14:54:05.362827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:54:05.362827Z digest=sha256:b26b73c9e6dda2525d811d158473321ba2396c1edaaac4a6704ed2127b02fe5b

Observation b2259c11-c277-4724-baf5-5408aa39b30c · inbound

Adaptive KV-Cache Compression without Manually Setting Budget cites this paper.

Adaptive KV-Cache Compression without Manually Setting Budget LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:56.339536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:56.339536Z digest=sha256:350159c2fa6b79ad596cd25459b2fac7461f34b089f8b2a0bdf0180e6f334a62

Observation 9594f869-0768-4162-b5ab-2164c42e49c5 · inbound

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs cites this paper.

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T18:36:44.119175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T18:35:01.328250Z digest=sha256:44a10dadbec0f30edd3c355114c5b03c86cd2ccf50712f3f0c3d329ef4369372

Observation 636a967b-4c93-4789-a3fc-863072673810 · inbound

Kimi Linear: An Expressive, Efficient Attention Architecture cites this paper.

Kimi Linear: An Expressive, Efficient Attention Architecture LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:ccfc255b0ff38b47c8a79197cf03298b25caac97e4c2662ea2d50032c2344cc1

Observation 45fe34df-18b6-4ff6-9f5c-a3ab67ead4cf · inbound

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding cites this paper.

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-16T22:21:18.662116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T22:20:53.856657Z digest=sha256:999edf2cedae350b878c473dac98796c6d91dbb48fe878a24a157a5871ed9afb

Observation 7f3bb36f-358e-44ec-94d2-3d7d9f0e9c42 · inbound

MultiPath Memory Access: Breaking Host-GPU Bandwidth Bottlenecks in LLM Services cites this paper.

MultiPath Memory Access: Breaking Host-GPU Bandwidth Bottlenecks in LLM Services LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-16T21:58:35.743753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T21:57:12.142867Z digest=sha256:30c569ce1febdc57bce656ee8349d957bbb8f1ec3728c968674b9f26053c7afd

Observation 92c910f7-7760-4bbd-9e90-44a160fa926e · inbound

RepetitionCurse: Measuring and Understanding Router Imbalance in Mixture-of-Experts LLMs under DoS Stress cites this paper.

RepetitionCurse: Measuring and Understanding Router Imbalance in Mixture-of-Experts LLMs under DoS Stress LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1967

Resolution
unresolved
no resolver link, observed 2026-08-03T13:34:27.680374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:34:27.680374Z digest=sha256:c7c6f0cbb8a1ca62da33a8d7fcd40461cb594fa87b8bb0692f12f1349c7125ae

Observation 06646c79-1746-44d7-9f6c-456058ff025b · inbound

Cost and Accuracy of Long-Term Memory in Distributed Multi-Agent Systems Based on Large Language Models cites this paper.

Cost and Accuracy of Long-Term Memory in Distributed Multi-Agent Systems Based on Large Language Models LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T11:01:43.574177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:01:43.574177Z digest=sha256:da9be4b480b6e5588ffa7045eb073fd4bbde58c104da002bbe053b0e6f6a5ca2

Observation f78d0af7-384d-4468-9dc7-3022f1be3623 · inbound

Efficient Evaluation of LLM Performance with Statistical Guarantees cites this paper.

Efficient Evaluation of LLM Performance with Statistical Guarantees LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T10:58:40.958435Z digest=sha256:efa3531dfed923c9c70a698f53f6a94ad7d2d7e5f66b9a2e0034c0f608f13b6a

Observation 7bdb59db-758b-4c0a-8486-921d6a48b5da · inbound

Kimi K2.5: Visual Agentic Intelligence cites this paper.

Kimi K2.5: Visual Agentic Intelligence LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T16:09:05.225767Z digest=sha256:8c75684cf70421a9782d151f55a34affab22d66dd99f48590a55e89dd6faa9e0

Observation 7b91b948-0cf2-44d4-a74d-11a23d9a0d67 · inbound

ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs cites this paper.

ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T03:37:50.089031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:37:50.089031Z digest=sha256:97b8c4629b5ab7c33e3a4094ebb3f38010e036d11c17aea4e5bc677829b52627

Observation 4890c304-24af-41cc-a209-9921afe4fce1 · inbound

S2O: Early Stopping for Sparse Attention via Online Permutation cites this paper.

S2O: Early Stopping for Sparse Attention via Online Permutation LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T19:32:52.948154Z digest=sha256:706d2b27310cc8198b4d996f265353f3f88c580d1f137985ee580d4aec581688

Observation c4c40748-529d-4a04-a0e7-a635babfade7 · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T23:53:19.148407Z digest=sha256:54b03bb635a60fc11fe00898bffc42489ecc0f9bf9f3a8b87e89d68e20d1f008

Observation f6c892d1-7dce-47ca-9976-cb510c35bd74 · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T15:50:49.083652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:50:49.083652Z digest=sha256:527d93a72b7266abb862d67bc363c3a53d7da47512c9b23241e58bbccabab172

Observation 3c305c41-13ce-4293-9a2a-a6a24235f38a · inbound

Too long; didn't solve cites this paper.

Too long; didn't solve LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:29:25.830962Z digest=sha256:a614713eee5f8d8790e4b84c526872326fefc184e2e31902cc3204e23635de2b

Observation e72bf671-1ebf-4e80-9136-6bfb28b2cbd2 · inbound

PolicyLong: Towards On-Policy Context Extension cites this paper.

PolicyLong: Towards On-Policy Context Extension LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:23:28.939977Z digest=sha256:177c14165055f7d76c3ce1ea38fa98a750ae7ee67ca235846699fab9ea3e7f69

Observation b345f64f-8b1a-43b0-b32f-40d76d646f3c · inbound

MemExplorer: Navigating the Heterogeneous Memory Design Space for Agentic Inference NPUs cites this paper.

MemExplorer: Navigating the Heterogeneous Memory Design Space for Agentic Inference NPUs LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T07:45:17.043107Z digest=sha256:3f60280b7610648faa16cb2c195c64690a4add41d9140e14d561293f23edc771

Observation 520eb4c7-d000-4ced-b693-27d41a2677bf · inbound

Evaluating Multi-Hop Reasoning in RAG Systems: A Comparison of LLM-Based Retriever Evaluation Strategies cites this paper.

Evaluating Multi-Hop Reasoning in RAG Systems: A Comparison of LLM-Based Retriever Evaluation Strategies LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T04:01:21.400171Z digest=sha256:de84efda077d5a8c4bc246841a496e54226b590705bd2fc229580e897f8cd9a8

Observation 0db1359a-c7a7-43e6-9160-40ef73940dbb · inbound

An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA cites this paper.

An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T02:45:16.595617Z digest=sha256:2b78e7a70b72416d1835b0130343daf721790cc04be64b055ce1afd5784a4b1d

Observation 5865a45b-4bf6-429d-883f-58466214d165 · inbound

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference cites this paper.

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T14:09:30.821354Z digest=sha256:66160a515ce89acc0d1056df2bdea6889b39d8ca0a08ab1ee7eacfe7624509e3

Observation 7f751f32-7ab7-44e7-8d7d-be1bd96ed09a · inbound

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving cites this paper.

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T13:15:21.201950Z digest=sha256:a624a7a16b6d0c0f0ac1a0d55d5efb50b22a6937971be0cee5d6e83cc577055f

Observation 2ebd3edf-c93c-4af5-823e-457c710691dc · inbound

Telegraph English: Semantic Prompt Compression via Structured Symbolic Rewriting cites this paper.

Telegraph English: Semantic Prompt Compression via Structured Symbolic Rewriting LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T17:41:35.869547Z digest=sha256:f03ae7dba2bba3c9f580fb5635cc364e4df66c53b5a4d8e15d0ac0e675babb33

Observation 916d6fb6-c423-4d7f-a955-21c04aa38f66 · inbound

An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference cites this paper.

An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T02:56:28.828593Z digest=sha256:384a64a0fc5007740ba36e55af359809be933847a27d52c0abac501e8cc7a8db

Observation 013b6995-1260-4408-ae8b-3a7b2f547ad1 · inbound

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction cites this paper.

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T05:02:25.513351Z digest=sha256:fd10348d9a6100cdc5c810802bf35adf5229bd6f49ce5ecb65518d22e68287a6

Observation 34a51c8f-96c3-4ef4-b6b8-8eff8fa91ef0 · inbound

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models cites this paper.

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-06-30T21:05:04.209777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T21:00:25.664841Z digest=sha256:37958c5f3aeb17f05a76cec1ed19e84a2c25b9060f515f2d7e98afbc9df73dd5

Observation 954920ce-7e86-49df-9e2f-0c543f769b18 · inbound

CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selection cites this paper.

CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selection LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-19T21:22:47.986335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T21:19:31.263068Z digest=sha256:43dc027f9aa6a0694d46f529628c4434d0b80d56a2a3d82bd80562ff5c6543f0

Observation 80e81a8e-ca55-48d0-aa30-10846a4447dd · inbound

Draft Less, Retrieve More: Hybrid Tree Construction for Speculative Decoding cites this paper.

Draft Less, Retrieve More: Hybrid Tree Construction for Speculative Decoding LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T07:03:06.083271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:622f52ecf13a14c08a090a78f4aa355d100bedbcd6c9a95b13dc7bb0822f13f2

Observation bfd33b8f-ffec-4025-a15d-5fc5cd592954 · inbound

Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs cites this paper.

Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T07:49:49.972396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T07:49:07.128814Z digest=sha256:f1b6dbe2088fdcccaf9a643ea4f7c33f0be699412b48c049f44c0c09f27dbeca

Observation e4582de8-1476-4444-b2c6-6de991640fe4 · inbound

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline cites this paper.

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-02T07:36:45.401425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T06:49:08.689588Z digest=sha256:00a5a26e1e05b3db3ce87a1e9cf641202b2d9191f9b382f76e4c4427241afa25

Observation 5cdc01c0-74e0-4f3e-a751-34e83ca4d5b7 · inbound

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models cites this paper.

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-02T16:17:09.134221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T22:52:13.337419Z digest=sha256:24827933cbf16058e6fd301374ae5dc1867b75c5a31b2c5e7e81732ef76e2bab

Observation 1c3b1429-4dd5-4416-9303-18173bbfd785 · inbound

Still: Amortized KV Cache Compaction in a Single Forward Pass cites this paper.

Still: Amortized KV Cache Compaction in a Single Forward Pass LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T16:47:10.149044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T22:21:44.541078Z digest=sha256:47212b64e6c48a035f3298715254df8b1d63c024c12c36c8bae5cd56286faeee

Observation 836676ee-4348-48f2-8274-839a03852910 · inbound

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design cites this paper.

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-03T07:27:44.466295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T12:06:46.806138Z digest=sha256:824bbb05e445394cae11e477056f57ec9e80086835cdd49dd718133d9efa59a6

Observation 2b30b787-9acb-431b-a4a0-f8244d4530d4 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-27T13:00:56.050362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:a5c9ce958f4fe846a6849b551181c76e50381f764120530863fafb4285857f3a

Observation 891732c6-0ce9-42f7-aa3b-507062626e86 · inbound

Exploration Structure in LLM Agents for Multi-File Change Localization cites this paper.

Exploration Structure in LLM Agents for Multi-File Change Localization LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:58:06.478688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T09:13:29.422048Z digest=sha256:c37b7d9217b5f23c92244ad6227f44bd7212c93499426c8887c64f11e458f59f

Observation 9d810950-5844-4ada-ba15-f840484c7556 · inbound

MemTrace: Probing What Final Accuracy Misses in Long-Term Memory cites this paper.

MemTrace: Probing What Final Accuracy Misses in Long-Term Memory LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T18:18:49.836881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T03:13:52.804489Z digest=sha256:e4816febce11ce121b0fc9290bc5c8ea98ee2af11b6dfa053d1d901b9c382177

Observation 65e1f3ea-7b7a-4118-99de-6378977ff1e5 · inbound

LegalWorld: A Life-Cycle Interactive Environment for Legal Agents cites this paper.

LegalWorld: A Life-Cycle Interactive Environment for Legal Agents LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 45

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T00:59:19.813499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T20:50:25.070041Z digest=sha256:a35ac2dab506d4700fd4806ca97bce350d3dc7681e398c055826dbba9ab14d88

Observation fb8f66b2-581a-48f3-8cbd-bb09ad82ce85 · inbound

Test-Time Training with Next-Token Prediction cites this paper.

Test-Time Training with Next-Token Prediction LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-04T07:09:37.450483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T13:52:15.078658Z digest=sha256:512127459ec662ea85451ca81c579e4f420601a06b2def2d461cfa24c2af08d1

Observation 01a883ed-e55e-41a0-853a-f8d757606e09 · inbound

Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization cites this paper.

Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-04T15:59:56.418032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T01:12:04.486570Z digest=sha256:f2869dd9775670200bc0127eb77221e46eeea2012465b244d910b01674093753

Observation c882c256-398c-473d-b5a9-ae26b17003f3 · inbound

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution cites this paper.

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-01T15:35:48.389368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T01:20:29.026464Z digest=sha256:7e929519a87b0fde0111ff90614a8b7bc41931b15fd1873885fdb4c4f69b57b3

Observation 90f3b55f-e311-467e-8f6b-c7624085a974 · inbound

CoCoScale: Leveraging Layer-wise Scaling to Unlock the Potential of Online LLM Serving cites this paper.

CoCoScale: Leveraging Layer-wise Scaling to Unlock the Potential of Online LLM Serving LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T21:08:24.159706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:08:24.159706Z digest=sha256:bcb2432fa3021de8eaa959defcc4662356b6fc5a41d1dd764bfc0517592c4501

Observation 6e6a31db-f8cf-41de-8c5c-683d0477421c · inbound

Uncertainty-gated selection for block-sparse attention cites this paper.

Uncertainty-gated selection for block-sparse attention LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T22:15:14.580916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T22:15:14.580916Z digest=sha256:c70817a72ff4b23b2b2ea10c1c046f77aef079c0d1b4593f1fbd7a1ce70d46d8

Observation b8b269e2-085e-4f76-b0aa-791ec2358132 · inbound

UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp cites this paper.

UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T10:51:16.019022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:51:16.019022Z digest=sha256:6ee43588f3b68c0b778b469a28bdc4234766a36b82dcb2985f553cace716cd1d

Observation a7f1a10d-9790-47e9-a803-8c988cda7288 · inbound

Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching cites this paper.

Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T23:14:53.715023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:14:53.715023Z digest=sha256:f0372c8b94d5dd8600b10e711a54ce700dfc81081fccb3b7f94d0357ad863ee4

Observation 70c27d55-abd2-41eb-87ea-cce04d98831c · inbound

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context cites this paper.

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T07:13:06.322955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:13:06.322955Z digest=sha256:c99d71e0cab6ff0f10b40a6dbd59c5050c616a1075b8fdaa423d80abe396fcd1

Observation 9d781903-13b5-4533-9198-444390d591e6 · inbound

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment cites this paper.

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T07:12:17.070993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:12:17.070993Z digest=sha256:6b6b3058aa9d6455f5443c9dbe9daa0cca35bee982c4e9e7ba72928ad03d5fa7

Observation 40e78ac4-1092-4d7b-8755-6881fb3e07e1 · inbound

CoSA: Accelerating Long-Context Inference via Proxy-Kernel Co-Designed Sparse Attention cites this paper.

CoSA: Accelerating Long-Context Inference via Proxy-Kernel Co-Designed Sparse Attention LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:55:47.646679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:55:47.646679Z digest=sha256:2933a0bcebf611f872649555a14afa5835f06bb0ba6327a3882b413fffc70025

Observation 95675119-b04a-4523-8845-d0dc42cdb2e7 · inbound

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition cites this paper.

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T02:55:23.352755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:55:23.352755Z digest=sha256:91339b61ebc7cb44b35e4dfc382c7d2798724b4a56e5165c0b27362833de96ed

Observation 8f1184d0-7747-431d-9ba4-9b1405d0be97 · inbound

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding cites this paper.

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-03T11:54:22.062287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:54:22.062287Z digest=sha256:26204a97a0d370ed9dbcd59feebc4d788da639c8c6041e82be59babf4b94feaa

Observation dad34323-354c-42b1-ad83-b3bf3c668532 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:50.544562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:50.544562Z digest=sha256:96b2486179b3691a50539ffcbf4af771974f39d359877cbac2cca8b3777dbdd8