Pith. sign in

Paper Citation Record · LEDGER

L-Eval: Instituting Standardized Evaluation for Long Context Language Models

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2307.11088.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.11088 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T12:42:30.058516Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d221e3e8-9e3c-49c4-bd7f-865d78f8f29e · inbound

LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding cites this paper.

LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:22:10.690217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-12T20:22:10.482509Z digest=sha256:22699876e7a037e19bbef402a0fd75a1e3f36a79e26c7b96ea48a22532a459cd

Observation c8d343fb-e86f-43e3-9e17-4839fb1389e0 · inbound

InternLM2 Technical Report cites this paper.

InternLM2 Technical Report L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 147

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:44:38.169759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-15T11:44:38.066501Z digest=sha256:c315ae2b44cbfc96df221459246a4eacce6b8f11f3fe8a6c123f7e7d8fef6cd0

Observation aa329656-4e69-4293-9776-6404ff6d1aec · inbound

Jamba: A Hybrid Transformer-Mamba Language Model cites this paper.

Jamba: A Hybrid Transformer-Mamba Language Model L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T14:11:27.181188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T14:11:27.156350Z digest=sha256:b38e5d1d1479e992d93dff97777f13d3a5b68367da3f284f8abb16e21ff00d63

Observation 348f8151-5ca1-48bc-9cdc-d068db0934af · inbound

SnapKV: LLM Knows What You are Looking for Before Generation cites this paper.

SnapKV: LLM Knows What You are Looking for Before Generation L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:57:43.115018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T12:57:43.036674Z digest=sha256:93775a029a1fd877f0bef547add26e49274650b268f0c6b1acb891047130bb4b

Observation 708c150c-2790-420c-91d6-24f836508ca8 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T17:35:44.190233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:80eb18fb6a8bb9a28742084be686554fffa44318dd97400f8e4a62dddfafd1b0

Observation 20812d0f-3b1c-40d3-b55f-2f68fbd2ea98 · inbound

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression cites this paper.

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T04:17:31.131311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T04:15:36.906263Z digest=sha256:ec9191c7daf893e1a1a4c47c643c80397b0746f8df668a55332c889557a83193

Observation 8e835c7a-4d13-4570-86c9-3b8f29da5fc3 · inbound

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs cites this paper.

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T12:42:30.058516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:42:30.058516Z digest=sha256:2bd41aa24fb4936745d2d2fa7d4bd5a9ee38bbed13725103854fcc638dbf60b3

Observation 797f398f-5c61-45d9-acde-db8426259afe · inbound

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives cites this paper.

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-10T01:04:50.201027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T01:00:41.543394Z digest=sha256:6e5a2d0badc49c6d534c2806c61a1d9c4df24f8510284f252c0a0faf3f988051

Observation 96ae3ef3-36cd-4cf3-b0b5-c422572678f0 · inbound

Efficient Training on Multiple Consumer GPUs with RoundPipe cites this paper.

Efficient Training on Multiple Consumer GPUs with RoundPipe L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:31:26.839097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T10:37:22.251566Z digest=sha256:27a1a4fb74d6ac72e4331e1ad65f4fc43b6919d8955acedee888c8ba52711c10

Observation d630b536-3192-4570-86ac-9e54bffbd585 · inbound

Tutti: Making SSD-Backed KV Cache Practical for Long-Context LLM Serving cites this paper.

Tutti: Making SSD-Backed KV Cache Practical for Long-Context LLM Serving L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:31:10.427339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T16:19:33.613685Z digest=sha256:6225f89b92c9644de5dce8640eab92ad19d8a0473c458424eb3baef865dd7345

Observation 8b982927-667f-479e-936f-73af253b2cc8 · inbound

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring cites this paper.

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T01:33:27.224830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T01:32:32.967922Z digest=sha256:2dece0fcea5698538af2bf0277cb0f61f45f7dc414cbe72e9333b3ee534d199c

Observation 1375033c-3ef9-4768-b005-fb1777f324a1 · inbound

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring cites this paper.

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T21:15:03.983326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T21:13:23.148984Z digest=sha256:aabe1445496b449502e182ab717b1fd1f22f73c621a1a87734c911db418d24e1

Observation caa6a0ab-d8f0-4702-aa93-ccb1681042ee · inbound

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks cites this paper.

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.921291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-25T04:58:15.184063Z digest=sha256:ab90ecd5877dcb38faaabb186d51aed3fd0b8b7c25c83ae73c36cccc0d9c32ee

Observation 9fb53eed-2b60-435b-b151-dde4fd5a5034 · inbound

JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data cites this paper.

JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:54:44.114006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T13:45:38.305767Z digest=sha256:6a68aca7a71de55511fd54b8bf7273a0600400d524674a7ca540956c8bd584f7

Observation 2c31c983-3808-4be8-bb80-78fd0a9f8580 · inbound

NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama cites this paper.

NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T19:28:52.462812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T01:48:08.794210Z digest=sha256:edf664a90563e7996bdfbe1f925397c2b0916d3eff36b5bfe04c55dfe4f08395

Observation dbaa4e11-876c-4eb7-a43f-42d4accbe3c0 · inbound

Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling cites this paper.

Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:03:51.865567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T04:59:00.304723Z digest=sha256:af43f60248accf5c51a3894dd7c580e2699594086925e7970295e530c6f74971

Observation a077d62c-96f8-46b0-b0c1-5554bd719916 · inbound

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment cites this paper.

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-01T07:12:16.866244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:12:16.866244Z digest=sha256:f3b517ea71ddc037280d2b68e9f934b2c1f79096f6736bc301924bf626b2d44a

Observation 1903ada6-956f-477c-8e7f-bdaae9c022d4 · inbound

Memory for Large Language Models cites this paper.

Memory for Large Language Models L-Eval: Instituting Standardized Evaluation for Long Context Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-01T02:37:54.577072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:37:54.577072Z digest=sha256:bef9d843bb95ec900870fc2bdb47df6b886624eaa6a4faef3638d847bf21728d