Pith. sign in

Paper Citation Record · LEDGER

Efficient Benchmarking of Language Models

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2308.11696.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.11696 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T23:49:57.580051Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:37:14.599642Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation aec9314c-16e6-4227-94d2-049a26902099 · inbound

Holmes: A Benchmark to Assess the Linguistic Competence of Language Models cites this paper.

Holmes: A Benchmark to Assess the Linguistic Competence of Language Models Efficient Benchmarking of Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-24T02:08:45.131464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T02:06:53.585629Z digest=sha256:4ece3d308e1b955bbb38b0c810de09b622401781bd0c06e7a8619a5ecd5c25ea

Observation 24dd9ceb-dede-4a83-8364-593207157b50 · inbound

Query-efficient model evaluation using cached responses cites this paper.

Query-efficient model evaluation using cached responses Efficient Benchmarking of Language Models

Reference 112

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:56:00.055951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T00:57:26.494031Z digest=sha256:74b8a2d0e48cc03add087a8a60f4f600c10c11e1d6266344cc82a0bcf674151b

Observation deaf8605-9742-4671-a893-d242a5ca1f18 · inbound

Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation cites this paper.

Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation Efficient Benchmarking of Language Models

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T23:52:49.477606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T23:49:57.580051Z digest=sha256:997041401e65fa30a1f9a2f0216753bb0c384748c47f8d8c79c9dd05ad6782fa

Observation 197d6ccd-ded6-4336-bc8d-c193e0de6b86 · inbound

Consistent and Distinctive: LLM Benchmark Efficiency via Maximum Independent Set Prompt Selection on Similarity Graphs cites this paper.

Consistent and Distinctive: LLM Benchmark Efficiency via Maximum Independent Set Prompt Selection on Similarity Graphs Efficient Benchmarking of Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:02:24.486175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T16:56:29.536912Z digest=sha256:189fa2764b6714b7b2a934c9425b709f462fa9981593d5362cfa6014a903eea4

Observation 7a49a4ea-4add-4ebd-b9be-34f0e11c3acd · inbound

Quantum-Inspired Trace-Augmented Evidence Selection for Reasoning over Structured Hypothesis Spaces cites this paper.

Quantum-Inspired Trace-Augmented Evidence Selection for Reasoning over Structured Hypothesis Spaces Efficient Benchmarking of Language Models

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:37:14.601224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T21:57:20.848868Z digest=sha256:b2f02b648f0ce6ee5bde1cca452d3b42940ea7e8cb07b27e187b7343a638c464