Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2104.02145.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:54:32.977436Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T21:16:13.549729Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 114205d4-b6f3-48d4-b9a4-18830077255b · inbound
GPAI Evaluations Standards Taskforce: Towards Effective AI Governance What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c93d6f05-293f-4694-83ec-c08040ca669e · inbound
Towards Effective Discrimination Testing for Generative AI What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f633ecc-145a-4fa1-99df-696893cb17e7 · inbound
Do Large Language Model Benchmarks Test Reliability? What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5161b5c4-8a12-4a33-a9a2-18c05efa0da6 · inbound
AUTOLAW: Enhancing Legal Compliance in Large Language Models via Case Law Generation and Jury-Inspired Deliberation What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aadb985-3932-4443-9b5c-20ebf50d25b2 · inbound
Potemkin Understanding in Large Language Models What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7b0d58c-24b5-4c16-808e-07ea8fb92706 · inbound
Private, Verifiable, and Auditable AI Systems What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8741e97a-d21d-4b4d-8341-40c8bcb021d2 · inbound
Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4dd1879b-a8f5-47f6-8591-e7ad02767f45 · inbound
The Case for Model Science: Verify, Explore, Steer, Refine What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 278c8f23-f6ca-4367-9807-70f3fc487741 · inbound
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA What Will it Take to Fix Benchmarking in Natural Language Understanding?
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.