Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:45:02.174033Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2506.23735.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:45:02.174033Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b59611fa-8696-4bb1-9e6f-0df2cd2a98a9 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 025d6689-70c1-4393-902d-71418d1ccd7c · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Deepseek-v3: Scaling open-source language models with mixture of experts
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a67eaf4-1d63-4ccc-9c49-ca13d65b5496 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Black-box generation of adversarial text sequences to evade deep learning classifiers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6813975-75c2-4d2f-ad10-1df82cbb37cd · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Measuring Massive Multitask Language Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e14cfab9-b745-4349-ba58-6caa471c5fef · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data C-Eval: A Multi-Level Multi-Discipline Chinese Evaluation Suite for Foundation Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea7df36b-156b-492d-ab37-f536b1963162 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Adversarial text generation by search and learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baaf731e-2d06-4501-9f47-67da1aceedf5 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Beyond Static Datasets: A Deep Interaction Approach to LLM Evaluation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53df4759-7c2d-4fb4-9e65-93c92e196499 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca5d8135-ce8e-42d7-bc25-ef1220d1d270 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Using Adversarial Attacks to Reveal the Statistical Bias in Machine Reading Comprehension Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b0ab9bf-24d6-45b8-b827-73b7a419d91d · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fab41cb-ecc3-4953-aca6-697cfc9ef00b · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data MedMCQA : A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c0eae8f-4d19-48d9-ac76-bfaf9f4e6033 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data Alcuna: Large language models meet new knowledge
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfd065d6-b13c-4793-8dbe-5967b7222084 · outbound
AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data ALCUNA: Large Language Models Meet New Knowledge
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.