Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 100 inbound Pith citation observations for arXiv:2206.04615.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T12:45:26.750295Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T13:57:06.851393Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 499806ea-bbf2-4490-bbe7-dd3bcd80cd63 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models MathQA: Towards interpretable math word problem solving with operation-based formalisms
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b3f33b82-cb8f-4a1c-8b27-ffb40ba3e906 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5b4d2b85-d7ba-4ca5-b9b0-fe4bed041872 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fccc22d1-56e2-4d45-a925-29e86d794ee7 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models On the Opportunities and Risks of Foundation Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1c6ef8b-9b33-45a4-9277-21817768be6d · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/W18-6433
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4146ebab-a66e-4011-adfe-a37e05701b55 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Simplicity: a unifying principle in cognitive science? , volume =
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 39fd2b3e-a7a1-45e4-835f-59cb7b62045b · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/W19-3824
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a79cc0b-eea7-4202-8267-ea89daf77dd8 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.1007/978-3-319-40566-7_4
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 97543141-53c3-408b-a2a8-81373b042ea0 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models overinformative
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d0676663-d3ef-47e7-8d71-97e29573918f · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bb376dd7-54e5-4a40-be04-f253ea6b370e · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Making sense of sensory input
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2cf03759-18fd-43ff-8cae-3ad10a984262 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/N19-1395
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ca183db-c8a3-481f-af4f-623814213246 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/P18-1082
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f6546cd8-dd12-4765-a18e-80c45a3645e1 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Fodor and Zenon W
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 524977f3-62e1-4c1f-be93-0d118172103f · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.5555/1625275.1625535
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b2f4698a-62f4-4395-ae66-d47c8c61a3f7 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 903fb3e9-d637-4c24-bcfe-a7d15167bd7c · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/N19-1061
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5b1d8f6c-a77e-45e7-99c8-776eda943dbe · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models URL https://doi.org/10.35111/0z6y-q265
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 761b008f-20c5-4bed-98d1-2e7d43ebf516 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models URL https://doi.org/10.1145/1925844.1926423
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b87c8db7-c871-425b-aa99-3f54658d6a31 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ce107d8-6f8b-4a64-bb32-ecb300aed9ad · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Henrich, S
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e6a67e0f-d32d-4156-b59e-4aa5616ccb6f · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 29) China Household Management Research Center, Ministry of Public Security
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f6f6bbaf-1890-4453-b1fb-fc4178b9beb9 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/2020.acl-main.164
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a0c8357e-077c-456f-95f3-fa0a7f76b497 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6f0dd71-dd83-4cce-9fc7-83a3b29def67 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ca22ed36-c860-4d89-9fec-1bb709f5652c · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models The N arrative QA reading comprehension challenge
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 38a947b5-7079-4c93-88ac-6a87896fabc0 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models URL https://doi.org/10.1007/s10992-020-09581-6
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6e2d620-b30b-4b4c-99d1-bb0b9819d509 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 67137944-82f1-4b14-ae6a-fbc76ee12077 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/W19-3005
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 107b8f27-6805-4849-bd17-bf3c36680518 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models and Rudinger, Rachel
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df711b25-1885-4ae8-aba9-904d944a6714 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Andere zeiten, andere lehren
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dfcc865a-7bcf-40cd-a00a-7ada0c416117 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 31) David Milne and Ian H
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 49d99f63-600b-47de-9374-a6d684d8fef2 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models URLhttps://www.aaai.org/Papers/Workshops/2008/WS- 08-15/WS08-15-005.pdf
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9d40431e-0780-4aa4-993e-42d853c25105 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models The Deep Bootstrap Framework: Good Online Learners are Good Offline Generalizers
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 20abb63c-278d-4e6e-8a6d-ed8e96ba45d8 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Cohen, and Mirella Lapata
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 980c77ab-506c-44b2-a1d0-7ee1e55d1df4 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models doi: 10.18653/v1/P19-1442
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 596f7243-5571-4f85-b020-89e74078c9ef · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models URL https://doi.org/10.1080/02724980443000566
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bf6af985-ea02-48ec-849b-ecfbb3a2c1f5 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 32) Judea Pearl.Causality: Models, Reasoning, and Inference
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41bf6c56-66ad-4d39-af61-31d4c40ffc41 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 29) Tony A
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 51ff55af-49cb-4d4d-9d6e-1e4318f94bfa · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 29) Robert Plutchik
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 245e1109-2da1-45d4-a2f6-03e8a52d39b9 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models URLhttps://aclanthology.org/2020.lrec-1.125
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 39592300-ccce-438b-a875-15853f495b46 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e4e5305e-3980-4d06-84c8-40c1e1b746b4 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 38) Zijian Wang and David Jurgens
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3302fd33-ab6b-4d8c-993a-441d920dcd70 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models assessing BERT’s syntactic abilities
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a492a40f-8a0a-4550-b874-08fb972b1b5a · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bfec096c-fe87-4344-b1ab-ee0c9365a972 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on p
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2993704c-b483-4fe5-83e6-579323448f3b · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models (cited on pp
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 09903080-9f60-479b-ac94-9d2368280315 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 31) Zhou Yu, Dejing Xu, Jun Yu, Ting Yu, Zhou Zhao, Yueting Zhuang, and Dacheng Tao
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fbb6d9ce-4ec6-4ee7-830c-64a5eebb8847 · outbound
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8c9721d4-2582-4b35-b622-46b8cd1bcff6 · inbound
Emergent Abilities of Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6fd99ee7-4c87-40a3-8c88-2e151787f520 · inbound
Language Models (Mostly) Know What They Know Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 210
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ca686e16-f7fd-4a9b-9aa8-209b6866ec99 · inbound
Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fc9fc5b9-7fca-4383-8567-25f4e0934cd4 · inbound
Large Language Models Can Self-Improve Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3fa233c2-a16c-462e-85c3-a27abbdd2d14 · inbound
Large Language Models Are Human-Level Prompt Engineers Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d0a8613-2704-41d5-ac87-4d3d39eb0b24 · inbound
BLOOM: A 176B-Parameter Open-Access Multilingual Language Model Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 141
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 81c0ff3f-125d-46a0-861a-946494458a01 · inbound
Galactica: A Large Language Model for Science Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba537a72-2496-4e2e-910e-77c977d362ee · inbound
Galactica: A Large Language Model for Science Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 237
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4648f200-ceaf-41e9-9635-e36dee31d635 · inbound
A Survey on In-context Learning Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 81c0c17a-e935-4f62-8c74-08f14b8f96bd · inbound
Progress measures for grokking via mechanistic interpretability Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e208446e-2900-4851-90ab-8b131f407bb0 · inbound
The Flan Collection: Designing Data and Methods for Effective Instruction Tuning Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 66c7c594-be97-4ee4-b1b0-a5f8e6b1f754 · inbound
A Multitask, Multilingual, Multimodal Evaluation of ChatGPT on Reasoning, Hallucination, and Interactivity Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7023981c-6e77-473f-80a4-3a3747e58a41 · inbound
ART: Automatic multi-step reasoning and tool-use for large language models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 152
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e0b65df4-9b71-448a-8545-f725732ae354 · inbound
BloombergGPT: A Large Language Model for Finance Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7da3c7be-ee12-4a6d-b4ec-5087c26ba17f · inbound
A Survey of Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ff7e6c23-fc25-4246-bf2a-fecdeafdedf2 · inbound
Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3b189186-8328-4d9c-b3ac-e79d18a634c5 · inbound
TinyStories: How Small Can Language Models Be and Still Speak Coherent English? Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1651c060-1637-4d28-937d-eb8a4fc78425 · inbound
Towards Expert-Level Medical Question Answering with Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6539ad68-7843-4cdc-9428-3fdda8b441bb · inbound
PaLM 2 Technical Report Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 255
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c48315cd-a1d6-454d-9eb1-1eb430b1ae93 · inbound
Evaluating the Performance of Large Language Models on GAOKAO Benchmark Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9efdb8f6-6d40-4ddf-9f14-5b160c602898 · inbound
Improving Factuality and Reasoning in Language Models through Multiagent Debate Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae218b2d-079e-4134-93ef-583a70e09dd1 · inbound
Scaling Data-Constrained Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7e1adb2e-beef-4bda-b5ac-906e746478fb · inbound
Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dac4623a-7cb0-49db-a419-f0ab19f0786f · inbound
Simple synthetic data reduces sycophancy in large language models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e79ed1ad-f787-416f-aff3-fedc40d44e70 · inbound
Reinforced Self-Training (ReST) for Language Modeling Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3f47cb8c-b37d-47d8-93b2-3f79c3b0c7ac · inbound
Large Language Models as Optimizers Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05828090-91fe-4df2-b661-0556f42135f3 · inbound
C-Pack: Packed Resources For General Chinese Embeddings Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 06822792-59bc-492c-9d91-e34fd508279e · inbound
Baichuan 2: Open Large-scale Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bd11d899-2e47-4818-8557-8e38eaf432fc · inbound
Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fac13bf6-5b03-4344-ba03-85efd1e97681 · inbound
Gemini: A Family of Highly Capable Multimodal Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cbfa9301-0159-4477-94f7-fca74eb99e7e · inbound
TinyLlama: An Open-Source Small Language Model Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc313612-5b75-41e4-9f01-a621e387e9fb · inbound
Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 114
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0cab335d-348b-41e8-b28a-fdcf70ca187d · inbound
KTO: Model Alignment as Prospect Theoretic Optimization Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4b20af0c-c671-447a-8d15-b69f1be61a32 · inbound
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 138
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 26b56491-f2e6-4bf5-a2ac-0f9180e7a2b1 · inbound
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 124
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bc9ed97e-d90f-4a90-99fe-84a9c60fcef2 · inbound
LLM Evaluators Recognize and Favor Their Own Generations Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 44b0fadf-d768-432a-9d07-5862ff58f81c · inbound
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f5a0b5e5-b0c3-4d9a-9b54-f14466339c26 · inbound
The Platonic Representation Hypothesis Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 162
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 66bdd477-3ef2-4772-a211-3ee151b39f68 · inbound
Lessons from the Trenches on Reproducible Evaluation of Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6dc02a4e-23d8-4905-8760-e3009390288a · inbound
MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1a0b8c6-96f0-4693-b834-2bcc59cefcec · inbound
ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 298e6435-eb7e-4b40-96b6-51b5bf4bc50e · inbound
LiveBench: A Challenging, Contamination-Limited LLM Benchmark Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f506b820-157b-43fa-9113-400f359ecf1b · inbound
Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3de9ad2b-c6f4-493e-8eee-02f35626dc37 · inbound
In Context Learning and Reasoning for Symbolic Regression with Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b070f90-0ee9-4746-900e-91bf360f1dd8 · inbound
Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42f5d02c-d974-48c2-8473-d4b4e8d1beab · inbound
A Survey on LLM-as-a-Judge Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 56f5e2ec-d951-42eb-9e15-3fbe4fc8bd70 · inbound
A ghost mechanism: An analytical model of abrupt learning in recurrent networks Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 697f7b4d-c536-4372-a270-c84f7110c643 · inbound
Small Language Models (SLMs) Can Still Pack a Punch: A survey (updated 2026) Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 412b4df4-37e4-4546-8fe9-9d919cadaec1 · inbound
Do generative video models understand physical principles? Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation af5e1950-1d1b-4cf1-9108-b245c7f21bcd · inbound
Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 138
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ceabc31-aa91-46dd-97e1-43cef4c1a185 · inbound
Humanity's Last Exam Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 06984b8b-98b9-438c-aa10-b52545e32b64 · inbound
Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a023a95d-9ed1-41f0-8143-ffd0049b7396 · inbound
Towards an AI co-scientist Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 122
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2ad8cbf1-a96d-4bee-a255-1e061f6aeede · inbound
Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0ddd7ef1-3f11-48bd-bd19-29aed95615ce · inbound
Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 75e71827-cdf9-4db2-967e-915f3839ce01 · inbound
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2361ad68-2484-43e5-966c-7c8206d5f9a9 · inbound
BacPrep: Lessons from Deploying an LLM-Based Bacalaureat Assessment Platform Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 204b904e-2e3c-45b2-8acd-7be9f63e6e8a · inbound
PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation db9e0a50-6b49-499d-8873-ce9c82f4dc3b · inbound
Bridging Brains and Machines: A Unified Frontier in Neuroscience, Artificial Intelligence, and Neuromorphic Systems Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 162
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation efceb3b3-d0ca-4cf2-af61-b6978a548905 · inbound
Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e7d87052-6883-4476-9245-600fb22d3202 · inbound
Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f6bd6156-db78-4c68-96ac-890b8f526955 · inbound
Designing Psychometric Bias Measures for ChatBots: An Application to Racial Bias Measurement Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e7b4900c-c7d6-4aac-a18f-5b204c5e0ccd · inbound
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cde58f09-7efa-495d-8171-387b557312c6 · inbound
Controlling the Risk of Corrupted Contexts for Language Models via Early-Exiting Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80be55b5-0748-45e0-a2ec-247eaa29da5c · inbound
Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 319e9b58-67a0-441e-9d78-f2a7366e8fd7 · inbound
Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3e4bc57-bbd7-42e9-a1b9-7b1ca3204f87 · inbound
The Art of Scaling Reinforcement Learning Compute for LLMs Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 854d13bd-fdbf-443a-a3c1-f9df365e5b8a · inbound
Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3b88ff94-1e23-46d1-a5ca-4b9005d66098 · inbound
DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c49015a-43ae-444c-bd7e-88131d4a7d83 · inbound
DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783aef56-55e9-4758-a1c9-d04e478c10a4 · inbound
Towards Real-World Validity in Generative AI Benchmarks: Understanding and Designing Domain-Centered Evaluations for Journalism Practitioners Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d7849b73-2840-4e4a-a39d-98d78d63ae77 · inbound
Contrastive vision-language learning with paraphrasing and negation Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9910e596-a17e-4a09-8ee8-bdde9bcf70d5 · inbound
Memory in the Age of AI Agents Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8cbe9bf7-32e8-43ca-aa7a-22c6c3e1a805 · inbound
Beyond Context: Large Language Models' Failure to Grasp Users' Intent Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d1f6dc17-0cf5-437c-a7b8-d7322cc96c0c · inbound
LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 145
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52b1f8f9-cd37-4b3d-9806-07d0c9eaf17f · inbound
When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f21422c-1069-42a7-88db-dca543e86cba · inbound
Mechanistic Evidence for Faithfulness Decay in Chain-of-Thought Reasoning Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 272130d4-2c34-4bf2-8dde-3596f4889729 · inbound
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c140785b-4a83-416f-b3cd-a0029e39f001 · inbound
Turbo Connection: Reasoning as Information Flow from Higher to Lower Layers Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c0aaf51-26a8-4fba-95cd-e6e0708aa674 · inbound
CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04320c7b-90d2-45b4-a803-b26a1c80f824 · inbound
Graph Property Inference in Small Language Models: Effects of Representation and Reasoning Strategy Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 301c55de-fbe7-4ca3-b28f-d157955d06e9 · inbound
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c175797-94bc-4917-bc83-75eac0c795d4 · inbound
FrontierFinance: A Long-Horizon Computer-Use Benchmark of Real-World Financial Tasks Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 54d18fb5-ad1e-498f-b34e-6e2e2af628a4 · inbound
Leveraging Weighted Syntactic and Semantic Context Assessment Summary (wSSAS) Towards Text Categorization Using LLMs Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation addbe7bf-e89d-4757-9468-bbcfaa47bc1f · inbound
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cd5c6406-627a-4879-a478-656a49e76187 · inbound
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec686702-ebf9-4b0b-97e6-cfd86ecedae2 · inbound
Parcae: Scaling Laws For Stable Looped Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bd2831a6-a122-4d3f-afd0-a44c195896cb · inbound
Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation faf54622-7ee2-43d9-8b27-785d8f0b568e · inbound
Consistency Analysis of Sentiment Predictions using Syntactic & Semantic Context Assessment Summarization (SSAS) Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 09f64f09-602a-467d-b4fc-b9500cd8316f · inbound
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7fcd0e40-c8dd-4b2a-90ed-bdcd4c874f6d · inbound
Measuring Representation Robustness in Large Language Models for Geometry Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 08069325-db17-4714-995d-f9947b29d698 · inbound
Calibrating Model-Based Evaluation Metrics for Summarization Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 155
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0a6f6160-40b6-4173-99cf-22707a57ec45 · inbound
Beyond Static Snapshots: A Grounded Evaluation Framework for Language Models at the Agentic Frontier Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d6df45f4-c7f2-41e9-8119-468aab917164 · inbound
QuickScope: Certifying Hard Questions in Dynamic LLM Benchmarks Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c0beaa9-fa61-49c3-847b-da814739ba65 · inbound
QuickScope: Certifying Hard Questions in Dynamic LLM Benchmarks Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 72c5f6d7-45bc-4714-985c-a48ff089c4f6 · inbound
AIT Academy: Cultivating the Complete Agent with a Confucian Three-Domain Curriculum Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9f57a9cf-c60c-4b9b-bd09-3b9fd8b18829 · inbound
TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 049bb1c8-3a57-416d-af92-7a0377c98006 · inbound
Evaluating Agentic AI in the Wild: Failure Modes, Drift Patterns, and a Production Evaluation Framework Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5fd3f028-5201-45c2-8662-d19a15e5ce82 · inbound
Complexity Horizons of Compressed Models in Analog Circuit Analysis Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 213dbb2c-6a05-4eff-bc1c-29ccfa692f48 · inbound
A Meta Reinforcement Learning Approach to Goals-Based Wealth Management Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 299
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.