Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 100 inbound Pith citation observations for arXiv:2107.03374.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
22 of 22 outbound references displayed
External citation measurements
1427
pith, observed 2026-08-05T02:28:24.338817Z
Observation 6775eca7-a7e3-4cf9-b02b-b213ab340f21 · outbound
Evaluating Large Language Models Trained on Code wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c79fa98-a5a9-4777-9499-944352cfa36b · outbound
Evaluating Large Language Models Trained on Code Generating Long Sequences with Sparse Transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 091eb877-cb95-4ef3-9bab-7174aa259756 · outbound
Evaluating Large Language Models Trained on Code Clarkson, M
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62706afc-d190-49d0-a0d6-d3a3bb9ae4d1 · outbound
Evaluating Large Language Models Trained on Code BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db5c3b46-2d3f-4b2f-ac9d-447c9c050d76 · outbound
Evaluating Large Language Models Trained on Code Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fcff2c7-3515-48fb-8c92-f5a2c60e3e84 · outbound
Evaluating Large Language Models Trained on Code Number of elements are less than k
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34ff45a2-277d-4626-996f-3099c9723d85 · outbound
Evaluating Large Language Models Trained on Code "" ### COMPLETION 1 (WRONG): ### return x if n % x == 0 else y ### COMPLETION 2 (WRONG): ### if n > 1: return x if n%2 != 0 else y else: return
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 848189d2-094e-41f1-824f-8b42eec8b8b7 · outbound
Evaluating Large Language Models Trained on Code remove all instances of the letter e from the string
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ada604cf-f260-47ec-8b10-c2e892e17738 · outbound
Evaluating Large Language Models Trained on Code replace all spaces with exclamation points in the string
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 666f331c-d1e6-40ff-a969-95deaf213e16 · outbound
Evaluating Large Language Models Trained on Code convert the string s to lowercase
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b3cb82b-438f-4526-9b40-14faeb483442 · outbound
Evaluating Large Language Models Trained on Code remove the first and last two characters of the string
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c13ab40a-4630-4b7d-8e62-b24d5e35edaf · outbound
Evaluating Large Language Models Trained on Code removes all vowels from the string
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 207afd72-1ae0-4b41-bf1c-141de88ec2bc · outbound
Evaluating Large Language Models Trained on Code remove every third character from the string
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 538c5d86-d1f0-4171-8b32-492f2e72f9ef · outbound
Evaluating Large Language Models Trained on Code drop the last half of the string, as computed by char- acters
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0970941c-eeda-48cc-8922-9bea0310597d · outbound
Evaluating Large Language Models Trained on Code replace spaces with triple spaces
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd7e41d6-b3c5-44a9-9338-36224d03dadc · outbound
Evaluating Large Language Models Trained on Code reverse the order of words in the string
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab3777a8-d2c2-4ca3-8f0f-7bd8d4a6e78a · outbound
Evaluating Large Language Models Trained on Code drop the first half of the string, as computed by num- ber of words
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3370eca5-4706-4b28-bd92-49941edc7f38 · outbound
Evaluating Large Language Models Trained on Code add the word apples after every word in the string
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a7d4dd1f-1f8f-4d10-b25a-efdfc37df6d1 · outbound
Evaluating Large Language Models Trained on Code make every other character in the string uppercase
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2558f076-28ad-4cd5-bb74-1da50af953fb · outbound
Evaluating Large Language Models Trained on Code delete all exclamation points, question marks, and periods from the string
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4652fab6-c693-4529-b937-a0249be20f49 · outbound
Evaluating Large Language Models Trained on Code When the prompt includes subtle bugs, Codex tends to produce worse code than it is capable of producing
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8eb67d4b-80ac-40f0-ba59-a3948cb7a762 · outbound
Evaluating Large Language Models Trained on Code 10") 10 >>> closest_integer(
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ede0bc2b-2c2b-4356-8d5b-574a40ce4c17 · inbound
Measuring Coding Challenge Competence With APPS Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fb2b0ab-3cf0-4775-908a-77122bc48125 · inbound
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8167052a-8ea5-4642-8a92-b2e5c8ed9a2d · inbound
TruthfulQA: Measuring How Models Mimic Human Falsehoods Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0208c989-8cfd-40fc-b419-10a23a247863 · inbound
Show Your Work: Scratchpads for Intermediate Computation with Language Models Evaluating Large Language Models Trained on Code
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eae23089-4d58-44f0-81ef-ef055ca36342 · inbound
A General Language Assistant as a Laboratory for Alignment Evaluating Large Language Models Trained on Code
Reference 218
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8f5e8b6-cb14-49b3-874c-93d6f5f17628 · inbound
Ethical and social risks of harm from Language Models Evaluating Large Language Models Trained on Code
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9859903a-28f8-4f62-a559-12d723f3133b · inbound
Text and Code Embeddings by Contrastive Pre-Training Evaluating Large Language Models Trained on Code
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5612270d-b438-4009-bd9a-5f6a12d4b0be · inbound
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models Evaluating Large Language Models Trained on Code
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3647d15-340e-40b3-b08b-3b98bd74bf7b · inbound
Quantifying Memorization Across Neural Language Models Evaluating Large Language Models Trained on Code
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 826508b5-3646-4866-85e0-2ffb1d738ed8 · inbound
Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language Evaluating Large Language Models Trained on Code
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adc4bb5e-3c9a-4c4c-9a81-7d17b4343d4c · inbound
PaLM: Scaling Language Modeling with Pathways Evaluating Large Language Models Trained on Code
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5642cd0b-22ab-44dc-99d7-cd88a40cb1ef · inbound
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback Evaluating Large Language Models Trained on Code
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb5cf1b4-3081-4255-a808-1e9d10bf9771 · inbound
InCoder: A Generative Model for Code Infilling and Synthesis Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9499c658-d77a-413c-9d6e-34664636ef2c · inbound
GPT-NeoX-20B: An Open-Source Autoregressive Language Model Evaluating Large Language Models Trained on Code
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2484b4e1-55a9-4e46-88ad-a885fe98a18a · inbound
Language Models (Mostly) Know What They Know Evaluating Large Language Models Trained on Code
Reference 293
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 01a60a53-9413-4c31-ad2c-f6edc4d484ec · inbound
Inner Monologue: Embodied Reasoning through Planning with Language Models Evaluating Large Language Models Trained on Code
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb97fade-0110-464f-970f-bf1335bcdfe2 · inbound
CodeT: Code Generation with Generated Tests Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 06c16986-b310-4bc6-955a-669e63512121 · inbound
Efficient Training of Language Models to Fill in the Middle Evaluating Large Language Models Trained on Code
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74b0ac66-0753-4936-9023-05fada10f7cc · inbound
Code as Policies: Language Model Programs for Embodied Control Evaluating Large Language Models Trained on Code
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36599aaf-edd3-4eb9-bc9a-94c66fab9273 · inbound
In-context Learning and Induction Heads Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d449c4a6-eda1-407a-a4b3-badd8195d43d · inbound
Automatic Chain of Thought Prompting in Large Language Models Evaluating Large Language Models Trained on Code
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6dccc3fc-f32c-40ab-ab2f-3159e18ec2cb · inbound
Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5181052-74b9-4b62-85b0-888dd0143191 · inbound
Large Language Models Are Human-Level Prompt Engineers Evaluating Large Language Models Trained on Code
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6764f59b-4615-411b-b4e6-a484b474e8fc · inbound
BLOOM: A 176B-Parameter Open-Access Multilingual Language Model Evaluating Large Language Models Trained on Code
Reference 216
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3fcae19-54a6-4b3d-9b9f-fab6720b085d · inbound
PAL: Program-aided Language Models Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82c4473d-ed10-4232-b015-4be5bb83161a · inbound
Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ec2040a-7fd2-4c63-8284-31f88fc54b0a · inbound
Solving math word problems with process- and outcome-based feedback Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52982979-213a-4ff0-be26-3f435df510de · inbound
Accelerating Large Language Model Decoding with Speculative Sampling Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47e3fefa-0ee8-4987-8652-9fbb600fda91 · inbound
Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents Evaluating Large Language Models Trained on Code
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ecc77b02-af9a-448b-8d25-13668bdbb813 · inbound
ViperGPT: Visual Inference via Python Execution for Reasoning Evaluating Large Language Models Trained on Code
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7eae23b-fb87-4e08-8650-99cff9ad5410 · inbound
ART: Automatic multi-step reasoning and tool-use for large language models Evaluating Large Language Models Trained on Code
Reference 154
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4c563659-1881-4c50-b6b7-c05579685761 · inbound
Reflexion: Language Agents with Verbal Reinforcement Learning Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1538bedc-91a6-442b-98e7-90d2e42c6447 · inbound
BloombergGPT: A Large Language Model for Finance Evaluating Large Language Models Trained on Code
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f5eb956-cc35-4f21-b392-fb867a056493 · inbound
Self-Refine: Iterative Refinement with Self-Feedback Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32eca035-f54f-44af-a2e2-d5e3466ffe5e · inbound
CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model Society Evaluating Large Language Models Trained on Code
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7094ea0-7540-476c-9035-18fbb3431e0a · inbound
A Survey of Large Language Models Evaluating Large Language Models Trained on Code
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90bc66e9-a994-42a3-81b1-f11b4bd08253 · inbound
Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling Evaluating Large Language Models Trained on Code
Reference 174
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 471a6759-2deb-4cb2-b67d-8222d50fdd44 · inbound
Teaching Large Language Models to Self-Debug Evaluating Large Language Models Trained on Code
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef3ea066-0f13-48cd-8186-9dc0bbfc0b41 · inbound
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3142dfeb-939e-48b8-b326-b526bf136cc4 · inbound
Usenix'23 Extended Version: Smart Learning to Find Dumb Contracts Evaluating Large Language Models Trained on Code
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6515b73d-d5ef-4b96-8183-8eef4c086a3a · inbound
LLM+P: Empowering Large Language Models with Optimal Planning Proficiency Evaluating Large Language Models Trained on Code
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6317b66e-5a3f-4d89-9c99-943eaed98f99 · inbound
WizardLM: Empowering large pre-trained language models to follow complex instructions Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f0de9db-4d67-41a8-ad68-edb9b52b77f3 · inbound
Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation Evaluating Large Language Models Trained on Code
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0565ccb-3356-4837-804b-e3faacff4760 · inbound
CodeT5+: Open Code Large Language Models for Code Understanding and Generation Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17618bc9-caf6-462c-b5d3-9d1332ad2364 · inbound
PaLM 2 Technical Report Evaluating Large Language Models Trained on Code
Reference 264
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6567f0fd-bfbd-4b6f-9153-a723e5876d83 · inbound
Gorilla: Large Language Model Connected with Massive APIs Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d3d2d01-8e4e-49e1-8a9a-d1204b262bbf · inbound
The False Promise of Imitating Proprietary LLMs Evaluating Large Language Models Trained on Code
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb29a2bb-cb55-4441-bd02-72f755dda231 · inbound
Voyager: An Open-Ended Embodied Agent with Large Language Models Evaluating Large Language Models Trained on Code
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 112537c4-7ddb-4d52-aa96-e2f4b5222e23 · inbound
RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems Evaluating Large Language Models Trained on Code
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4505964-88d7-4bcd-8a49-358f75ddea94 · inbound
Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ce7a911-ee76-488e-9550-e17528221591 · inbound
Textbooks Are All You Need Evaluating Large Language Models Trained on Code
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea4a1be1-5845-463d-aa61-6400ae1fcd26 · inbound
A Comprehensive Overview of Large Language Models Evaluating Large Language Models Trained on Code
Reference 141
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d37fbb32-2122-4138-829c-888fe3e76966 · inbound
Towards General Text Embeddings with Multi-stage Contrastive Learning Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fa05cf2-719f-4576-b13f-217c51a461fd · inbound
A Survey on Large Language Model based Autonomous Agents Evaluating Large Language Models Trained on Code
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da439efc-6a4a-4058-bf30-485a9a3ea435 · inbound
LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding Evaluating Large Language Models Trained on Code
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 05d46dd0-c7a1-4155-8ad5-65f64a18c6cd · inbound
Cognitive Architectures for Language Agents Evaluating Large Language Models Trained on Code
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c3163cb-1430-4a58-b9f6-7db7f94f6a48 · inbound
Textbooks Are All You Need II: phi-1.5 technical report Evaluating Large Language Models Trained on Code
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 924c57c6-2071-4f27-aba3-a8b1cca610d9 · inbound
MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa3293c6-1d77-49e6-addc-9ad3407bc999 · inbound
Efficient Memory Management for Large Language Model Serving with PagedAttention Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99d3b1b9-f0be-452d-939a-be6269844720 · inbound
Baichuan 2: Open Large-scale Language Models Evaluating Large Language Models Trained on Code
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed5ea7e5-b033-4bab-a875-b2100c4b73a4 · inbound
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models Evaluating Large Language Models Trained on Code
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4eb5ab85-a01b-4997-add8-ade633abde6d · inbound
UltraFeedback: Boosting Language Models with Scaled AI Feedback Evaluating Large Language Models Trained on Code
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 486b472f-e689-4023-a6aa-ae37e5da4134 · inbound
Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs Evaluating Large Language Models Trained on Code
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1670ac2a-2c7c-447a-b2de-eefb37389b0c · inbound
Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c43e08e5-d1d3-4a8a-b658-9a665e2ca7be · inbound
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b5da458-943f-4ea4-8ff2-cbceb8884217 · inbound
Instruction-Following Evaluation for Large Language Models Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fb12c4d-9e66-4641-819d-85b575b9fb02 · inbound
GAIA: a benchmark for General AI Assistants Evaluating Large Language Models Trained on Code
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86df73b4-9a1e-4a98-b88a-513abe34604e · inbound
The Falcon Series of Open Language Models Evaluating Large Language Models Trained on Code
Reference 258
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2aea54e5-25c2-4a7b-92f2-0b9db08103a8 · inbound
Gemini: A Family of Highly Capable Multimodal Models Evaluating Large Language Models Trained on Code
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 392b939c-156e-4b0a-b9b7-a620cad38e2b · inbound
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16881213-f8ee-4503-97ac-bbe7fe8677cb · inbound
AppAgent: Multimodal Agents as Smartphone Users Evaluating Large Language Models Trained on Code
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6796b5b-97de-47ea-9eff-ceb1d1ecc1fb · inbound
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8837baa4-7275-4218-9361-69eb425cd406 · inbound
Mixtral of Experts Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34f954ac-2378-40c7-b8f0-952895f46ceb · inbound
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models Evaluating Large Language Models Trained on Code
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c213a35c-e268-4bec-a6db-a6046d8994d7 · inbound
Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads Evaluating Large Language Models Trained on Code
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43f51345-189e-4143-afc2-1be62d58aca8 · inbound
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a71d9742-1d95-49b4-8097-036a2ab81432 · inbound
EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty Evaluating Large Language Models Trained on Code
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19c9ef4a-20ee-4e15-978b-e790df07fbf1 · inbound
KTO: Model Alignment as Prospect Theoretic Optimization Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 278ac788-5422-4d60-a9ab-f9e4c9134f90 · inbound
CodePori: Large-Scale System for Autonomous Software Development Using Multi-Agent Technology Evaluating Large Language Models Trained on Code
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da414c94-9568-468d-9a03-b31188e6f4a0 · inbound
Understanding the planning of LLM agents: A survey Evaluating Large Language Models Trained on Code
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 699fccd2-077e-417f-8499-5ee20b57a0c9 · inbound
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6709153-ff5b-481d-99be-d9fb52509ddc · inbound
Large Language Models: A Survey Evaluating Large Language Models Trained on Code
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa7b790d-53b2-4273-8abe-1a9e95f6fe11 · inbound
Massive Activations in Large Language Models Evaluating Large Language Models Trained on Code
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97f96a46-0a2e-4bdf-975e-a3427da73404 · inbound
StarCoder 2 and The Stack v2: The Next Generation Evaluating Large Language Models Trained on Code
Reference 178
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3090d910-ae40-4419-89ad-3daa61d6b655 · inbound
Retrieval-Augmented Generation for AI-Generated Content: A Survey Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40cef2d6-289a-462e-bf7e-66a1147d7f6e · inbound
Yi: Open Foundation Models by 01.AI Evaluating Large Language Models Trained on Code
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c386a1a3-9b7d-4d51-83e2-b0868df0439d · inbound
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code Evaluating Large Language Models Trained on Code
Reference 249
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dfae1096-f9d7-4dd2-a0a3-620d0a9f95f6 · inbound
Gemma: Open Models Based on Gemini Research and Technology Evaluating Large Language Models Trained on Code
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 699041f8-d3d4-4e55-a540-635c8858ca67 · inbound
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair Evaluating Large Language Models Trained on Code
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd78022a-b462-4c63-9b6c-e89b5843aa7b · inbound
InternLM2 Technical Report Evaluating Large Language Models Trained on Code
Reference 191
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 02139f0a-55b8-4909-bfc4-901f25b862aa · inbound
Jamba: A Hybrid Transformer-Mamba Language Model Evaluating Large Language Models Trained on Code
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8b9e1a8-b2e7-43ed-85cb-27d49b46c16f · inbound
Assessing, Exploiting, and Mitigating Syntactic Robustness Failures in LLM-Based Code Generation Evaluating Large Language Models Trained on Code
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b3b96ea-2cd0-4073-8d08-9f28f1020137 · inbound
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies Evaluating Large Language Models Trained on Code
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d430552-d115-4ebb-9305-7448586a8337 · inbound
A Survey on Retrieval-Augmented Text Generation for Large Language Models Evaluating Large Language Models Trained on Code
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07ef32fc-0e55-4c63-95a6-257ed1fefc4f · inbound
A Survey on Efficient Inference for Large Language Models Evaluating Large Language Models Trained on Code
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1ef9dcd8-8003-41ee-bb8f-2402bdfc3d81 · inbound
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6649110d-78d1-477a-b3b6-ee9a4fc1006d · inbound
Better & Faster Large Language Models via Multi-token Prediction Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b4d7b31-7be9-435f-bf76-02761533340a · inbound
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6692694d-d432-4fa5-99a5-fb7d40a69996 · inbound
Lessons from the Trenches on Reproducible Evaluation of Language Models Evaluating Large Language Models Trained on Code
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a5ca1a7-5fb6-4eed-951a-31f158165c7b · inbound
A Survey on Large Language Models for Code Generation Evaluating Large Language Models Trained on Code
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.