Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:46:20.720359Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2411.17927.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:46:20.720359Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 416a0a09-8e82-45ef-8f66-1d9a829c47e1 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Emergent abilities of large language models,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5bc978b2-c081-4760-a448-9185af02ba53 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Measuring Data
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61a3bf9a-5e2c-4cbc-a950-ebcc70b0facc · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Codegen: An open large language model for code with multi-turn program synthesis,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 711e452f-e0a4-4044-a102-69c20d0a9b10 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Codegen2: Lessons for training llms on programming and natural languages,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 668ce3b1-fe11-4e53-8a0a-79854a11e0d2 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Are emergent abilities of large language models a mirage?
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0c6cc93c-b5de-42ef-96f5-49d776dd93db · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Are emergent abilities in large language models just in-context learning?
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 60f56d2f-ec20-49d0-9401-6fff90858473 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Beyond the imitation game: Quantify- ing and extrapolating the capabilities of language models,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b83ed9ba-fb9c-4d6d-ba85-8e463f8b74a7 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Beyond accuracy: Behavioral testing of nlp models with checklist,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f5abe65a-cb60-4d74-a6e0-bb1042d0500e · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? A Structured Review of the Validity of BLEU,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59fd066c-88d0-4a62-adbb-23a430fbfd9c · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? ORANGE: a method for evaluating automatic evaluation metrics for machine translation,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e9d362b6-e235-47ec-9ee4-4f28c95c61ec · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Codebleu: a method for automatic evaluation of code synthesis,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b9992ad7-473d-4397-92b5-725acb4cd55b · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Bleu: a method for automatic evaluation of machine translation,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5b8c493f-07b1-42b4-83ba-283f5fffc07c · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4674ae1c-6d29-474e-84dc-8053821ecbf2 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Pypi/k4black/codebleu,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b18a490e-abf3-40d8-95a8-6e2dbbca4656 · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Commit message generation for source code changes,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5577396c-4775-498a-b4ba-2dc8edd8f87b · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Using large language models for commit message generation: A preliminary study,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8529a88b-818d-46be-844e-261a1e4fd7fb · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Emergent capabilities of LLMs for software engineering,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation aa2e705e-ee08-4f5c-b3d3-af45a455abbc · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? On the evaluation of commit message generation models: An experimental study,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c5ed9eb3-ab0f-4eeb-b045-b066387e313b · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? On the dangers of stochastic parrots: Can language models be too big?
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a61fe14-d7c4-4141-bbcf-41c14aea55bc · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Detecting emergent intersectional biases: Contextualized word embeddings contain a distribution of human-like biases,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9fe2d3b-8c2d-4803-b330-362a36cf7ccc · outbound
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We? Social biases in NLP models as barriers for persons with disabilities,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.