Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:04:55.962152Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 1 inbound Pith citation observation for arXiv:2506.09099.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:04:55.962152Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T02:02:06.788056Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T07:46:25.975143Z
14 of 14 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8161e917-cecd-4e3a-a2a4-1c889c3a655e · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57c1af3c-30f3-4ba1-8fa7-3fe26dac7beb · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers 9, 2024); accessed May 19,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04fcb939-cc24-4f4e-89c4-98b4e2b86369 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers doi: 10.1016/j.neunet.2024.106550
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d55ccfad-9b27-495b-b0a9-0810362e5e9e · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Training language models to follow instructions with human feedback
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e040c504-c48e-47f9-8502-6c11da144c51 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 170e1a5b-513d-4023-b290-12eee9943dc2 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d414a4b0-7877-4fb0-bf7d-2931d92c6792 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Memorization Without Overfitting: Analyzing the Training Dynamics of Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43e43de8-b55d-4d46-8ce2-d425768a8620 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Grokked Transformers are Implicit Reasoners: A Mechanistic Journey to the Edge of Generalization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456e782f-fb2c-42de-9a69-8828c499f5c4 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Exploring Memorization in Fine-tuned Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48de51b3-0a53-43ee-91df-0e18c07519cb · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Goldilocks zone
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6083ae0c-f92c-4cd0-93e5-190a11b12064 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Gemini: A Family of Highly Capable Multimodal Models
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f86a55e6-0981-4687-a4ba-da29f92be046 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers Towards Understanding Grokking: An Effective Theory of Representation Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d4bf7da-300f-490b-b06e-29a15f30c65a · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers GPT-4o System Card
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b355b9-aa0d-4b4d-9c9b-0e98309b49f6 · outbound
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a6da3b7-5808-4654-9ae7-d0e6edac71f7 · inbound
Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.