Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2405.18392.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:06:23.521445Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 5c6e8e3e-a02d-4d8b-959e-cbcb79d340f2 · inbound
Optimization Hyper-parameter Laws for Large Language Models Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d8d53b71-3169-4cce-ae0e-ce8fad15b0d4 · inbound
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 141
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9faea4ab-6f22-4e48-8bd1-ffc020328510 · inbound
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 148
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8481af53-3f09-4cf1-b860-802cd22b8d79 · inbound
SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 177
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 50ca0f07-5562-4a8f-9e7f-74178882b260 · inbound
Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1428323-a73c-4f23-8e36-d0a5ebcdb450 · inbound
BioClinical ModernBERT: A State-of-the-Art Long-Context Encoder for Biomedical and Clinical NLP Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbf646a2-602d-4105-bf75-187a47e1043a · inbound
Is your batch size the problem? Revisiting the Adam-SGD gap in language modeling Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a237938b-8a27-45c8-befd-83775802303a · inbound
The Automated LLM Speedrunning Benchmark: Reproducing NanoGPT Improvements Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a50d57de-612e-4e2e-b028-99df93492290 · inbound
AbbIE: Autoregressive Block-Based Iterative Encoder for Efficient Sequence Modeling Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ab7824-863b-428a-92a9-768cb60b62eb · inbound
Analysis of Schedule-Free Nonconvex Optimization Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f919b332-930b-47ca-94f3-762fe9bce088 · inbound
Foundation Models for Discovery and Exploration in Chemical Space Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 281
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation faa0eda6-b091-4192-941c-e878bdf1ea6c · inbound
Scaling-Aware Data Selection for End-to-End Autonomous Driving Systems Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0d3a9710-b345-4969-9776-c3b3da0ec585 · inbound
Scaling Laws for Mixture Pretraining Under Data Constraints Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1c0c4aa7-a990-4f60-8fe7-430a2084560b · inbound
Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bc7b8a2f-2565-494d-a659-9a50a5bd435c · inbound
Anytime Training with Schedule-Free Spectral Optimization Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 96f3dcb0-b37b-470b-a982-57501ec69192 · inbound
Mellum2 Technical Report Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 137949ea-f578-451a-b46b-ffe82b532b32 · inbound
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 605173fd-8a1f-4072-8768-fee7b95d5b2d · inbound
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db6076a0-f5b2-49bf-89e4-fff2dbae6c71 · inbound
Scaling Point-in-Time Language Models Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.