Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2405.18392.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:32:05.909089Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 5c6e8e3e-a02d-4d8b-959e-cbcb79d340f2 · inbound
Optimization Hyper-parameter Laws for Large Language Models Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation df95fbbe-38b5-4452-b272-d08a71de1ef3 · inbound
PoM: Efficient Image and Video Generation with the Polynomial Mixer Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4912248d-357f-4082-a29c-8c9bb19db3b5 · inbound
INTELLECT-1 Technical Report Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06527b53-fad6-4c37-9a85-92dbe8aa7470 · inbound
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8d53b71-3169-4cce-ae0e-ce8fad15b0d4 · inbound
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 141
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9faea4ab-6f22-4e48-8bd1-ffc020328510 · inbound
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 148
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6a403ff7-89b4-4607-92e6-f7666643267d · inbound
YuLan-Mini: An Open Data-efficient Language Model Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 112eb45b-9638-4c86-8dcf-ab089d0b216b · inbound
METAGENE-1: Metagenomic Foundation Model for Pandemic Monitoring Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed6f8b7d-04b4-425d-acf6-25bc03cc08c1 · inbound
Proxies for Distortion and Consistency with Applications for Real-World Image Restoration Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8481af53-3f09-4cf1-b860-802cd22b8d79 · inbound
SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 177
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 50ca0f07-5562-4a8f-9e7f-74178882b260 · inbound
Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b1cf965-f320-43ba-be7d-75782d723954 · inbound
Trillion 7B Technical Report Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42786b4f-a28b-49c1-8c5f-56eee43f816b · inbound
BioVFM-21M: Benchmarking and Scaling Self-Supervised Vision Foundation Models for Biomedical Image Analysis Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1428323-a73c-4f23-8e36-d0a5ebcdb450 · inbound
BioClinical ModernBERT: A State-of-the-Art Long-Context Encoder for Biomedical and Clinical NLP Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbf646a2-602d-4105-bf75-187a47e1043a · inbound
Is your batch size the problem? Revisiting the Adam-SGD gap in language modeling Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a237938b-8a27-45c8-befd-83775802303a · inbound
The Automated LLM Speedrunning Benchmark: Reproducing NanoGPT Improvements Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a50d57de-612e-4e2e-b028-99df93492290 · inbound
AbbIE: Autoregressive Block-Based Iterative Encoder for Efficient Sequence Modeling Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ab7824-863b-428a-92a9-768cb60b62eb · inbound
Analysis of Schedule-Free Nonconvex Optimization Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f919b332-930b-47ca-94f3-762fe9bce088 · inbound
Foundation Models for Discovery and Exploration in Chemical Space Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 281
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation faa0eda6-b091-4192-941c-e878bdf1ea6c · inbound
Scaling-Aware Data Selection for End-to-End Autonomous Driving Systems Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0d3a9710-b345-4969-9776-c3b3da0ec585 · inbound
Scaling Laws for Mixture Pretraining Under Data Constraints Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1c0c4aa7-a990-4f60-8fe7-430a2084560b · inbound
Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bc7b8a2f-2565-494d-a659-9a50a5bd435c · inbound
Anytime Training with Schedule-Free Spectral Optimization Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 96f3dcb0-b37b-470b-a982-57501ec69192 · inbound
Mellum2 Technical Report Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 137949ea-f578-451a-b46b-ffe82b532b32 · inbound
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 605173fd-8a1f-4072-8768-fee7b95d5b2d · inbound
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db6076a0-f5b2-49bf-89e4-fff2dbae6c71 · inbound
Scaling Point-in-Time Language Models Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 861e5366-811d-4e69-be36-4504f6cc5566 · inbound
Compute-Optimal Is Not Cluster-Optimal: Systems-Aware Scaling for Sparse Mixture-of-Experts Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.