Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2503.04715.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T05:27:10.359838Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 5d54bad5-dbfe-43d1-a3f7-0e20e17efeee · inbound
Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d0f628a8-daea-4fbf-8d14-e5bcf9243c97 · inbound
SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e25eb6b-12cf-4bb2-9d36-37a8aecf8ea8 · inbound
OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa01726a-0612-4346-824f-30a2507a8b6f · inbound
Rethinking Language Model Scaling under Transferable Hypersphere Optimization Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89c6de2c-1130-4202-8f08-e3cd7811a13d · inbound
AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration - Learning from Cheap, Optimizing Expensive Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c3d8b653-6d90-4b50-b188-c87cb9dda4be · inbound
AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration - Learning from Cheap, Optimizing Expensive Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c272cd62-9e05-4b36-84e8-3bb687330044 · inbound
BoLT: A Benchmark to Democratize Black-box Optimization Research for Expensive LLM Tasks Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 931597dd-a8b5-4142-8954-839468e0a7bc · inbound
Quantifying Hyperparameter Transfer and the Importance of Embedding Layer Learning Rate Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 78f9b9c1-6564-4ea0-a71e-b01b07db99e1 · inbound
Staged Factorial Screening for Budget-Constrained Micro-Pretraining Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a77351ce-5419-44df-8d6a-cae45a7be5d3 · inbound
Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b36f3396-bd91-48a9-827c-ec445c2739c4 · inbound
Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation beb7a600-b152-4956-909c-b2676abab6d1 · inbound
MultiHashFormer: Hash-based Generative Language Models Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 23542f01-bb32-401a-9b90-90b1e45102bb · inbound
On the Nonlinearity of Learning Rate Scaling for LLM Training Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9e8febb8-fbee-4d59-87cf-6fc4f48cb2cf · inbound
How to Allocate Your Tokens? Scaling Laws with Training Steps and Batch Size Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f9a9a766-4339-4587-a696-5daaaf3a386d · inbound
Convolution for Large Language Models Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92ff9c0d-a7ff-4c5c-9f4c-26e6bab97084 · inbound
SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a71e5f3-88c8-4657-8a6b-b9ed478aebca · inbound
LLM-Based Generative Retrieval for Snapchat Content Recommendation Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.