Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T05:30:11.389756Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2510.18245.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T05:30:11.389756Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1da1a75f-503e-45d4-8765-a892a0188f69 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Phi-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 709620e1-3cdb-448e-ae3e-15838ae8e350 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Parameters vs FLOPs: Scaling Laws for Optimal Sparsity for Mixture-of-Experts Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bdc0747a-befe-4b43-9c4a-b213055b852e · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a6499279-3b04-4907-899c-77f49c282663 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 145ea757-62b2-414d-bff1-47c0818dd646 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Program Synthesis with Large Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b109582b-e94e-4858-853b-129ad6b632fc · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scaling Inference-Efficient Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3155b484-d3d6-4f13-9fd1-b30a2f9a2bf6 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5d40e9f-ebb8-4d07-be6e-17c8a44c3b17 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4891e28d-626d-4305-8fa6-0820b87ffeac · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Exploring Diffusion Transformer Designs via Grafting
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 83b1f712-d611-498f-9b56-7fe78a8936e6 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Evaluating Large Language Models Trained on Code
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 29230e97-6ab5-4f76-b4cd-0ae41a7c2b4f · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scaling Law for Quantization-Aware Training
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 34b87522-9e74-4c47-b8f5-bfbac89ef4f7 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Reducing the carbon impact of generative ai inference (today and in 2035)
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f971779c-a9fc-4fd9-a9fd-738e40178cc9 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ff32e271-e228-46d2-9303-efb746bd19f2 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Training Verifiers to Solve Math Word Problems
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6f7ef904-a7d1-46ba-b1c4-818fae591acc · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e85a2c3f-e98e-4305-be55-f6fefa8f9f0c · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs The Llama 3 Herd of Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5b57a8b-8ff2-4038-9194-3c93348bc488 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Language models scale reliably with over-training and on downstream tasks
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b9e6101-b575-4ecc-8611-f7a4b355e517 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs He, B., Yin, L., Zhen, H.-L., Liu, S., Wu, H., Zhang, X., Yuan, M., and Ma, C
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2b12d6ce-7fe1-40c5-baf4-25f423b5945b · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d14a1da3-bbd1-4d8c-b2d8-95c2ae267dc2 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 26c4dd41-d428-496f-94a5-2159be8a4e24 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Measuring Mathematical Problem Solving With the MATH Dataset
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 313ae315-8dad-482c-899a-14b5dbf82c82 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Training Compute-Optimal Large Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f662aaaf-2814-40c7-9a9f-294ddf7af60d · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 87954259-ef6d-4d24-a4b2-7bf3539ab4e2 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 830e0122-ba17-4433-8bad-3fd466451386 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scaling Laws for Neural Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8e35aa42-8db1-4f09-a7d3-54e90d3459ab · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scaling Laws for Fine-Grained Mixture of Experts
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 77a1a9f5-b51f-4a18-b02d-d7fb973e58e6 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scaling Laws for Precision
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e24042eb-c1cc-4cb0-bca9-681c1e1b0b19 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs DeepSeek-V3 Technical Report
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21a2512e-a594-4953-aefc-4f2821587e21 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b62a3903-ef66-4613-baba-02b1013818cf · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs The LAMBADA dataset: Word prediction requiring a broad discourse context
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cda07176-c96d-4846-9525-0d478dccbbda · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs The Impact of Depth on Compositional Generalization in Transformer Language Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e49369ee-2d24-4817-8bbe-d18c2ee685ec · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f272d83c-a6b2-4c07-b2f3-06e2045b63b0 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Observational Scaling Laws and the Predictability of Language Model Performance
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 64973a38-ee69-4354-a109-9d52931c3803 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 988fb2ff-4019-4f6b-9cd8-5ddf3e42bffa · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 02ad5651-0047-4b35-8e23-3585d10ce9f5 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e7ebcbc1-7631-4364-8326-91ff1f732d3c · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a349a68-e32e-4bc8-9752-804a537431a2 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ad6a4eb1-b724-46c8-9c8d-59507364db17 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b10b1dce-e4e7-4b07-b21d-a88466423d0d · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Gemma: Open Models Based on Gemini Research and Technology
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 693b4fc7-6cb5-4544-bafe-17a44bd934c4 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs LLaMA: Open and Efficient Foundation Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5d0bb1cb-a0b8-4006-b2cc-121d4b7df4e1 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a32cf55a-0359-4b57-b055-0f8711cce19d · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Emergent Abilities of Large Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 808db350-a29a-451b-8c6a-d6ad43d91182 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Crowdsourcing Multiple Choice Science Questions
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3dc8d93e-5ac6-4414-b19f-b16be7165640 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Efficient Streaming Language Models with Attention Sinks
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3da75dd6-ac2e-46ec-b23c-ecce9857d078 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 72fca583-44f3-4719-980d-b715d56f87f3 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Qwen3 Technical Report
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 366e5d4d-bf24-4113-9445-f6f00f570e4d · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 817612f6-1772-48df-8973-acaaf0cd85c5 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3e18282a-a525-43b9-8ffc-343806a7f11b · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c338478b-de5f-4260-961b-ead6fc97eb4e · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs TinyLlama: An Open-Source Small Language Model
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 85388c2f-c1de-4033-879f-dc63fd0b9bcc · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs It was not used to generate research ideas
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5fc007a8-5807-40d6-89ee-8db1be66b5e0 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Across varying batch sizes and model scales, larger hidden sizes yield higher inference throughput under a fixed parameter budget
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2b9ec17b-1782-4749-8eca-43101dfe8541 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Moreover, while multiplicative and additive calibrations differ in formulation, their MSE and Spearman values remain nearly identical
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8cfb120f-2eff-4147-9a34-e2b24c757cf4 · outbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
No inbound Pith citation observations are available.