Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:32:47.917300Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2602.19938.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:32:47.917300Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c5649613-9af0-490d-aee7-85a81c5d09e0 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Deep Rewiring: Training very sparse deep networks
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26f304c9-6c42-4ac7-8b60-0fc3bfc08edf · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fba3b89-ab0e-4b0e-a054-683de3c8271a · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Towards MoE Deployment: Mitigating Inefficiencies in Mixture-of-Expert (MoE) Inference
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be9d72b0-5d09-4624-a639-712d70c7d671 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs MathPrompter: Mathematical Reasoning using Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a16666b4-85b6-4f95-8b37-244edddc5327 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Scaling Laws for Neural Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf7cb693-3216-476d-aae0-c18d55c558fc · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17699da9-e3dd-44f2-b44e-2c75b053d9d0 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecd4fd35-43a1-4ae2-ad23-d9c4901028dd · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs TruthfulQA: Measuring How Models Mimic Human Falsehoods
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e87c76d-2f96-4b48-a16d-4167011f5afb · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs AlphaPruning: Using Heavy-Tailed Self Regularization Theory for Improved Layer-wise Pruning of Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e41dd5c-4aa4-40b2-b200-12d63ee3e27c · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs LLM-Pruner: On the Structural Pruning of Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a8a08df-5f80-47b7-9ffc-862559cd3f3f · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Pruning Convolutional Neural Networks for Resource Efficient Inference
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3ec516e-e2c7-4ccf-849a-da6909f3db15 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs https://aclanthology.org/Q19-1016/
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0240ea56-1bcd-4876-9813-92b052dcdf30 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Systematic Characterization of LLM Quantization: A Performance, Energy, and Quality Perspective
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 640d47b3-c827-4530-8609-4c1171b90e33 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs MathCoder: Seamless Code Integration in LLMs for Enhanced Mathematical Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5747fa34-0ae2-40c9-a46c-b9e4e25a1ec5 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Auxiliary-Loss-Free Load Balancing Strategy for Mixture-of-Experts
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd601836-1c38-4c44-8c73-b5ac115df73c · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dab9dbf-2ff2-4733-a74f-14cdf17ce36f · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Incremental Network Quantization: Towards Lossless CNNs with Low-Precision Weights
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0bd5800-dffa-442c-b380-5b2bef93cdc7 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Exploring Sparse MoE in GANs for Text-conditioned Image Synthesis
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72bf5afb-77d3-4ae0-af99-d779c209ab0a · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs ST-MoE: Designing Stable and Transferable Sparse Expert Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a40d2930-fdbf-445b-9d2c-2df26913bf5a · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Mixtral of Experts
Reference 1991
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba0ceff-4272-4239-832b-8fa7681de94f · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs GRACE-MoE: Grouping and Replication with Locality-Aware Routing for Efficient Distributed MoE Inference
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c82cef0-921c-4e14-b83f-4b266164cd0b · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12f5d8b2-ac3f-44c6-8603-a738968efac1 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs PIQA: Reasoning about Physical Commonsense in Natural Language
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30dd69dd-16ed-488e-9087-118f33bdf256 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a13b915e-d7a6-4cbf-8d5e-9333396d095c · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Load balancing mixture of experts with similarity preserving routers, 2025.https://arxiv.org/abs/2506.14038
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e2c223-ba81-467f-ae84-d88585e7db20 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Sergey Zagoruyko and Nikos Komodakis
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e7364a-526e-43d0-b9ca-bdd52b718f6d · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs WinoGrande: An Adversarial Winograd Schema Challenge at Scale
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91ba37c3-0a0d-4940-9a6c-5cc4c030616b · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs PaLM: Scaling Language Modeling with Pathways
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60a34feb-a324-4079-a8c2-30d24acc0b46 · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa8540e2-59da-45f2-89cc-8733b036db1e · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5098165-2ac2-4d9e-aede-d79affeade7e · outbound
A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs Zachary Doucet, Rishi Sharma, Martijn de Vos, Rafael Pires, Anne-Marie Kermarrec, and Oana Balmau
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.