Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2312.17238.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:58:49.567474Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:39:56.479057Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b7b9dc76-8433-4408-b2b3-edbc5d6a2998 · inbound
Mixtral of Experts Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 57cbe660-d453-4c4b-b85b-fd89038d3988 · inbound
Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55ab2dfb-5b33-4648-897b-1a4f4fa7ad5f · inbound
On-the-Fly Adaptive Distillation of Transformer to Dual-State Linear Attention Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a3c1a15-ecce-4ede-abb9-dc516e42cffe · inbound
Cache Management for Mixture-of-Experts LLMs -- extended version Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50f6c9cc-1a75-4bd0-85ca-829cd62620d5 · inbound
MoE-Compression: How the Compression Error of Experts Affects the Inference Accuracy of MoE Model? Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 875b6dce-b865-416e-833d-d33894ade016 · inbound
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cdaeee-e4d1-4928-81bc-99f65c599ee5 · inbound
Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55c79c1e-5af5-4994-ab70-2257e512a681 · inbound
A Scheduling Framework for Efficient MoE Inference on Edge GPU-NDP Systems Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5ea1806-5533-4d32-b8f8-4fa5b08bc178 · inbound
ZipMoE: Efficient On-Device MoE Serving via Lossless Compression and Cache-Affinity Scheduling Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 279b913d-45ce-4afd-96e1-dee673e1c68f · inbound
Temporally Extended Mixture-of-Experts Models Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0d89220-01c7-4c6a-8d56-df888f7fefab · inbound
Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8d31969-13bf-4095-be85-c883a2df96c4 · inbound
VisMMOE: Exploiting Visual-Expert Affinity for Efficient Visual-Language MoE Offloading Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27eaf72f-a32c-4b38-86b6-2340c878028f · inbound
TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a20efcfd-4283-4bd5-b036-b37f7283b9fc · inbound
PALS: Power-Aware LLM Serving for Mixture-of-Experts Models Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd0db0fd-e662-4314-b0aa-de69969497cd · inbound
Does Mixture-of-Experts Actually Help Inference on Consumer and Edge Hardware? An Empirical Study Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 77fe61f6-e4f1-485d-a0f7-e05a193591eb · inbound
Does Mixture-of-Experts Actually Help Inference on Consumer and Edge Hardware? An Empirical Study Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ad6a6fe-0125-4b4e-8410-e7ef99d3df00 · inbound
WiSP: A Working-Set View of Mixture-of-Experts Serving on Extremely Low-Resource Hardware Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b47a5de4-1eeb-449d-bbfc-1e1b32b55484 · inbound
GeMoE: Gating Entropy is All You Need for Uncertainty-aware Adaptive Routing in MoE-based Large Vision-Language Models Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b72e20b-0132-473b-bccb-7cc1e8163784 · inbound
Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 778f6169-e786-4b4a-afd2-887bcfc3461e · inbound
Broadcast Rate Limits in Wi-Fi: A Forgotten Bottleneck for Collaborative Edge LLM Inference Fast Inference of Mixture-of-Experts Language Models with Offloading
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.