Pith. sign in

Paper Citation Record · LEDGER

MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2411.01016.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.01016 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:09.523179Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T17:14:57.406110Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0272af74-5132-471d-816a-2bfd2ee463fa · inbound

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging cites this paper.

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:09.523179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:09.523179Z digest=sha256:ec2940722df90f27bebc87c0fa05f28cb465736d6e74165ec827f931dc20e162

Observation adc35558-fd31-43f6-910d-11c348de5fea · inbound

LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference cites this paper.

LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T11:30:14.314029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:30:14.314029Z digest=sha256:79d0f2980bc470c3ed4ab0175eeca33c7c81116de754b835568eb019cd062828

Observation c3df6d3c-33f2-43a4-89fc-75718c3ca28c · inbound

Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs cites this paper.

Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T17:57:25.224160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:57:25.224160Z digest=sha256:318e5d29e504e1b1a327a058d5486006853c6dfc5d95ee8413f3e60e8684174b

Observation c986254d-4e84-45a7-81e5-e58639686ea9 · inbound

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference cites this paper.

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T23:41:52.927468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T23:41:52.927468Z digest=sha256:8801d95d4de4209c30045f304c11ef49b24e731e40f285a7d034dec8701fda51

Observation ba5cfb40-d5cc-4cea-be5f-101809737376 · inbound

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits cites this paper.

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T09:32:55.511235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:32:55.511235Z digest=sha256:806f134a9db36e3b809003196e602a1a5ad4108805d2c8df4c88f8e1b00f7a9d

Observation cd1471e9-d9dd-431f-b701-928b761f5709 · inbound

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE cites this paper.

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:35:55.646451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T14:34:48.524592Z digest=sha256:69300bae32032cf7303cc503c5333430ce05f96a15d2e09ab3dc103763880d03

Observation db7cfda7-5df8-4f90-82bf-f193ef37c6fb · inbound

Path-Constrained Mixture-of-Experts cites this paper.

Path-Constrained Mixture-of-Experts MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:19:54.276593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T09:16:06.226566Z digest=sha256:da0e838dd34692b982768e132ad48d061d95fb927f4af48d0a6a82bf9469e88a

Observation 95615c89-1ce1-4a93-948e-e8717e670abb · inbound

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving cites this paper.

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:13.280550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:16:16.466375Z digest=sha256:754ee047381239259937c496977fa73c54dcf18df3ea9a0cbbcfc4e440f1e50d

Observation d58bc009-1d59-459c-9739-63e84a7a37cd · inbound

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization cites this paper.

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:44:48.466223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T15:38:18.616792Z digest=sha256:f2726d37f723fd2a4fecd4f53da668f4956038ad7a9937676d2bcc9834d021a7

Observation 528da36e-b37d-4f86-af60-f3647c4a628a · inbound

Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference cites this paper.

Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:14:57.407554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T04:27:02.915854Z digest=sha256:74fa8b77fec3eede35caec98756d040137468ff6cefe66026eeb72f746cf7223

Observation 49ffc646-cb1b-4427-b0e8-255ff2f9e063 · inbound

Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference cites this paper.

Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T08:35:22.347459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:35:22.347459Z digest=sha256:56ffcf8e5b73de899809f9d178a898d046927453b2593aac08d24a645e7a0f06