Pith. sign in

Paper Citation Record · LEDGER

SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2402.09025.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.09025 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T11:47:24.885060Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:16:16.844716Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 49953636-8b04-4da7-a906-40a9c50a41a3 · inbound

DOCS: Quantifying Weight Similarity for Deeper Insights into Large Language Models cites this paper.

DOCS: Quantifying Weight Similarity for Deeper Insights into Large Language Models SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T11:47:24.885060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:47:24.885060Z digest=sha256:71151115496ee606564ae17909a8d1fc4c1a0b84a884cbee19f24cc97321efde

Observation 1c7748c0-9aa1-4f27-b552-43f5c056fbe4 · inbound

SlimLLM: Accurate Structured Pruning for Large Language Models cites this paper.

SlimLLM: Accurate Structured Pruning for Large Language Models SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:20.434754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:30:20.434754Z digest=sha256:fad1ace87056568c36e42eead6b4b7230f74f2bac4bee5dffa42bf932c87cdab

Observation b3007589-1249-4ed5-8778-84e7a8302faf · inbound

ProcrustesGPT: Compressing LLMs with Structured Matrices and Orthogonal Transformations cites this paper.

ProcrustesGPT: Compressing LLMs with Structured Matrices and Orthogonal Transformations SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:20:43.592215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:20:43.592215Z digest=sha256:053c4338574b7737e5b182e3ac4817fafa500932fa7cf9fac7274d6923fa279f

Observation 5b973fc4-827a-4292-85b3-0f0987a2249b · inbound

SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling cites this paper.

SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:51.611505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:51:51.611505Z digest=sha256:e94ae3c420da4c3a5d35aa00d05f3f66081482674f898e1c5898314a21ffe314

Observation a9097cef-a1d4-41e4-be55-c8db94e02f57 · inbound

DipSVD: Dual-importance Protected SVD for Efficient LLM Compression cites this paper.

DipSVD: Dual-importance Protected SVD for Efficient LLM Compression SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:58:09.407929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:58:09.407929Z digest=sha256:0287df5ce2f791f5f36598dd1bbfba4fe73163865662ee7f5782ed3a50d6fd07

Observation e6567f3c-3e5f-4d7b-ba81-7140ec8adb65 · inbound

Prune&Comp: Free Lunch for Layer-Pruned LLMs via Iterative Pruning with Magnitude Compensation cites this paper.

Prune&Comp: Free Lunch for Layer-Pruned LLMs via Iterative Pruning with Magnitude Compensation SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T14:48:30.383090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:48:30.383090Z digest=sha256:e7c0f8c8dd67d1861d000b2963a8ecab28e6547503ce88a72b37157fd34a132f

Observation ec0219c8-ea7a-437e-b6a6-643da53ca0c8 · inbound

Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers cites this paper.

Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:55:18.296382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:55:18.296382Z digest=sha256:929698adde70e94ef0b8bc0ed2fd28ae5d907e38839bcb32e917dc916e4537c2

Observation e3204733-27a3-425a-ace5-15b4671b5ec8 · inbound

Numerical Simulation for General Relativistic Magnetohydrodynamics in Dynamic Spacetimes cites this paper.

Numerical Simulation for General Relativistic Magnetohydrodynamics in Dynamic Spacetimes SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T20:23:27.311851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:23:27.311851Z digest=sha256:c3c2b5605325cbfa8b54297418c7ae3998af74604b592dec3242bc577f969734

Observation 4907db18-a43a-4d6d-bcc8-7e140a468017 · inbound

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models cites this paper.

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:40:46.132071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T08:40:24.822863Z digest=sha256:480d5ca8fdd60b69f5741245e3cb391147ab5d0400da3a9d817d4793173b9f74

Observation 038d7afa-0ff8-4214-83fd-66a97b849786 · inbound

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions cites this paper.

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:25:53.867897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T02:23:52.589354Z digest=sha256:e15396da180705a4453f3899519553a95bfc316ade2d509ea683703170da7731

Observation 79b74e79-8a38-4aae-9d17-4e830e76c8b0 · inbound

TAPIOCA: Why Task- Aware Pruning Improves OOD model Capability cites this paper.

TAPIOCA: Why Task- Aware Pruning Improves OOD model Capability SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:11:23.065258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-22T10:09:46.259358Z digest=sha256:c2f8297b94099d72b3382fea5e235c09591f4c64c3ca156b15fde4dae90078da

Observation 4077d087-7a4f-422a-9bb8-e93db051816a · inbound

No Free Swap: Protocol-Dependent Layer Redundancy in Transformers cites this paper.

No Free Swap: Protocol-Dependent Layer Redundancy in Transformers SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:33:43.492621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T20:29:38.004776Z digest=sha256:94fd9718a2befeea16e21922777082f990c9fbb5284f2f2f89ec86390e7538d6

Observation f066fd47-9d56-42ae-a50b-9c90e16842ba · inbound

DOT-MoE: Differentiable Optimal Transport for MoEfication cites this paper.

DOT-MoE: Differentiable Optimal Transport for MoEfication SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:16.846434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T15:30:34.142428Z digest=sha256:0cd3d01d4d3abf73d992d7f8f479a27beb68c76d526195769871770b0e56f584