Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2403.17887.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T14:29:13.847195Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 15924654-a6e9-4bb0-a154-96407715c746 · inbound
MixLLM: LLM Quantization with Global Mixed-precision between Output-features and Highly-efficient System Design The Unreasonable Ineffectiveness of the Deeper Layers
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8278ec27-444a-4a53-b670-5bf17e096aa5 · inbound
Exploiting Sparsity for Long Context Inference: Million Token Contexts on Commodity GPUs The Unreasonable Ineffectiveness of the Deeper Layers
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0879de0-20b5-4512-9560-2606bf0d14a3 · inbound
DarwinLM: Evolutionary Structured Pruning of Large Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2a0077f-9906-4d40-a7c2-332cf89e403b · inbound
SHARP: Accelerating Language Model Inference by SHaring Adjacent layers with Recovery Parameters The Unreasonable Ineffectiveness of the Deeper Layers
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed77fdf-3f4e-48ae-b846-e454c23d7916 · inbound
MUDDFormer: Breaking Residual Bottlenecks in Transformers via Multiway Dynamic Dense Connections The Unreasonable Ineffectiveness of the Deeper Layers
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d7c48b-a93a-4468-b8f5-be1252e80b84 · inbound
Void in Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65109521-ad22-4ee9-9643-971dc58d939b · inbound
Leveraging Stochastic Depth Training for Adaptive Inference The Unreasonable Ineffectiveness of the Deeper Layers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4e4d684-fa8f-416f-b29c-851b1e4d532f · inbound
SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling The Unreasonable Ineffectiveness of the Deeper Layers
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2b52689-eac5-4880-893e-9e4446e96386 · inbound
GPTailor: Large Language Model Pruning Through Layer Cutting and Stitching The Unreasonable Ineffectiveness of the Deeper Layers
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b716bc9a-db6d-4c36-bcd7-c917ef91e4cc · inbound
GPAS: Accelerating Convergence of LLM Pretraining via Gradient-Preserving Activation Scaling The Unreasonable Ineffectiveness of the Deeper Layers
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0a084e5-e6ef-4663-85d2-3565f95c0be8 · inbound
Towards Distributed Neural Architectures The Unreasonable Ineffectiveness of the Deeper Layers
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39e6cb2-95ed-417a-99fd-7ee09908d0f3 · inbound
A Survey on Latent Reasoning The Unreasonable Ineffectiveness of the Deeper Layers
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4779f547-3305-4fa8-922a-a5d87963794a · inbound
PUMA: Layer-Pruned Language Model for Efficient Unified Multimodal Retrieval with Modality-Adaptive Learning The Unreasonable Ineffectiveness of the Deeper Layers
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcab9834-a370-4cff-9f4f-0d1f01636d0d · inbound
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers The Unreasonable Ineffectiveness of the Deeper Layers
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 428b471c-e54d-40c7-bb57-69bac7a49e01 · inbound
Amber Pruner: Leveraging N:M Activation Sparsity for Efficient Prefill in Large Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df68b59e-b75e-416b-9d0d-c110f7247e5c · inbound
On the Limits of Layer Pruning for Generative Reasoning in Large Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e0df9383-f5c5-4323-bde1-0dd286d3095f · inbound
Inverse Depth Scaling From Most Layers Being Similar The Unreasonable Ineffectiveness of the Deeper Layers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff7b1b89-f86b-40df-af12-51f254be6010 · inbound
Attention Residuals The Unreasonable Ineffectiveness of the Deeper Layers
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 05093803-5c9d-4837-b839-8b3b9182506b · inbound
When Does Sparsity Mitigate the Curse of Depth in LLMs The Unreasonable Ineffectiveness of the Deeper Layers
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d33c3ed4-d16b-4b5c-9096-fdb75cea654a · inbound
Do Transformers Use their Depth Adaptively? Evidence from a Relational Reasoning Task The Unreasonable Ineffectiveness of the Deeper Layers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e8104843-579d-42a9-9d15-eb0b8db47ea3 · inbound
LASER: Low-Rank Activation SVD for Efficient Recursion The Unreasonable Ineffectiveness of the Deeper Layers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36982239-45d0-44d2-8e0a-541c046845e1 · inbound
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales The Unreasonable Ineffectiveness of the Deeper Layers
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da4b0a3c-dd7c-472f-8b1e-611296fcf2c8 · inbound
Why Geometric Continuity Emerges in Deep Neural Networks: Residual Connections and Rotational Symmetry Breaking The Unreasonable Ineffectiveness of the Deeper Layers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 968f2024-b415-4611-830c-e3e2d21054b6 · inbound
Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions The Unreasonable Ineffectiveness of the Deeper Layers
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9edb290a-8f87-4e97-beaf-00ab32135e20 · inbound
A Hamiltonian-Inspired Local-Operator Ansatz for Slimming Large Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f089641-195d-4c4e-863e-279db3fc46d7 · inbound
A Hamiltonian-Inspired Local-Operator Ansatz for Slimming Large Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b9f4ddb-81d9-427a-9268-c1049d67dde0 · inbound
Dense2MoE: Pushing the Pareto Frontier of On-Device LLMs via Unified Pruning and Upcycling The Unreasonable Ineffectiveness of the Deeper Layers
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2c2190d2-c753-4317-80b6-fb56bc47b5c2 · inbound
Complementary Attention Head Pruning for Efficient Transformers The Unreasonable Ineffectiveness of the Deeper Layers
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 906b38bd-28c8-4533-8690-63b81b2d6f74 · inbound
Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think The Unreasonable Ineffectiveness of the Deeper Layers
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 038e7b6e-7223-4dda-bee5-2689f8499486 · inbound
Tapered Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb2b8ee1-e747-4e0f-9231-b0b9414c71b2 · inbound
Neural Scaling Universality: If Exponents Are Fixed, Time to Understand Coefficients The Unreasonable Ineffectiveness of the Deeper Layers
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 75daf596-2a9d-485b-bb1d-3b0c5e5ed129 · inbound
CascadeFormer: Depth-Tapered Transformers Motivated by Gradient Fan-in Asymmetry The Unreasonable Ineffectiveness of the Deeper Layers
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 089336c5-735d-4e29-a588-97efa8dd67c0 · inbound
Gradient Smoothing: Coupling Layer-wise Updates for Improved Optimization The Unreasonable Ineffectiveness of the Deeper Layers
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation afafa4bd-10a8-460a-8c10-0b46a478daa9 · inbound
CHERRY: Compressed Hierarchical Experts with Recurrent Representational Yield The Unreasonable Ineffectiveness of the Deeper Layers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1b3ff6bd-0d27-4ead-bada-7607901cac3b · inbound
CHERRY: Compressed Hierarchical Experts with Recurrent Representational Yield The Unreasonable Ineffectiveness of the Deeper Layers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 448ee395-7f79-4a6c-a86d-f39707d2bc37 · inbound
Rethinking Depth Pruning for Vision Transformers: A Heterogeneity-Aware Perspective The Unreasonable Ineffectiveness of the Deeper Layers
Reference 128
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db6df2d3-4bbb-45c9-9a50-b67f2d74f697 · inbound
Wrong Before Right: Late Rescue and Interface Failure in Aligned Language Models The Unreasonable Ineffectiveness of the Deeper Layers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a05f125-ea59-4a1e-9499-0f213ae44cec · inbound
Latent Communication Between Language Model Agents: Channels, Alignment, and the Limits of Text The Unreasonable Ineffectiveness of the Deeper Layers
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2735368-4d99-448e-977c-6c8e1e2645b6 · inbound
Bekko Embedding: Parameter-Efficient Multilingual Retrieval with Ultra-Compact Encoders The Unreasonable Ineffectiveness of the Deeper Layers
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.