Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:31:06.702424Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 97 of 97 outbound references and 0 inbound Pith citation observations for arXiv:2510.00192.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:31:06.702424Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
97 of 97 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dc40d292-4cf2-46ee-a940-4023f01d7983 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Binarybert: Pushing the limit of bert quantization
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 829f914f-a6b3-4a26-93fe-7c48cd0be5c6 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Federated fine-tuning of large language models under heterogeneous tasks and client resources.Advances in Neural Information Processing Systems, 37: 14457–14483, 2024
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69e5e4f8-3aa1-43f4-b374-c0775893f143 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Ce-lora: Computation-efficient lora fine-tuning for language models.CoRR, 2025
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b10f8df-b062-44a4-b9db-5e73b59e7c13 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation LoRAShear: Efficient Large Language Model Structured Pruning and Knowledge Recovery
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5aa05aaa-d78d-4e87-abba-d1fb33ab20c8 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Hessian-free optimization for learning deep multidimensional recurrent neural networks.Advances in Neural Information Processing Systems, 28, 2015
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347cdee2-d545-4bb6-abfe-ac372b75b811 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Heterogeneous LoRA for Federated Fine-tuning of On-Device Foundation Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e785c09b-20d2-4a6f-991d-29182f7982a0 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Training Verifiers to Solve Math Word Problems
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb65cf2-809b-434b-8327-d4561dd6a95a · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81998eaa-18f5-4c34-8de5-bfeb10e6d1a8 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Predicting parameters in deep learning.Advances in neural information processing systems, 26, 2013
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8593886f-bb44-4ff9-ae21-7058745dd726 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Coreset-based neural network compression
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4212534-6d08-43e8-ac53-500660fabe92 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Reducing transformer depth on demand with structured dropout
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3e5cb77-57db-4f56-aca9-a09d56084375 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Depgraph: Towards any structural pruning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a553de21-3638-415d-a056-b6fac7f67c54 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Geometric measure theory
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a998bcb7-ccb8-48c2-9103-5237a4d3ae33 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Mixture-of-LoRAs: An Efficient Multitask Tuning for Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8bfe133-019a-433d-af1d-5816d365fa01 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Optimal brain compression: A framework for accurate post-training quantization and pruning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c27ab0c7-20d9-419c-86f9-48c760c5a6bf · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Sparsegpt: Massive language models can be accurately pruned in one-shot
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38c1c143-6cce-43d0-8c2e-8905d447a0a7 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation The Llama 3 Herd of Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60746ec9-aa14-41a6-9c04-27fd519c4acb · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Reweighted Proximal Pruning for Large-Scale Language Representation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a8af465-7787-4c78-8712-2ac545802ffc · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Learning both weights and connections for efficient neural network
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd0a66d6-9342-4826-b44a-9505b07b4879 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Flora: Low-rank adapters are secretly gradient compressors
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b7bbc3-f206-4ec3-8529-8ba3ad957d01 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Second order derivatives for network pruning: Optimal brain surgeon.Advances in neural information processing systems, 5, 1992
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b66a6d9-281c-4180-8916-7e5201e9d6fa · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Optimal brain surgeon and general network pruning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea315db3-c540-4089-97b4-19033bee0107 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lora+ efficient low rank adaptation of large models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db7e1ab-1d6f-487a-ac22-a1175654dbad · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Subspace optimization for large language models with convergence guarantees
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82da0f9e-fa4f-4c04-bde8-d07bdae9ff6c · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Measuring mathematical problem solving with the math dataset
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1f34cec-dee5-4380-8315-9e35d55661e1 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cfcee99-6076-4fd8-9bd8-e294b84fd331 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d104d56-4bae-4022-8f64-9cf8a47b2b6d · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Accelerated sparse neural training: A provable and efficient method to find n: m transposable masks.Advancesin neural information processing systems, 34:21099–21111, 2021
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d7aa9df-d1ad-421d-93af-ead81aab9d49 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation A Rank Stabilization Scaling Factor for Fine-Tuning with LoRA
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce9c9923-2138-4087-adf9-6dd6bf2f73b1 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Some fundamental aspects about lipschitz continuity of neural networks
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea43b0f6-c49e-4693-8e1d-6dfd56a352fa · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation The optimal bert surgeon: Scalable and accurate second-order pruning for large language models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc589ecf-70d9-4690-90ee-b25976b1e18c · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Ziplm: Inference-aware structured pruning of language models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd0e990b-98d0-4b0d-9226-683156e44756 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Sparse fine-tuning for inference acceleration of large language models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa481466-fcc0-46aa-a3de-6c80f060e2d2 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Albert: A lite bert for self-supervised learning of language representations
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd2b4232-c958-4914-99c0-78fd768a553c · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lipschitz constant estimation of neural networks via sparse polynomial optimization
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46224fc7-20ab-4d6d-93e2-2889a5679241 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Optimal brain damage.Advances in neural information processing systems, 2, 1989
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d5b2695-62ab-4310-b536-33e6b58d6be7 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation T\’yr-the-pruner: Unlocking accurate 50% structural pruning for llms via global sparsity distribution optimization.arXiv preprint arXiv:2503.09657, 2025
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f377ed9a-4110-41d9-9aa4-d46ba17cac09 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Pruning filters for efficient convnets
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67fbccff-3515-48f2-a62c-8b650db8c896 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation SepPrune: Structured Pruning for Efficient Deep Speech Separation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02eaee91-d8c1-467b-ba31-d265e0f9b270 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Memory-efficient llm training with online subspace descent
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78ce60b2-284a-4e14-b5cd-b0e312e95d29 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Dynamic adaptation of lora fine-tuning for efficient and task-specific optimization of large language models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cee76c2c-d269-4488-b224-d10321023029 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation GaLore$+$: Boosting Low-Rank Adaptation for LLMs with Cross-Head Projection
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a19f575a-b0a2-4c96-8290-454a52ffdc12 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Dora: Weight-decomposed low-rank adaptation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42116b73-24c9-4af0-889e-1b473d258b69 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Ebert: Efficient bert inference with dynamic structured pruning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e949072-23a0-403c-b972-8d0394646aa3 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Learning efficient convolutional networks through network slimming
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f1e7130-2b04-48c4-9699-0a982d3c7b13 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Decoupled weight decay regularization
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5852614b-c80a-48b6-af93-a3ccaf18e4df · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation LCM-LoRA: A Universal Stable-Diffusion Acceleration Module
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04174519-2574-4dd5-9a98-aa8e09c27a20 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Llm-pruner: On the structural pruning of large language models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed1f517f-41bf-4dc7-8107-87eb88756a67 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Pissa: Principal singular values and singular vectors adaptation of large language models.Advances in Neural Information Processing Systems, 37:121038–121072, 2024
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 594ba95e-f619-4659-879d-f769ab4372e0 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Pointer sentinel mixture models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cecd9cf-dc79-41b9-aa5f-9e68b937cf78 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Accelerating Sparse Deep Neural Networks
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b116ae71-3485-4aba-b171-d45906e821dd · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Pruning convolutional neural networks for resource efficient inference
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 628ae958-cc78-481d-8b24-d10852a51e77 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Importance estimation for neural network pruning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb9fdb9-f12a-4cb4-9ffe-f53d0ad78993 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Sosp: Efficiently capturing global correlations by second-order structured pruning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78eb8584-3cdc-4233-a5dc-9230d16c2916 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Gradient-free structured pruning with unlabeled data
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d9d37ea-f880-4ce3-8101-013c7a10833a · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Meta-kd: A meta knowledge distillation framework for language model compression across domains
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4eaa6c2-9814-4e3c-a220-a47cafac1877 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 151baf2e-ef0a-48ff-8549-056c48550fa5 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Structural pruning via latency-saliency knapsack
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a898b301-55d1-458a-b5c6-098fa3d90597 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Parameter Efficient Reinforcement Learning from Human Feedback
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e895c56-e7f8-4dd0-9ba5-5b4538ecc5ce · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Woodfisher: Efficient second-order approximation for neural network compression
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeaff8f7-a551-48ec-a25b-1818c93bbbf2 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation A simple and effective pruning approach for large language models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b80c776c-8e10-4a80-8266-6ea6aa176c55 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Patient knowledge distillation for bert model compression
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9768038-dcf6-4def-bb72-9a0c0cf783a8 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Contrastive distillation on intermediate representations for language model compression
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 698325b9-6deb-43ea-8a43-ad4849f0bff1 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Improving LoRA in Privacy-preserving Federated Learning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aaf5e46-c21b-4668-9174-283dad6bbc6a · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Mobilebert: a compact task-agnostic bert for resource-limited devices
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f06c780-054b-4572-b047-7384f857302a · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation DarwinLM: Evolutionary Structured Pruning of Large Language Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0216e233-01c7-4c2e-89a7-237199be87e7 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation LLaMA: Open and Efficient Foundation Language Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2a2da0b-79bc-48d7-9642-64b67cd3d4bd · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 921b7e6b-9ba1-4951-89f4-9ebaf443aaa5 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Attention is all you need.Advances in neural information processing systems, 30, 2017
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97b8fa66-7df6-4169-98db-b4ada96729d5 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Glue: A multi-task benchmark and analysis platform for natural language understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f842fe44-9d50-4586-9b46-808aa377678d · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lora-ga: Low-rank adaptation with gradient approximation.Advances in Neural Information Processing Systems, 37:54905–54931, 2024
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b914099e-0017-4294-b5d7-c0e39b07384a · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lora-pro: Are low-rank adapters properly optimized? In The Thirteenth International Conference on Learning Representations, 2024
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1487b643-c4e3-42b9-9ef9-b4087d23ba37 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Assessing the brittleness of safety alignment via pruning and low-rank modifications
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0984fd7f-d6e7-4b3d-8e53-868410bd99d6 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Structured optimal brain pruning for large language models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54288fbb-9f33-44eb-bda4-7ad7a22ba123 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation The iterative optimal brain surgeon: Faster sparse recovery by leveraging second-order information.Advances in Neural Information Processing Systems, 37:139621–139649, 2024
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9aa7571-1b63-427b-9019-1ca5061b3c0b · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Structured pruning learns compact and accurate models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 004f3f9e-bb2a-4b06-bc39-57ed81f83761 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd4e1ad9-f87e-4464-adfd-1e93b5909e6b · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Deebert: Dynamic early exiting for accelerating bert inference
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c20982d-41d1-4529-98da-8cc4468b79c7 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Rethinking network pruning–under the pre-train and fine-tune paradigm
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26c72107-9900-4c67-93a8-b16ccdd102bf · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Theoretical characterization of how neural network pruning affects its generalization.OpenReview, 2023
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 310b74ce-55c8-4d1a-b4c6-717d9304bea0 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Mtl-lora: Low-rank adaptation for multi-task learning
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c78340-18b5-47cd-8366-53db9232320f · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Wanda++: Pruning large language models via regional gradients
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ede60751-dd98-42db-9867-a69701b04380 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Gradient-based intra-attention pruning on pre-trained language models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff35142-5fa4-4932-9f00-d9b61fffd795 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Zeroquant: Efficient and affordable post-training quantization for large-scale transformers.Advances in neural information processing systems, 35:27168–27183, 2022
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5541c67e-4196-4196-b9d8-c229384ac352 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lora done rite: Robust invariant transformation equilibration for lora optimization
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e77a342-d52c-4a35-92c5-13246ca3e522 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Metamath: Bootstrap your own mathematical questions for large language models
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50972c07-37a0-4a48-94c9-1c93fd4060dc · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Understanding the Statistical Accuracy-Communication Trade-off in Personalized Federated Learning with Minimax Guarantees
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00e775b1-f0b2-40cb-aa5c-af2c90e37031 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Altlora: Towards better gradient approximation in low-rank adaptation with alternating projections.arXiv preprint arXiv:2505.12455, 2025
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e17a03dc-21e9-4c78-a211-0c9296c0c804 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Q8bert: Quantized 8bit bert
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6ff9d6a-c44d-496f-80fd-7ddb175e5aec · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Block-diagonal Hessian-free Optimization for Training Neural Networks
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04a1f2bc-20fc-45e9-8af4-1adcf7e87a18 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Loraprune: Struc- tured pruning meets low-rank parameter-efficient fine-tuning
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2d8d4ed-11ae-45ac-8387-58b1465c45bf · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Adap- tive budget allocation for parameter-efficient fine-tuning
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 684e44fd-643e-43a4-9cf3-f192963eacdb · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Lora-one: One-step full gradient could suffice for fine-tuning large language models, provably and efficiently
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58561001-168f-4cb2-968d-059b871e41c0 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Galore: Memory-efficient llm training by gradient low-rank projection
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75ef3e90-57fa-4ee7-bd45-c9a061a9e25f · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Adaptive Activation-based Structured Pruning
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db05630d-57f8-4472-8431-ac7392949e70 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation Each Rank Could be an Expert: Single-Ranked Mixture of Experts LoRA for Multi-Task Learning
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41485d92-3dbb-484b-bc63-31ad17e3a477 · outbound
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation To prune, or not to prune: exploring the efficacy of pruning for model compression
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.