Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T19:29:57.522816Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2501.10054.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T19:29:57.522816Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
70 of 70 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 47ae6c4c-283a-4f1e-9acf-e30db799f366 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network org/wiki/Constant_folding
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 55d8c0b9-e687-4110-8705-e7aae8c255d7 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2024- 11-01]
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 56ef685d-66dd-4872-95bb-ee4cf618b79d · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; ac- cessed 2025-01-09]
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 686e192f-7b29-4b0b-a239-ada71ce30b8a · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network https://gcc.gnu.org/onlinedocs/gcc/ Optimize-Options.html
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f4881877-eaa7-4200-bd99-cd7d376b614b · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network https://github.com/python/cpython?tab= readme-ov-file
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d15c553d-6a0f-44a9-a9a5-443a57333b4e · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2024-11-24]
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 258f282e-e002-4763-8c7e-f22858ebea94 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2025-01-14]
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b697275b-1576-48bc-b4a7-87f5c83864ac · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2024-12-16]
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b3ca9336-7a7b-4891-b59e-1e533fe61770 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network apache.org/sql/
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f11c88ed-e790-4a31-99d8-9faa3c59d64a · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network https://huggingface.co/docs/ transformers/index
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bbbec6b4-63bf-4896-afab-21dc14a73e46 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network https://docs.rapids.ai/api/ cuml/stable/
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 010db526-213b-4a13-863c-b8a8ac4d08ba · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2025-01-09]
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1c70fca2-9dc9-4be8-bea8-015e1d594fab · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2024-11-01]
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6cd4a652-4e7e-4b9d-ad5d-9e01c31131e7 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4b978dbe-baaa-4f16-bbf0-1369b183af35 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; ac- cessed 2025-01-13]
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bff2270c-2764-47fa-b352-87e0f84d5b2e · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [On- line; accessed 2025-01-09]
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 88f6008b-4556-4b6c-9501-13585b7a9590 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2025-01-09]
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f68de9b5-8556-493f-bb2a-c550855ce721 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network [Online; accessed 2025-01-08]
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b4b13c58-2e90-40d2-a3d1-7788b69cd37f · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network The falcon series of open language models, 2023
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b27dfc9a-5afb-407a-8e45-e09b9b3961f3 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Piqa: Reasoning about physical commonsense in natural language
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 017692bf-2f49-48e6-956a-d395eeab2110 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Language Models are Few-Shot Learners
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fe328c4-ecac-403e-9e30-8f3396adca7f · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network DSFormer: Effective Compression of Text-Transformers by Dense-Sparse Weight Factorization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd6b491b-2a39-4729-8837-64d16d867ec8 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs)
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fee8dd64-3654-468a-a685-0ded02614e57 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Language modeling with gated convolutional networks
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fdb5e811-bcee-4fb0-9bc1-7163780e91f5 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Llm.int8(): 8-bit matrix multiplication for transformers at scale
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 154393a1-edb9-481e-9196-f3b62b4bdda5 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network The Llama 3 Herd of Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc653491-8c84-4b7e-9393-42bd93d952fa · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Sigmoid- weighted linear units for neural network function ap- proximation in reinforcement learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5c263d24-44cb-4662-9dfb-f2adc114ba08 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Optq: Accurate quantization for generative pre-trained transformers
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 71bedc88-247f-47da-ba68-465f7b400241 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Gptq: Accurate post-training quantization for generative pre-trained transformers, 2023
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 397f3daa-a141-4857-8a3c-a925630621b3 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network A framework for few-shot language model evaluation, 07 2024
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 432917f4-daa3-4f17-ab2b-76eaafc1b8e9 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network MiniLLM: On-Policy Distillation of Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90c98b53-7fd3-4a8e-8af1-59f0c6f11aee · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network How to access global memory efficiently in cuda c/c++ kernels | nvidia technical blog, 4 2014
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f750f5dc-3400-42b7-8dcc-2652d26c3692 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Gaussian Error Linear Units (GELUs)
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e47903f-e0b9-494a-8a65-9aed501f3e0a · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network FineQuant: Unlocking Efficiency with Fine-Grained Weight-Only Quantization for LLMs
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7d81d1e-463c-4a16-921d-6908ca8824ed · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Self-normalizing neural networks
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 33d35ab0-8ac1-4c4d-8c6e-3994e1105e14 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Pruning vs quanti- zation: Which is better? Advances in neural information processing systems, 36:62414–62427, 2023
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e1665b7b-dcf2-4388-bee8-447535897847 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Efficient memory man- agement for large language model serving with page- dattention
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 46eac479-8591-4048-99c4-f9ce8f111840 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Bloom: A 176b-parameter open-access multilingual language model
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6f325b37-05c8-47d8-9ace-ba2c8e724dd9 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Losparse: Structured compression of large language models based on low-rank and sparse approximation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2abd438c-11be-4ea1-8af6-a31b7b742fe9 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network A method for calculating the deriva- tive of activation functions based on piecewise linear approximation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 78b89c51-3c3d-4ab5-97e9-b9f7f859bc78 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Awq: Activation- aware weight quantization for on-device llm compres- sion and acceleration
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c0b0bb5e-2b4f-4781-8d1a-e1987cdb7b8b · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Deja vu: Con- textual sparsity for efficient llms at inference time
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e129649-d204-4878-9880-b13ae5e80f01 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Llm- pruner: On the structural pruning of large language mod- els
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 18ebb02b-b103-4fac-93df-5ede178bdc9d · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network The penn treebank: an- notating predicate argument structure
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ff2b605b-e130-4127-a010-9f2274bc6588 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Pointer Sentinel Mixture Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d11ea08-45f2-4d27-8897-e34db63e4800 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Relu strikes back: Exploiting activation sparsity in large language models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e100caf0-bd02-4053-a847-c090aa904677 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Implementation of a digital neuron with nonlinear activation function using piecewise linear approximation technique
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 46d902b8-e5af-4a9b-914c-f5884ea100a3 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network On estimation of a probability den- sity function and mode
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5f6f5a42-0057-40da-8706-9c5f87352fcc · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fc26d064-0da5-4483-9ead-a1bae1cf91bb · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Ma- trix compression via randomized low rank and low pre- cision factorization
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d986af69-3217-4710-9389-ec1baf7816b8 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Glu variants improve transformer, 2020
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20f0c52a-d450-469e-ab28-166c9950f701 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Powerinfer: Fast large language model serving with a consumer-grade gpu
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 71785d44-2d13-4158-8154-fe9dea67d553 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network A simple and effective pruning approach for large lan- guage models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 87730f5e-6e14-4f50-8332-7167b5ca0885 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Variable kernel density estimation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9e34aa2c-5175-4468-b5a0-c474dde97254 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Baby Llama: knowledge distillation from an ensemble of teachers trained on a small dataset with no performance penalty
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9cb247a-04b8-46ed-907b-da86dcfe1aac · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Norm (mathemat- ics) - wikipedia, 9 2004
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5d5e99fd-5c1c-4bd4-a127-7de77fb525ce · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Llama: Open and efficient foun- dation language models, 2023
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b0d9b4fa-f029-4ca5-ba28-14f4c28ee7d3 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Model Compression and Efficient Inference for Large Language Models: A Survey
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c208a281-ae88-4a78-8607-fa24e89187bb · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 196022ff-cf7d-4578-ba83-6afb1afda3c3 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network A Survey of Resource-efficient LLM and Multimodal Foundation Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d809c6b8-c51d-4e3d-a727-71dfc47701ed · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Qwen2.5 Technical Report
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1f05672-b123-4df4-a139-3846fb852591 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Llmcbench: Benchmarking large language model compression for efficient deployment
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 880d2a4f-b811-4098-81fe-7d2af96ea0c1 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Zeroquant: Efficient and affordable post-training quantization for large-scale transformers
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b61e8c3d-98b9-45e5-b58f-7bab992c0e90 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Outlier weighed layerwise sparsity (OWL): A missing secret sauce for pruning LLMs to high sparsity
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a7f3fe86-1afd-43a3-9585-2e055337003d · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae8c5ab6-16aa-4c53-8d15-ab52856f56d3 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network A novel sigmoid function approx- imation suitable for neural networks on fpga
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e6207a2a-4e7b-49e9-a5c3-422b04599a81 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Inves- tigating layer importance in large language models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c3fd90f2-2ebd-4c22-aec3-8dbf4e0f84e1 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Plug-and-play: An efficient post-training pruning method for large lan- guage models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 16628c73-4829-48be-bc05-9294ae5df128 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network A Survey on Efficient Inference for Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f360ec6-00c8-4083-9147-66736c8aa1f8 · outbound
Accelerating Large Language Models through Partially Linear Feed-Forward Network Unresolved cited work
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.