Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:57:33.640328Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2507.12205.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:57:33.640328Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 48a63018-7cf9-4643-95a2-162f8bdf7e07 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage https://docs.nvidia.com/cuda/cublas/index
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e74c6229-dcd0-44da-bda5-5477f16a7be9 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage M., Buluç, A., Williams, S., and Y ang, C.Optimizing sparse matrix- multiple vectors multiplication for nuclear configuration interaction calculations
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c97a557a-fac1-401c-aa9a-0df47fb8909c · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Fast sparse matrix-vector multiplication on gpus for graph applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 490c7c82-b1d0-472b-95b2-eeab2e50fd88 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Efficient sparse matrix-vector multiplication on cuda
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3464eca1-787c-4aea-a21d-876979d64e4f · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage On the relations between ilus and factored approx- imate inverses
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cde669b4-9605-4140-bbfd-e8e2ff075da4 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In SC22: International Conference for High Performance Computing, Networking, Storage and Analysis (2022), IEEE, pp
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 603493ed-e13f-4a82-9779-850e44bf4070 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Scaling algorithms for weighted matching in general graphs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba33c097-54d3-4f54-97a3-a3bb9d37a992 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Spinfer: Leveraging low-level sparsity for efficient large language model inference on gpus
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 02d83cfa-cdc3-4bb5-b99a-fb6a4782ba62 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Sparsegpt: Massive language models can be accurately pruned in one-shot
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4bdbb461-922d-4b48-9ef7-e9429071805c · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Leveraging index compression techniques to optimize the use of co-processors
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f04e6b05-ae1f-4b3c-b175-a542ff16c034 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Sparse GPU kernels for deep learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ced03fe2-8687-4b8a-beb0-a7f9fcee5273 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage ggerganov/llama
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 57e6c840-e4e6-4e2b-a172-e08de5a78e03 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage L., and Daga, M
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b4ff62c-ddbe-4048-ba7d-d7c0d0eab7c5 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Y., Leng, J., Qiu, Y., Guan, Y., W ang, Z., Jia, X., Li, X., Guo, M., and Zhu, Y
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f701053-2b7c-44d3-849d-4b4e92d299a4 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In Proceedings of the 24th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming, PPoPP 2019, Washington, DC, USA, February 16-20, 2019 (2019), ACM, pp
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 241a9bab-12fb-49a2-a192-1c1353043314 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Flashdecoding++: Faster large language model inference with asynchronization, flat gemm optimization, and heuristics
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66d04aa1-90c5-4445-b1c0-22621663ff4f · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In Proceedings of the 25th ACM SIGPLAN symposium on principles and practice of parallel program- ming (2020), pp
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba390c32-a1f8-4c43-b859-fd9fa46e476a · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Maximum bounded 3-dimensional matching is max snp-complete
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2a7b3a5-33ed-44cb-b088-a48c1f71cee7 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Computational complexity of the perfect matching problem in hypergraphs with subcritical density
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c631827-6d0e-499c-a6ff-359e5020e677 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Optimizing sparse matrix-vector multiplication using index and value compression
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80d4e512-f945-4554-ac46-978fa9d1234f · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage ACM Transactions on Architecture and Code Optimization (2024)
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0204ba87-e55d-4302-a816-413f1822313f · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Csr5: An efficient storage format for cross-platform sparse matrix-vector multiplication
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9653f73c-a23b-4de5-9b01-490916a57d2e · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Spp: Sparsity-preserved parameter-efficient fine-tuning for large language models, 2024
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7222110e-eb3a-4c27-96e4-c5d2073cebdb · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Dasp: Specific dense matrix multiply-accumulate units accelerated general sparse matrix-vector multiplication
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74dd5b41-ec91-4b1a-94c8-35bb361fd81f · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Llm-rec: Personalized recommendation via prompting large language models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 147c739a-f584-4fc4-a0ae-2abc12432e22 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Advances in neural information processing systems 36 (2023), 21702–21720
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fef18ad-bb73-4c57-bd46-bfc9e3d79c97 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Adell: An adaptive warp-balancing ell format for efficient sparse matrix-vector multiplication on gpus
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfa0a9c9-41af-44e9-bd41-56fcef23a7cb · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Merge-based parallel sparse matrix-vector multi- plication
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a7886ae0-cc7d-4f5f-928c-98078b721c07 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In GPU Technology Conference (2010), vol
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a13808ca-0101-4bb2-9ddf-06cddf59714b · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In2021 IEEE International Parallel and Distributed Processing Symposium (IPDPS) (2021), IEEE, pp
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e860b4a-708f-43e3-aad1-8ed942bdf3a0 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c498357d-ccaf-46b9-a17f-a4e0c07a3377 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage A Simple and Effective Pruning Approach for Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31b7b88b-7f35-485f-a93d-e0f572a4a012 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In 2011 International conference on parallel processing (2011), IEEE, pp
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ebc2f897-3a61-4a56-b034-62f6b6b6b329 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c648bf5c-9f23-4d22-ba94-2a440a2ef3c5 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af78bfaf-f28a-4812-a87d-d0c97c12f52a · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage W., and Yelick, K
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd9bd963-9ffc-4a45-9cbc-5c49340c05fb · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage PrivateLoRA For Efficient Privacy Preserving LLM
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18427a5f-1fb6-4a2a-b5cf-eacdb605087a · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage M.Register tiling for unstructured sparsity in neural network inference
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc708c7e-9cb0-43fa-8382-8f44deeb7737 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Accelerating sparse matrix computations via data compression
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8036e1d-99f5-4bce-a1b8-5ebb9735bdf1 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae5af087-a1b0-45e5-83f0-b0aa3f0d4416 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03df9fe-b017-417c-b89e-ea99fa139406 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Besa: Pruning large language models with blockwise parameter- efficient sparsity allocation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a3f4dea-9ab7-4c07-9a9d-97e23b7c3b0c · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage A survey on large language model (llm) security and privacy: The good, the bad, and the ugly
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0438e405-d73c-4f6d-a8f8-526e234a7762 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage OPT: Open Pre-trained Transformer Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 839236d8-5228-4f8f-9019-a9802689158d · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Dynamic Sparse No Training: Training-Free Fine-tuning for Sparse LLMs
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4a225ad-ab69-455a-b4a7-18e8b33ccf67 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Acc-spmm: Accelerating general-purpose sparse matrix-matrix multiplication with gpu tensor cores
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e83db51-a3e4-45a8-b62b-8ff6ef5584e5 · outbound
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage A Survey of Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.