Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:56:15.620948Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2411.09242.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:56:15.620948Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8c46a14b-1056-44cc-81f3-b66999537102 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67ddb99e-1cbf-4df8-9ae1-2784b73a2ecf · outbound
FluidML: Fast and Memory Efficient Inference Optimization GPT-NeoX: Large Scale Autoregressive Language Modeling in PyTorch , 9 2023
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4a371ee6-a733-4e5e-b3d7-2a65a34d81a2 · outbound
FluidML: Fast and Memory Efficient Inference Optimization High Performance Code Generation in MLIR: An Early Case Study with GEMM
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d28679-13a3-4408-b167-bf29d096d115 · outbound
FluidML: Fast and Memory Efficient Inference Optimization The slab allocator: an object-caching kernel memory allocator
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ba640ada-42f8-4be8-bdcd-296794f8006e · outbound
FluidML: Fast and Memory Efficient Inference Optimization J., Leary, C., Maclaurin, D., Necula, G., Paszke, A., Vander P las, J., Wanderman- M ilne, S., and Zhang, Q
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35733c97-50cf-4b60-b5e6-f7310f69d9d3 · outbound
FluidML: Fast and Memory Efficient Inference Optimization TVM: An Automated End-to-End Optimizing Compiler for Deep Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2cac0c7-83e4-447c-9e2a-aa4d963e85af · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b706b7-3ec3-4e60-80e3-3799153b1ffa · outbound
FluidML: Fast and Memory Efficient Inference Optimization Intel(r) math kernel library for deep neural networks (intel(r) mkl-dnn)
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8ff634b7-3931-496e-b56f-49dffeac24c3 · outbound
FluidML: Fast and Memory Efficient Inference Optimization TensorFlow Lite Micro: Embedded Machine Learning on TinyML Systems
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e0dda72-379b-451d-9e03-21bb2eb33b8b · outbound
FluidML: Fast and Memory Efficient Inference Optimization Tensorflow lite micro: Embedded machine learning for tinyml systems
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6d102845-b61d-4a27-a197-b8df7e7fb21f · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7bb98dc0-72c8-4adf-a419-8b276598b219 · outbound
FluidML: Fast and Memory Efficient Inference Optimization BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7921ea0-f8ae-4312-a12f-90cf00bbc586 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Algorithms for compile-time memory optimization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8a9c8a04-cd6f-4927-a43d-cb41d951dbcb · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ab60ea0e-a3ba-4910-a915-cdc994f06cc6 · outbound
FluidML: Fast and Memory Efficient Inference Optimization and Geijn, R
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6efccda-0494-45be-95ab-ca700345e725 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e013fa76-f195-47ba-8175-c2b45a6a39ca · outbound
FluidML: Fast and Memory Efficient Inference Optimization Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 330c517a-6c1b-4ce8-8a1b-4e1dfa0b060e · outbound
FluidML: Fast and Memory Efficient Inference Optimization Distilling the Knowledge in a Neural Network
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec94146e-6de9-4959-97f2-88dbe04eb24c · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 47a0e427-e4dc-4539-9b3a-4b2a3a4a5a33 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Openvino
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 54873ebe-6de1-4abc-afe2-11d86425316a · outbound
FluidML: Fast and Memory Efficient Inference Optimization ConvBERT: Improving BERT with Span-based Dynamic Convolution
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3266c7c-d358-4813-83ac-b72e9fd23524 · outbound
FluidML: Fast and Memory Efficient Inference Optimization FBGEMM: Enabling High-Performance Low-Precision Deep Learning Inference
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0647653-e7a9-4bb6-aef9-ce4697c75ab2 · outbound
FluidML: Fast and Memory Efficient Inference Optimization W., and Keutzer, K
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0204436d-8b31-4502-83f9-59772d9e8992 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 29196095-6593-4a14-8ddd-c66b900bb85b · outbound
FluidML: Fast and Memory Efficient Inference Optimization CMSIS-NN: Efficient Neural Network Kernels for Arm Cortex-M CPUs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75fe1ef6-a324-49c9-8f69-17b1186b8030 · outbound
FluidML: Fast and Memory Efficient Inference Optimization MLIR: A Compiler Infrastructure for the End of Moore's Law
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e905d47-f3a4-4035-837d-b946e51f15c1 · outbound
FluidML: Fast and Memory Efficient Inference Optimization MLIR : Scaling compiler infrastructure for domain specific computation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 867c6d68-25de-4b0d-9dcd-11f0f7f0c093 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Compiling ONNX Neural Network Models Using MLIR
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a0f5617-e628-427f-8c59-a7f86c1579a8 · outbound
FluidML: Fast and Memory Efficient Inference Optimization On-Device Neural Net Inference with Mobile GPUs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd8231fb-ed7e-4a60-917d-9c73b56192eb · outbound
FluidML: Fast and Memory Efficient Inference Optimization MCUNetV2: Memory-Efficient Patch-based Inference for Tiny Deep Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc4f2966-590c-4e7e-aaf0-b1f1d064f4ad · outbound
FluidML: Fast and Memory Efficient Inference Optimization and Deng, W
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f86e7b1-706e-4333-960f-05d662e9c11b · outbound
FluidML: Fast and Memory Efficient Inference Optimization Optimizing CNN model inference on CPUs
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4b5caed0-8a29-4675-bb2f-cfc91d4af710 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Rammer: Enabling holistic deep learning compiler optimizations with rTasks
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa0c3e2f-d18e-493c-ba61-766a99ef7dae · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 11a49a10-119a-4226-bb22-260ccd0aad17 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a3969b7f-7fca-457e-95a5-da03fa5cc8e7 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Tensorrt
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3a37fea0-d617-4a59-9ac2-5226e0d6a1d4 · outbound
FluidML: Fast and Memory Efficient Inference Optimization PyTorch: An Imperative Style, High-Performance Deep Learning Library
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f52dbcd3-e8e0-4f12-a81e-859755eea8d2 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Efficient Memory Management for Deep Neural Net Inference
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b8ccd284-3882-4be1-899b-067239ed600e · outbound
FluidML: Fast and Memory Efficient Inference Optimization Halide: a language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61996e79-8239-4cac-94e3-19e8a18c9c0d · outbound
FluidML: Fast and Memory Efficient Inference Optimization Glow: Graph Lowering Compiler Techniques for Neural Networks
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af7e3904-88dc-4e36-9328-81398812238f · outbound
FluidML: Fast and Memory Efficient Inference Optimization Xla : Compiling machine learning for peak performance, 2020
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c954bb67-6291-4847-b3c3-0945b98bf92e · outbound
FluidML: Fast and Memory Efficient Inference Optimization Efficient Transformers: A Survey
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a777bded-2662-4b06-8289-855dd326dab3 · outbound
FluidML: Fast and Memory Efficient Inference Optimization IREE , September 2019
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 363f2bf0-1ae2-457f-82bf-79e47cab334d · outbound
FluidML: Fast and Memory Efficient Inference Optimization Attention Is All You Need
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7855c6d-940a-4e13-a323-193a0f8ade81 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Augem: Automatically generate high performance dense linear algebra kernels on x86 cpus
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51ffc36d-2cdd-4f0f-ac66-1aa7e8a9dca7 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 67f7592a-28c4-406a-909f-8afcd4bb9f33 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Model-driven level 3 blas performance optimization on loongson 3a processor
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a66df72-f32a-41b1-ba77-4ab3357bfea5 · outbound
FluidML: Fast and Memory Efficient Inference Optimization DeepCPU : Serving RNN-based deep learning models 10x faster
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9656f4ed-f265-46c7-9bdf-6434388e5888 · outbound
FluidML: Fast and Memory Efficient Inference Optimization H., Haj-Ali, A., Wang, Y., Yang, J., Zhuo, D., Sen, K., Gonzalez, J
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0f583c32-1c31-45a0-8511-a2fd3c574681 · outbound
FluidML: Fast and Memory Efficient Inference Optimization vMCU: Coordinated Memory Management and Kernel Optimization for DNN Inference on MCUs
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1ee2c716-5ce3-4f87-8c01-a730e8241a36 · outbound
FluidML: Fast and Memory Efficient Inference Optimization Neural Architecture Search with Reinforcement Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4ac0616-10e2-4c9c-9291-206640762832 · outbound
FluidML: Fast and Memory Efficient Inference Optimization write newline
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.