Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:44.749503Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 2 inbound Pith citation observations for arXiv:2505.06302.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:44.749503Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:34:04.652782Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T10:34:05.589341Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2dbcf31a-de3d-4ba5-b5a7-b0b267b60031 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Precision-energy-throughput scal- ing of generic matrix multiplication and discrete convolu- tion kernels via linear projections
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4e71d887-c2bb-4e1a-b429-9e63b9021176 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Nvidia A100 Tensor Core GPU: Performance and innovation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 078d9e27-69e7-47a1-8c9f-9e71dbaae483 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Claude 3.5 Sonnet
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d5ce0f7e-68d7-41b5-8317-d7308e20cd17 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives DeepSeek-V3 Technical Report
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1cf2130d-582f-4c2e-b354-8f2a315c786a · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Flexible Performant GEMM Kernels on GPUs
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a247c96e-2745-47d1-8462-b17bfa7e353b · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Tensorir: An abstraction for automatic tensorized program optimization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4a99e110-7f9a-4caa-a4d3-0949bf5eff74 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives The Llama 3 Herd of Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7e55381-991b-4d36-ab0e-29f6d127aa54 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Automatic generation of ARM NEON micro-kernels for matrix multiplication
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7c0a71b1-da33-4b23-81cf-989ba57641fb · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Textbooks Are All You Need
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85936515-9da2-41be-b5ed-ea8fd1b681b7 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Deep Residual Learning for Image Recognition
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29677837-0684-499d-9317-d2ad4bce9a96 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives In-datacenter performance analysis of a tensor processing unit
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ca501492-ed13-4231-bce1-311566df0905 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Exploit- ing Intel® Advanced Matrix Extensions (AMX) for Large Language Model Inference
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e0b4b441-d88a-4fd3-871c-a474488fa2bc · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives oneAPI Open-Source Math Library Interface
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 76f4a096-07b3-4d0c-b5f1-a58cb1dfe12b · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Autotuning GEMM kernels for the Fermi GPU
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a1e57917-efd3-4c51-ba74-71f5605e314c · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Cambricon: An instruction set architecture for neural networks
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d368b010-54b1-46ec-8a82-81c184a557b0 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Introducing Llama 3.1: Our most capable models to date
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 99270a18-0545-47af-b6e1-17c6329ee676 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives ReACC: A Retrieval-Augmented Code Completion Framework
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd84b7c5-0477-4dff-b655-b956a10d0f5a · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives NVIDIA Tensor Core Programmabil- ity, Performance & Precision
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1c61b06a-ecd3-4ddc-8407-362ffae22470 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives CUBLAS LIBRARY user guide v12.1
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e7b4d3c8-5337-410b-9b69-4d600cb9fdce · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives [OpenAI, 2025] OpenAI
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1af83234-57db-430b-a16b-4b20c374cb66 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives U-Net: Convolutional Networks for Biomedical Image Segmentation,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 937666da-9beb-4962-a93d-e22800cfe1d7 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Code Llama: Open Foundation Models for Code
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2568c05e-bf8b-44ce-bcd7-4b9c36560a4e · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Tensor program optimization with probabilistic pro- grams
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 23b197ec-800c-431b-a284-d9f5b7faaf02 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Very Deep Convolutional Networks for Large-Scale Image Recognition
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e09e2411-15b7-4890-a416-1e4299dbeb77 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Efficient processing of deep neural net- works: A tutorial and survey
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1b6e8658-6d68-4bb5-80ec-4e9314207a1e · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Fast implementation of DGEMM on Fermi GPU
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a483d327-5d02-4316-9148-de81756764a2 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Chain-of-Thought Prompting Elicits Rea- soning in Large Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5f5970df-732d-4d7a-98f6-6a4d21d991c1 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives The storage hierarchy is not a hierarchy: Optimiz- ing caching on modern storage devices with orthus
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 27d3c2ea-8e12-40ee-9f8f-eefc3f201ec9 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives autoGEMM: Pushing the Limits of Irregular Matrix Mul- tiplication on Arm Architectures
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fa4f498f-db61-40c0-aa70-d03617cc2235 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Model-driven level 3 BLAS performance opti- mization on Loongson 3A processor
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b6b38813-4ce7-40e9-9a1d-0e0254b8844a · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Hasco: Towards agile hardware and software co-design for tensor computation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e886e66c-4928-46d3-929b-521a4bd46c78 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Large Language Models Meet NL2Code: A Survey
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 78f2a07b-c858-444b-9282-d1b6c39ef2f6 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives TLP: A Deep Learning-Based Cost Model for Tensor Pro- gram Tuning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e29177c9-be1c-4973-acf1-47480008f44c · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Enabling Tensor Language Model to Assist in Generating High-Performance Tensor Programs for Deep Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2c2def90-37a9-4ab0-a763-ef28f46024eb · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Ansor: Generating High-Performance Ten- sor Programs for Deep Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cc68d893-7766-472e-8c09-670f0556a0a3 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives LLaMA: Open and Efficient Foundation Language Models
Reference 2011
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f6e40fc-b231-4712-80dd-6782e27a6f62 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives StarCoder: may the source be with you!
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2d3ac28-05d7-4368-b368-5a434d8bbbcc · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Multi-lingual Evaluation of Code Generation Models
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e165f11c-31e6-4ed3-987f-e05cfd57aaaf · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives A new golden age for computer architecture
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f5fc90e5-fb11-4be8-a73f-b4940708d20f · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives High-Performance Tensor Learning Primitives Using GPU Tensor Cores
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d665f4be-373a-4e91-8cf6-48bc84dbfc97 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives High Performance GPU Code Generation for Matrix-Matrix Multiplication using MLIR: Some Early Results
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5e535f0d-b2f0-418a-be6e-867b4aa58450 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Xuantie-910: A Commercial Multi- Core 12-Stage Pipeline Out-of-Order 64-bit High Per- formance RISC-V Processor with Vector Extension
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8337a8e4-6be7-4847-96be-b5661da30e86 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Automatic Generation of Micro-kernels for Performance Portability of Matrix Multiplication on RISC-V Vector Processors
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1baca3a5-41db-4662-ac39-62fef05596e0 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Experiments and optimizations for TVM on RISC-V Architectures with P Extension
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 89cbbf1b-26d6-4480-8df6-fd28a40b36f3 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Heron: Automatically Constrained High-Performance Li- brary Generation for Deep Learning Accelerators
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0f3d3397-4ea9-4b82-b7f4-e32807f6cc87 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Program Synthesis with Large Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a22d3f3-a42a-4f90-bb3e-a7ffa02d6020 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives Tackling the Matrix Multiplication Micro-Kernel Generation with Exo
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 75920c96-a955-4875-a855-39e35cad169b · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives [Dally et al., 2021] William J Dally, Stephen W Keckler, and David B Kirk
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0bb652b8-dc32-4cab-8794-ff74f2b22ab6 · outbound
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives [Ragan et al., 2013] Jonathan Ragan, Connelly Barnes, and Andrew Adams et al
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0613ad5c-6843-4a3b-9358-e656ec082d4e · inbound
QiMeng: Fully Automated Hardware and Software Design for Processor Chip QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3be5b7e5-042b-4262-8c59-928d32faf87c · inbound
Towards Automated Kernel Generation in the Era of LLMs QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.