Integrating RVV tensor intrinsics into TVM's MetaSchedule autotuner yields AI kernels that are 29-50% faster than hand-written muRISCV-NN and 35-46% faster than compiler autovectorization on tested RVV 1.0 hardware.
banana pi bpi-f3 [Online]
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Tensor Program Optimization for the RISC-V Vector Extension Using Probabilistic Programs
Integrating RVV tensor intrinsics into TVM's MetaSchedule autotuner yields AI kernels that are 29-50% faster than hand-written muRISCV-NN and 35-46% faster than compiler autovectorization on tested RVV 1.0 hardware.