Sparse Matrix-Matrix Multiplication (SpMM) is a fundamental operation in graph computing and analytics. However, the irregularity of real-world graphs poses significant challenges to achieving ...
Abstract: Fine-grained structured sparsity, especially the 2:4 pattern, is a practical way to reduce the computation and memory cost of deep learning workloads while preserving model accuracy. NVIDIA ...