# Accelerating Sparse CNN Inference on GPUs with Performance-Aware Weight Pruning

> Research article (Proceedings of the ACM International Conference on Parallel Architectures and Compilation Techniques, 2020) · cited 26× · AI/ML

**Wikidata**: [openalex:W3091174280](https://www.wikidata.org/wiki/openalex:W3091174280)  
**Source**: https://4ort.xyz/entity/accelerating-sparse-cnn-inference-on-gpus-with-performance-aware-weight-pruning
