Machine learningML & AI Methods
Weight Block Sparsity: Training, Compilation, and AI Engine Accelerators
The paper proposes a system that applies weight block sparsity in Deep Neural Networks, halving the weight with minimal accuracy loss and doubling the speed of inference.
Featured in No. 57 on 17 Jul 2024 · 5 days after release · 2 citations today
- Released
- 12 Jul 2024
- First featured
- No. 57 · 17 Jul 2024
- Citations (Semantic Scholar)
- 2
- Influential citations
- 0
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 10
- Identifier
- arXiv:2407.09453
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).