ML-QuantSubscribe

Machine learningML & AI Methods

Weight Block Sparsity: Training, Compilation, and AI Engine Accelerators

The paper proposes a system that applies weight block sparsity in Deep Neural Networks, halving the weight with minimal accuracy loss and doubling the speed of inference.

Featured in No. 57 on 17 Jul 2024 · 5 days after release · 2 citations today

Released
12 Jul 2024
First featured
No. 57 · 17 Jul 2024
Citations (Semantic Scholar)
2
Influential citations
0
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
10
Identifier
arXiv:2407.09453

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page