상세 보기
High-performance Sparsity-aware NPU with Reconfigurable Comparator-multiplier Architecture
- Ryu, Sungju;
- Kim, Jae-Joon
Citations
WEB OF SCIENCE
2Citations
SCOPUS
2초록
-Sparsity-aware neural processing units have been studied to exploit computational skipping on less important features in the neural network models. However, neural network layers typically show various matrix densities, so the hardware performance varies depending on the layer characteristics. In this paper, we introduce a reconfigurable comparator-multiplier architecture, so we can dynamically change the number of comparator/multiplier modules. The proposed reconfigurable architecture increases the throughput by 1.06-17.00x compared to the previous sparsity- aware hardware accelerators.
키워드
Neural processing unit; sparse matrix; multiplier; weight pruning; hardware accelerator
- 제목
- High-performance Sparsity-aware NPU with Reconfigurable Comparator-multiplier Architecture
- 저자
- Ryu, Sungju; Kim, Jae-Joon
- 발행일
- 2024-12
- 유형
- Article
- 권
- 24
- 호
- 6
- 페이지
- 572 ~ 577