A Novel Mixed-Precision Quantization Approach for CNNs
Model size and inference speed have brought about major challenges for the deployment of Convolutional Neural Networks (CNNs) in many applications. An effective approach to address this issue is model quantization, which achieves network compression and inference speedup by reducing the parameters b...
Zapisane w:
| Główni autorzy: | , , , |
|---|---|
| Format: | Artigo |
| Język: | Inglês |
| Wydane: |
IEEE
2025-01-01
|
| Seria: | IEEE Access |
| Hasła przedmiotowe: | |
| Dostęp online: | https://ieeexplore.ieee.org/document/10929039/ |
| Etykiety: |
Nie ma etykietki, Dołącz pierwszą etykiete!
|
