Mã QR

Learning Low-Precision Structured Subnetworks Using Joint Layerwise Channel Pruning and Uniform Quantization

Pruning and quantization are core techniques used to reduce the inference costs of deep neural networks. Among the state-of-the-art pruning techniques, magnitude-based pruning algorithms have demonstrated consistent success in the reduction of both weight and feature map complexity. However, we find...

Mô tả đầy đủ

Đã lưu trong:
Chi tiết về thư mục
Những tác giả chính: Xinyu Zhang, Ian Colbert, Srinjoy Das
Định dạng: Artigo
Ngôn ngữ:Inglês
Được phát hành: MDPI AG 2022-08-01
Loạt:Applied Sciences
Những chủ đề:
Truy cập trực tuyến:https://www.mdpi.com/2076-3417/12/15/7829
Các nhãn: Thêm thẻ
Không có thẻ, Là người đầu tiên thẻ bản ghi này!