Hardware-Aware DNN Compression via Diverse Pruning and Mixed-Precision Quantization

Open in new window