AutoCompress automatically prunes DNNs to ultra-high compression rates without accuracy loss.
problem Efficiently compressing deep neural networks to reduce storage and computation requirements.
method Automatic hyperparameter determination, ADMM-based structured weight pruning, purification step, heuristic search.
result Achieves ultra-high pruning rates on weights and FLOPs, up to 33x in pruning rate.