neural-compressor. Intel® Neural Compressor (formerly known as Intel® Low Precision Optimization Tool), targeting to provide unified APIs for network compression technologies, such as low precision quantization, sparsity, pruning, knowledge distillation, across different deep learning frameworks to pursue optimal inference performance.

github.com/patrickvonplaten/neural-compressor

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.