Techniques for reducing model size and improving inference speed through quantization, pruning, and optimization
Techniques for reducing model size and improving inference speed through quantization, pruning, and optimization
amnadtaowsoam
cli
free
Others in the same category, ranked by how often they are opened.