Use when deploying custom ML models on-device, converting PyTorch models, compressing models, implementing LLM inferenc…
Use when deploying custom ML models on-device, converting PyTorch models, compressing models, implementing LLM inference, or optimizing CoreML performance. Covers model conversion, compression, stateful models, KV-cache, multi-function models, MLTensor.
CharlesWiltgen
cli
free
Others in the same category, ranked by how often they are opened.