4B open-weight LLM by rapid-mlx
Qwen3.8-Flash-Next-4bit is an open-weight text-generation model with roughly 4B parameters published by rapid-mlx on the Hugging Face Hub. It has 1,307 downloads. Runs with mlx.
rapid-mlx
api
free
Others in the same category, ranked by how often they are opened.