1.6B open-weight LLM by inference-optimization
Qwen3-1.6B-A0.9B is an open-weight text-generation model with roughly 1.6B parameters published by inference-optimization on the Hugging Face Hub. It has 10,988 downloads. Runs with transformers.
inference-optimization
api
free
Others in the same category, ranked by how often they are opened.