Efficient Qwen model for fast chat, extraction, and high-volume workloads
Efficient Qwen model for fast chat, extraction, and high-volume workloads. Supports a context window of 1,000,000 tokens. Supports extended reasoning. Supports tool calling. Input priced at $0.05 per million tokens. Weights are not publicly released; access is via the provider API.
Alibaba
api
paid
No benchmark results have been added yet.