Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks. Supports a context window of 1,000,000 tokens. Supports extended reasoning. Supports tool calling. Input priced at $0.15 per million tokens. Weights are not publicly released; access is via the provider API.
Alibaba
api
paid
Others in the same category, ranked by how often they are opened.