Qwen3.5-Flash
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention m…
Description
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the 3 series, these models deliver a leap forward in performance for both pure text and multimodal tasks, offering fast response times while balancing inference speed and overall performance. This model is served by Alibaba Cloud Int. from Singapore. Save 10% on input tokens and 8% on output tokens compared to standard API rates. Notes: - Context Window: 1,000,000 - Text, Image, & Video input are supported - Built-in tool calls are not supported with video attachments This bot supports optional parameters for additional customization.
Author
EmpirioLabs AI
Platform
web
Pricing model
subscription
Categories
Tags
Capabilities
- Text input
- Image input
- Video input
- Text generation
- By EmpirioLabs AI