AgentHub

Qwen3.5-397B-A17B

The Qwen3.5 series 397B-A17B native vision-language model is based on a hybrid architecture design that integrates line…

Live
Open / InstallLast updated July 27, 2026

Description

The Qwen3.5 series 397B-A17B native vision-language model is based on a hybrid architecture design that integrates linear attention mechanisms with sparse Mixture-of-Experts (MoE), achieving higher inference efficiency. Across a variety of tasks—including language understanding, logical reasoning, code generation, agentic tasks, image understanding, video understanding, and graphical user interface (GUI) interaction—it demonstrates exceptional performance comparable to current top-tier frontier models. Possessing robust code generation and agentic capabilities, it exhibits strong generalization across various agent scenarios. File Support: Text, Markdown, Image, Video and PDF files Context window: 262k tokens Optional parameters: Enable thinking about the response before giving a final answer: toggle it `on`, otherwise it is `off` by default. Set temperature to control randomness in the response: Set number from 1 to 2. This is set to `0.7` by default. Lower values make the output more focused and deterministic. Set max output tokens: Set number from 1 to 64000. This is set to 64000 by default.

Author

Novita AI

Platform

web

Pricing model

subscription

Categories

Coding
Design
Research
Productivity

Tags

poe
novita-ai
multimodal
text
image
video

Capabilities

  • Text input
  • Image input
  • Video input
  • Text generation
  • By Novita AI