AgentHub

GLM-4.6V-N

GLM-4.6V represents a significant multimodal advancement in the GLM series, achieving state-of-the-art visual understan…

Live
Open / InstallLast updated July 27, 2026

Description

GLM-4.6V represents a significant multimodal advancement in the GLM series, achieving state-of-the-art visual understanding accuracy for models of its parameter scale. Notably, it's the first visual model to natively integrate Function Call capabilities directly into its architecture, creating a seamless pathway from visual perception to executable actions. This breakthrough establishes a unified technical foundation for deploying multimodal agents in real-world business applications. File Support: Text, Markdown, Image and PDF files Context window: 131k tokens Optional parameters: Enable Thinking - Toggle this on for the model to think before providing a response. This is disabled by default Temperature - Controls randomness in the response. Lower values make the output more focused and deterministic. Select from 0 to 2 range. This is set to 0.7 by default. Max Output Tokens: Maximum number of tokens to generate in the response. This can be set from 1 to 32768. Set to Max token at 32768 by default.

Author

Novita AI

Platform

web

Pricing model

subscription

Categories

Design
Productivity

Tags

poe
novita-ai
multimodal
text
image

Capabilities

  • Text input
  • Image input
  • Text generation
  • By Novita AI