GLM 4.6Von Krater.ai
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts
Released December 8, 2025
Specifications
GLM 4.6V at a glance
Context
131.1K tokens
Max output
32.8K tokens
Input
Image, Text, Video
Output
Text
Input price
$0.30 / 1M
Output price
$0.90 / 1M
Capabilities
What it is good at
Catalog capabilities
- Accepts image input
- Generates text responses
- Supports tool calling
Good for
Compare
Compare GLM 4.6V
GLM 4.6V vs GLM 5V Turbo
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.
Open comparisonGLM 4.6V vs GLM 4.5V
GLM-4.5V is a vision-language foundation model for multimodal agent applications.
Open comparisonGLM 4.6V vs GLM 4.5
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications.
Open comparisonFAQ
Questions about GLM 4.6V
Clear answers based on the model catalog and Krater access.
Use GLM 4.6V in your API
Connect with Krater's unified API.