Zhipu AI

GLM-4.6V — Benchmark Scores, Pricing & Performance Analysis

Output Speed
68 tok/s
Input Cost
$0.30/1M
Output Cost
$0.90/1M
Context Window
131K
Max Output
33K
Knowledge Cutoff
Apr 2025
Accepts
text, image, video

GLM-4.6V by Zhipu AI demonstrates competitive pricing. View detailed benchmark data including scores across coding, math, reasoning, speed, and cost metrics.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

Speed Benchmarks

Time to First Token
770 ms69th of 92

Latency before first token arrives

Aggregated
Output Speed
68 tok/s57th of 94

Tokens generated per second

Aggregated

Cost Benchmarks

Input Cost
$0.30204th of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$0.90190th of 413

Cost per 1M output tokens

OpenRouter API
Cached Input Cost
$0.0648th of 136

Cost per 1M cached input tokens — the price that actually applies to a long agent conversation

OpenRouter API

Context Benchmarks

Context Length
131K212th of 451

Maximum context window size

OpenRouter API

Available from 13 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 2 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
ZenMux$0.14$0.42$0.03200K—
302.AI$0.145$0.43—128K—
DevPass (LLM Gateway)$0.3$0.9$0.05131K—
Eden AI$0.3$0.9$0.05131K—
Kilo Gateway$0.3$0.9$0.055131K—
LLM Gateway$0.3$0.9$0.05128K—
NanoGPT$0.3$0.9$0.15128K—
NovitaAI$0.3$0.9$0.055131K—
OpenRouter$0.3$0.9$0.055131KSupported
Z.AI$0.3$0.9—128K—
Zhipu AI$0.3$0.9—128KSupported
Zhipu AI Coding Plan$0.3$0.9—128K—
Poe———131K—

GLM-4.6V — Frequently Asked Questions

How much does GLM-4.6V cost?

GLM-4.6V costs $0.30 per 1M input tokens and $0.90 per 1M output tokens. This makes it one of the more affordable models.

How fast is GLM-4.6V?

GLM-4.6V generates output at 68 tokens per second, which is slower, prioritizing quality over speed compared to other models. The time to first token is 770 ms.

What is the context window of GLM-4.6V?

GLM-4.6V has a context window of 131K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created GLM-4.6V?

GLM-4.6V was created by Zhipu AI. It is classified as a mid model in our catalogue.

Is GLM-4.6V open source?

No, GLM-4.6V is a proprietary model. It is available through Zhipu AI's API and compatible providers.