Skip to content

GLM-4V

zhipu/glm-4v
Vision

Vision-language member of the GLM family for grounded image understanding.

Model overview

Context window

64K

tokens

Input price

quoted on request

Output price

quoted on request

Weekly volume

tokens / week

Open weights

No

API only

Capabilities

2

of 11 tags

Description

GLM-4V is a vision model from Zhipu AI. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. Zhipu AI’s stated focus is open agentic GLM family. The weights are not published, so it is API-only.

Capabilities

Vision input
Accepts images alongside text in the prompt.
Streaming
Emits tokens as they are generated, so answers appear progressively rather than all at once.

Strengths & weaknesses

Strengths

  • Solid grounding and localisation
  • Consistent with GLM tooling
  • Fair pricing

Weaknesses

  • Smallest context of the vision group
  • No tool calling

Pricing

Free
Pricing for GLM-4V, per 1M tokens
RatePriceUnit
Input$0.00per 1M tokens
Output not charged

Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.

Context window

64Ktokens

covers prompt and response together, so a long input leaves less room for the answer.

64K tokens against a peer range of 32K to 128K, median 128K. Compared across 5 vision models.

Supported features

Feature support for GLM-4V
FeatureSupport
Tool callingNot supported
JSON modeNot supported
StreamingSupported
Vision inputSupported
Audio inputNot supported
Long contextNot supported
Open weightsNot supported
ReasoningNot supported
Fine-tunableNot supported
BatchNot supported
CachingNot supported

Example use cases

  • Image QA

    Solid grounding and localisation

  • Accessibility descriptions

    Consistent with GLM tooling

  • Visual grounding

    Fair pricing

Comparison

GLM-4V compared with Kimi VL and Qwen VL Max
AttributeGLM-4VKimi VLQwen VL Max
Context window64K128K128K
Input price$0.00$0.00$0.00
Output price not charged not charged not charged
Tool callingNot supportedNot supportedSupported
JSON modeNot supportedNot supportedNot supported
Vision inputSupportedSupportedSupported
Open weightsNot supportedNot supportedNot supported
Prompt cachingNot supportedNot supportedNot supported

Best value in each row is highlighted. Illustrative placeholder data.

Provider information

View Zhipu AI
Zhipu AI3 models

Zhipu AI · 智谱AI · Beijing, China · est. 2019

Born from Tsinghua University research, Zhipu's GLM family became the open ecosystem's agentic specialist: models tuned for tool use and multi-step work, released under permissive licences. Its free Air tier made capable open models accessible to every developer.

Recent releases

    Documentation

    Frequently asked questions

    Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Every figure on this page is an illustrative placeholder.