GLM
GLM-5V Turbo
z-ai/glm-5v-turbo
Z.ai's vision-coding model with a 200K context — turns screenshots and designs into code, debugs visually, and drives GUI automation.
Pricing
| Input | $1.20per 1M tokens |
|---|---|
| Output | $4.00per 1M tokens |
| Cached input | $0.24per 1M tokens read from cache |
Specifications
| Context window | 200Ktokens |
|---|---|
| Max output | 66Ktokens |
| Provider | GLM |
| Capabilities | Vision · Reasoning · Function calling |
| Input types | text, image |
Use GLM-5V Turbo
Point your existing OpenAI SDK at xKiro and pass z-ai/glm-5v-turbo as the model. Nothing else in your code changes.
from openai import OpenAI
client = OpenAI(
base_url="https://api.xkiro.com/v1",
api_key="YOUR_XKIRO_KEY",
)
response = client.chat.completions.create(
model="z-ai/glm-5v-turbo",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Supported parameters
max_tokens · temperature · top_p · stop · frequency_penalty · presence_penalty · seed · stream · tools · tool_choice · response_format · structured_outputs · reasoning · include_reasoning
Related models
Call GLM-5V Turbo through the same endpoints you already use.
Get an API key