Z
GLM-5V-Turbo
zai-org/glm-5v-turbo
available
chat
hosted: global
zero_retention
Zhipu's GLM models are capable bilingual generalists; the Air variants trade a little quality for substantially lower cost.
Context window
204,800
tokens
Input
₹131.10
per million tokens
Output
₹437.00
per million tokens
Licence
varies
verify before production use
Call this model
from openai import OpenAI
client = OpenAI(
base_url="https://thaliai.in/api/v1",
api_key="thali-sk-...",
)
completion = client.chat.completions.create(
model="zai-org/glm-5v-turbo",
messages=[{"role": "user", "content": "Hello"}],
)
Works with any OpenAI SDK. Streaming, fallback lists and provider preferences are documented in the routing guide.
This model is served via global infrastructure. For workloads
that must stay in India, filter the catalog for
hosted_in: "in".