T
GLM-4-32B-0414
thudm/glm-4-32b-0414
available
chat
hosted: global
zero_retention
A chat model in the thudm family, served through Thali's OpenAI-compatible API with rupee billing and per-request cost attribution.
Context window
32,000
tokens
Input
₹59.80
per million tokens
Output
₹181.70
per million tokens
Licence
varies
verify before production use
Call this model
from openai import OpenAI
client = OpenAI(
base_url="https://www.thaliai.in/api/v1",
api_key="thali-sk-...",
)
completion = client.chat.completions.create(
model="thudm/glm-4-32b-0414",
messages=[{"role": "user", "content": "Hello"}],
)
Works with any OpenAI SDK. Streaming, fallback lists and provider preferences are documented in the routing guide.
This model is served via global infrastructure. For workloads
that must stay in India, filter the catalog for
hosted_in: "in".