Zhipu GLM

GLM-5 Turbo API

CN Zhipu GLM | released 2026-06
Try in Playground → Get API key

GLM-5 Turbo is the latency-optimized version of the GLM-5 family. It keeps the family’s strong tool-use and bilingual ability while cutting time-to-first-token, making it well suited for real-time assistants and interactive features.

FastTool UseJSON Mode

Pricing (official list, per 1M tokens)

Input
$0.71/M
Output
$3.14/M

Specifications

Context window200K tokens Max output4.096K tokens Input modalitiesText StreamingYes APIOpenAI-compatible

How to use GLM-5 Turbo

Drop-in OpenAI replacement. Change only the base_url and model.

# Python -pip install openai
from openai import OpenAI

client = OpenAI(
    base_url="https://meshtok.com/v1",
    api_key="sk-your-MeshTok-key",
)

resp = client.chat.completions.create(
    model="bigmodel/glm-5-turbo",
    messages=[{"role":"user","content":"Hello!"}],
)
print(resp.choices[0].message.content)
# cURL
curl https://meshtok.com/v1/chat/completions \
  -H "Authorization: Bearer sk-your-MeshTok-key" \
  -H "Content-Type: application/json" \
  -d '{"model":"bigmodel/glm-5-turbo","messages":[{"role":"user","content":"Hello!"}]}'

Related models

Other models you might consider.