GLM API Pricing — KeyoAPI Rates & Zhipu Context
This page puts current KeyoAPI sell rates for the Zhipu GLM 5.x chat family next to official Zhipu context, with GLM ASR and TTS on the same key.
Price comparison table
| Model ID | Typical official / context | KeyoAPI | Notes |
|---|---|---|---|
glm-5.3 |
Zhipu GLM 5.3 — verify on open.bigmodel.cn | ~$0.24 / $0.77 per 1M in/out | Current-generation everyday GLM lane |
glm-5.3-flash |
Zhipu GLM 5.3 Flash — verify on open.bigmodel.cn | ~$0.22 / $0.72 per 1M in/out | Cheapest GLM chat lane on Keyo |
glm-5.2 |
Zhipu GLM 5.2 — verify on open.bigmodel.cn | ~$1.40 / $4.40 per 1M in/out | Own band; paid twin of glm-5.2:free |
Prototype free, then meter GLM on the same key
For $0 prototyping on the same key, use glm-5.2:free from /free-models — a $0 fixed-ID twin whose calls draw down your $10 welcome credit (fair-use rate limits). When you meter GLM traffic, switch model= to glm-5.3-flash or glm-5.3 without changing base URL or auth.
Rates and how to call
The GLM family on KeyoAPI covers three chat lanes and two speech lanes on one OpenAI-compatible key: glm-5.3 and glm-5.3-flash for chat, glm-5.2 at its own band, plus GLM-ASR (~$0.007 per request) and GLM-TTS (~$0.014 per request). No separate Zhipu account, no second invoice.
Routing pattern: put glm-5.3-flash under high-volume classification, drafts, and summarization; use glm-5.3 as the everyday product lane; keep glm-5.2 where your stack already pins that exact ID. Rates above are indicative sell prices — confirm live rows on /pricing-list before volume contracts.
How to call: keep your OpenAI SDK and set the base URL to https://www.keyoapi.xyz/v1, then pass model=glm-5.3-flash (or any lane above) to POST /v1/chat/completions with your Bearer key. Read usage.prompt_tokens and usage.completion_tokens from each response to reconcile spend against the table above.
For broader planning: compare GLM rates against OpenAI-class dollars on /openai-api-pricing, DeepSeek on /deepseek-api-pricing, and Claude on /claude-api-pricing — all on the same prepaid wallet. Side-by-side planning: /compare.
curl https://www.keyoapi.xyz/v1/chat/completions \
-H "Authorization: Bearer $KEYO_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"glm-5.3-flash","messages":[{"role":"user","content":"Hello"}]}'
FAQ
How do I call the GLM API on KeyoAPI?
Set baseURL to https://www.keyoapi.xyz/v1, Authorization: Bearer YOUR_KEY, and model=glm-5.3, glm-5.3-flash, or glm-5.2 on POST /v1/chat/completions.
Is there a free GLM tier?
Yes — glm-5.2:free is a fixed $0 catalog ID on /free-models, covered by your $10 welcome credit with fair-use rate limits. Switch model= to a metered GLM lane on the same key when you outgrow it.
Do you carry GLM ASR and TTS?
Yes. GLM-ASR (~$0.007 per request) and GLM-TTS (~$0.014 per request) run on the same Keyo key — see /brand/keyo-docs.html for their request shapes.