Model catalogue / GLM
GLM-5.3 Flash
GLM-5.3 Flash on VipAI costs $0.70 per 1M input tokens and $2.20 per 1M output. It has a 1M context window.
$0.70Input / 1M tokens
$2.20Output / 1M tokens
$0.13Cache read / 1M tokens
1MContext window
Price against list
The list column is the provider's own published rate, so the discount can be checked rather than taken on trust.
| Rate | Provider list | VipAI | Discount |
|---|---|---|---|
| Input / 1M tokens | $0.70 | $0.70 | — |
| Output / 1M tokens | $2.20 | $2.20 | — |
| Cache read / 1M tokens | — | $0.13 | — |
Endpoints it answers on
The router does not translate protocols silently, so use the shape the family expects. The full matrix is in models & routing.
| Protocol | Method | Path |
|---|---|---|
| OpenAI Chat Completions | POST | /v1/chat/completions |
Calling it
curl https://api.vipai.site/v1/chat/completions \
-H "Authorization: Bearer $VIPAI_API_KEY" \
-H "Content-Type: application/json" \
-d '{ "model": "glm-5.3-flash", "messages": [{ "role": "user", "content": "Ship it." }] }'Use it
- Create an API key — it works with every model on the key, not just this one.
- API integration — headers, bodies and streaming for each protocol.
- Billing questions — why one message can produce several billing rows.
- All model prices — the full catalogue against list.
Prices are live and move with upstream costs. This page reflects the rate in effect now.