Free test tokens — try any model

Model catalogue / GLM

GLM-5.3 Flash (free)

GLM-5.3 Flash (free) on VipAI costs $0.007 per 1M input tokens and $0.022 per 1M output — 99% off the provider's published rate. It has a 1M context window.

$0.007Input / 1M tokens
$0.022Output / 1M tokens
$0.0013Cache read / 1M tokens
1MContext window

Price against list

The list column is the provider's own published rate, so the discount can be checked rather than taken on trust.

RateProvider listVipAIDiscount
Input / 1M tokens$0.70$0.00799% off
Output / 1M tokens$2.20$0.022—
Cache read / 1M tokens—$0.0013—

Endpoints it answers on

The router does not translate protocols silently, so use the shape the family expects. The full matrix is in models & routing.

ProtocolMethodPath
OpenAI Chat CompletionsPOST/v1/chat/completions

Calling it

request.sh
curl https://api.vipai.site/v1/chat/completions \
  -H "Authorization: Bearer $VIPAI_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{ "model": "glm-5.3-flash-free", "messages": [{ "role": "user", "content": "Ship it." }] }'

Use it

Prices are live and move with upstream costs. This page reflects the rate in effect now.