GLM 5 Turbo

glm/glm-5-turboAvailable

GLM-5-Turbo by Z.ai — speed-optimized GLM-5 for high-throughput agentic workloads. OpenAI-compatible: pass model=glm/glm-5-turbo to /v1/chat/completions.

Retail priceinput $2.4e-06 · output $8e-06 · cached input $4.8e-07
Providerglm
Capabilitychat
InterfacePOST /v1/run

Request example

Every paid call needs a unique Idempotency-Key. The same key and payload replay the original result without a second charge.

cURL
curl https://router.tycoon.cool/v1/run \
+  -H "Authorization: Bearer $TYCOONROUTER_API_KEY" \
+  -H "Content-Type: application/json" \
+  -H "Idempotency-Key: $(uuidgen)" \
+  -d '{"model":"glm/glm-5-turbo","inputs":{"messages":[{"role":"user","content":"Hello"}]}}'

Response contract

Synchronous routes return the provider result with TycoonRouter request metadata. Async routes return a request ID plus status and cancel URLs.

HeaderPurpose
x-request-idImmutable Router request identity
x-idempotent-replayTrue when the response is a replay
x-tycoonrouter-charge-microsSettled retail charge when available

Production safeguards

Unknown prices, unbounded requests and uncertain provider outcomes fail closed before a second paid dispatch. Usage and wallet events remain queryable from the account page.