Skip to content
Models/zhipu/GLM 5.3 Flash

GLM 5.3 Flash

glm-5.3-flashText·

Zhipu GLM 5.3 Flash vision model, 18B active parameters, image and video input, 1M context, always-on thinking

Context Window
1.0M
Input Price /M in
$0.12$0.06/M in
Output Price /M out
$0.43$0.22/M out
Cached Input Price /M
$0.02/M
Max Completion
131K
Input Modalities
text, image
Output Modalities
text
reasoningFunction callingChatVisionJSONStreamingNew-50%fastGeneralProgramming

Description

Zhipu GLM 5.3 Flash vision model, 18B active parameters, image and video input, 1M context, always-on thinking

Available Providers

AllToken can route requests to the providers below based on route priority and policy.

ProviderContextInputOutputCached / MLatencyThroughput

Best For

Zhipu GLM 5.3 Flash vision model, 18B active parameters, image and video input, 1M context, always-on thinking

How To Use This Model

Use the exact model ID shown below. This is the safest way to avoid call failures, variant mismatches, or incorrect route assumptions.

curl https://api.alltoken.ai/v1/chat/completions \
  -H "Authorization: Bearer sk-your-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "glm-5.3-flash",
    "messages": [
      {"role": "user", "content": "Hello!"}
    ]
  }'
Supported Parameters
temperaturetop_pmax_tokenstoolsresponse_formatreasoning_effort
API Key Setup
Smart Routing

Let the platform choose the best provider path automatically.

Default Model

If a request does not specify a model, default the key to glm-5.3-flash.

Forced Model

Always override incoming requests and lock the key to glm-5.3-flash.