Models/Zhipu AI/GLM 5.3
Z
Zhipu AI

GLM 5.3

zhipu-ai/glm-5.3
Live on DIT
Get API Key

$1.40 input / $4.40 output per 1M tokens; $0.26 per 1M cached input tokens.

ModalitiesText Text
DIT in / out$1.12 / $2.64per 1M tokens
Context1Mreviewed context window
Max output131.1Ktokens per response
ReleasedAug 14, 2026OPENAI compatible
Capabilities
Reasoning
Tool calling
Structured outputs

Overview

Recommended use cases and a precise technical profile, generated from the same reviewed model record used throughout this page.

Technical profileReviewed model metadata
Model IDglm-5.3ProviderZhipu AI
ProtocolOPENAI compatibleDIT availability
2 active routes
InputTextOutputText
Context window1,000,000 tokensMaximum output131,072 tokens
ReleasedAug 14, 2026Knowledge cutoff
Reasoning controls
Official identifiers and selectable reasoning effort
Aliases
No alias published
Reasoning effort
No default published
Supported API endpointsOfficial model API surface
EndpointPathStatus
Chat Completions/v1/chat/completions
Supported

Providers

Compare DIT market routing with the official upstream reference across price, speed, and availability. Supplier identities remain private.

Market coverage2 active routes
Selection strategyPrice + health + automatic failover
Runtime window30 days
ProviderInput /MOutput /MCache read /MLatencyThroughputUptime
DIT Market
40% off
Aggregated supply · 2 active routes
$1.4$1.12$4.4$2.644923 msP50 TTFT56.4 tpsP50 output44.17%30-day success
ZZhipu AI official
Reference
Direct upstream · reviewed list price
$1.4$4.4Awaiting dataAwaiting dataNo status source

Prices are USD per 1M tokens. DIT performance and uptime appear only after 50 eligible requests; supplier identities and customer traffic remain private.

Pricing

Compare DIT market rates directly with reviewed official prices so savings remain visible and auditable.

Current price comparison
Effective DIT rate compared with the reviewed official list price.
Effective
USD / 1M tokens
Up to 40% lower
Direct rate comparison · reviewed standard prices per 1M tokens
Price composition
DIT effective rate mix
$3.76combined rate
InputOutput
Relative rate composition, not estimated invoice share
Price comparisonReviewed source ↗
UsageOfficialDIT marketSavingBilling basis
Input tokens$1.4$1.12~40%Per 1M tokens
Output tokens$4.4$2.64~40%Per 1M tokens
Cached inputPer 1M cached tokens

Performance

Latency, throughput, activity, and request success use privacy-safe DIT runtime aggregates only when the eligibility threshold is met.

Throughput
56.4 tok/s
P50 across eligible DIT traffic
Latency
4923 ms TTFT
Lower is better
Runtime sample
240
30-day privacy-safe window
All locations
P50 / P95
30-day eligible DIT traffic · privacy-safe aggregates
Latency distribution
Time to first token and end-to-end
Lower is better · aggregated eligible requests only
Traffic activity
Privacy-safe normalized request index
Relative shape only; customer traffic totals remain private
Request success
Eligible DIT request completion rate
44.17%
System probes and user fallback traffic excluded
DIT runtime distribution30-day eligible traffic
MetricP50P95Reading
Time to first byte4923 ms7779 msLower is better
End-to-end latency9592 msLower is better
Output throughput56.4 tpsHigher is better

Uptime

A concise service-health view with scope and source stated next to the number.

44.17%30-day eligible request success
~240 eligible requestsSystem probes and user fallback traffic excluded

Quick Start

A copy-ready request for the compatible DIT endpoint using this model's reviewed identifier.

cURLOPENAI compatible
curl https://api.dit.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "glm-5.3",
    "messages": [{"role":"user","content":"Explain dynamic model routing."}]
  }'

Sources

Field-level provenance keeps pricing, limits, capabilities, and benchmarks independently reviewable.

Reviewed sources3 references · reviewed Sep 18, 2026
SourceTypeCoversChecked
Zhipu AI model documentation
Official
Context, Modalities, CapabilitiesSep 18, 2026
models.dev catalog
models.dev
Context, Modalities, PricingSep 18, 2026
DIT live model catalog
DIT catalog
DIT availability, PricingSep 18, 2026

Official sources take priority. Third-party metadata is retained only when its field and review date are shown.

Models like GLM 5.3

Reviewed models ranked by capability overlap, protocol compatibility, and context similarity.

Frequently asked questions

Answers use the same reviewed facts and pricing fields presented in the tables above.

What is GLM 5.3?+

$1.40 input / $4.40 output per 1M tokens; $0.26 per 1M cached input tokens.

How much does GLM 5.3 cost?+

The reviewed official rate is $1.4 per 1M input tokens and $4.4 per 1M output tokens. DIT pricing is shown only when a live DIT route exists.

What is the context length of GLM 5.3?+

GLM 5.3 supports a reviewed context window of 1,000,000 tokens and up to 131,072 output tokens.

What capabilities does GLM 5.3 support?+

Reviewed capabilities include Reasoning, Tool calling, Structured outputs. Consult the linked official documentation for endpoint-specific limits.

What inputs and outputs does GLM 5.3 support?+

Reviewed input modalities: Text. Reviewed output modalities: Text.

Which API endpoints and reasoning levels does GLM 5.3 support?+

GLM 5.3 is documented for Chat Completions.

Is GLM 5.3 available on DIT?+

GLM 5.3 currently has 2 active routes. DIT selects eligible supply by price and health, with automatic failover when another route is available.

When was GLM 5.3 released?+

GLM 5.3 was released on Aug 14, 2026.