Zhipu active

GLM-5.3-flash API Pricing and Package Multipliers

Model glm-5.3-flash

GLM-5.3-flash

Pay-as-you-go

GLM-5.3-flash API Pricing

Official base: Input 0.15 / Output 0.5 / Cache read 0.03 / Cache write 0
Channel Billing unit Quota multiplier Final discount
Zhipu USD / 1M 5.376x 20% off

Cache hits are billed at the cache read price. Cache writes use the cache write price, and the Console usage detail shows the final cost breakdown.

Package Availability

GLM-5.3-flash available packages and multipliers

Package Channel Billing unit Quota multiplier Final discount
Plus Zhipu USD / 1M 42.3191x 8.3% off
Max Zhipu USD / 1M 42.3191x 26.6% off
Super Ultra Zhipu USD / 1M 42.3191x 40.1% off
Lite Zhipu USD / 1M 42.3191x 1.02× official
Pro Zhipu USD / 1M 42.3191x 14.4% off
Ultra Zhipu USD / 1M 42.3191x 38.8% off
SU500 Zhipu USD / 1M 42.3191x 43.5% off
SU750 Zhipu USD / 1M 42.3191x 44.3% off
SU1000 Zhipu USD / 1M 42.3191x 43.5% off
Integration

GLM-5.3-flash OpenAI-compatible API access

After creating a key in the console, call this model with the unified model name glm-5.3-flash. When package and pay-as-you-go balance are both available, package quota is used first.

model: "glm-5.3-flash"
Best For
  • Good as a general-purpose model for text generation and everyday reasoning.
  • Good when you want package quota first, with pay-as-you-go enabled by key configuration.
Not Ideal For
  • Not ideal for legacy integrations that only accept OpenAI-compatible APIs.
  • Not ideal when you need dedicated SLA terms, private capacity, or a fixed-cost enterprise commitment.
Migration Notes

This model is available through the public catalog and unified gateway pricing configuration.

Open this model in the console

The public model page is for discovery and comparison. Actual API access, key creation, and model scopes are managed after sign-in.

Related Models

Models worth comparing