<- All Models

GLM 5.3 Prime

NEW

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

1M context·$2.80 in·$8.80 out·Sep 2026 release·Reasoning, Tools / functions, Web search

Available on

Providers

Where GLM 5.3 Prime runs in Big-AGI, with each service's own model ID and USD-per-1M-token rates. The summary above prefers the creator's claim; rates below are each service's own.

ServiceModel IDContextInputOutputCache read

z-ai/glm-5.3-prime

1M

$2.80

$8.80

$0.56

Specs as reported by each service · usage over the last 2 days · refreshed hourly

Run GLM 5.3 Prime in Big-AGI.

Connect your own key on any of the services above, at their rates, no markup - and run it side by side with every other model.

Launch Big-AGI

© 2026 Token Fabrics·Built with passion in San Diego