<- All Models

GLM-5.3 Flash (1M)

HOTNEW

Multimodal Flash on a new 320B MoE base (18B activated, hybrid sparse+linear attention). Image inputs, thinking always on with low/high/max effort. 1M context, 128K output.

1M context·$0.15 in·$0.50 out·Aug 2026 release·Vision, Reasoning, Tools / functions

Providers

Where GLM-5.3 Flash (1M) runs in Big-AGI, with each service's own model ID and USD-per-1M-token rates. The summary above prefers the creator's claim; rates below are each service's own.

ServiceModel IDContextInputOutputCache read

Z.ai

creator

glm-5.3-flash

1M

$0.15

$0.50

$0.03

z-ai/glm-5.3-flash

1.3M

$0.08

$0.25

$0.02

zai-org/GLM-5.3-Flash

1M

$0.15

$0.50

$0.03

Usage mix in Big-AGI: OpenRouter 87% · Community 13%Specs as reported by each service · usage over the last 2 days · refreshed hourly

Run GLM-5.3 Flash (1M) in Big-AGI.

Connect your own key on any of the services above, at their rates, no markup - and run it side by side with every other model.

Launch Big-AGI

© 2026 Token Fabrics·Built with passion in San Diego