<- All Models
Multimodal Flash on a new 320B MoE base (18B activated, hybrid sparse+linear attention). Image inputs, thinking always on with low/high/max effort. 1M context, 128K output.
Available on
Where GLM-5.3 Flash (1M) runs in Big-AGI, with each service's own model ID and USD-per-1M-token rates. The summary above prefers the creator's claim; rates below are each service's own.
Connect your own key on any of the services above, at their rates, no markup - and run it side by side with every other model.
Launch Big-AGIBIG-AGI
Resources
© 2026 Token Fabrics·Built with passion in San Diego