GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
Los precios se actualizan automáticamente; se verifican a diario contra las tarifas oficiales de los proveedores.
Artificial Analysis mide 289 de los 424 modelos rastreados aquí — este aún no está cubierto. La posición de precios anterior se mide a través de todo el catálogo y no depende de las referencias.
GLM 5.3 FlashX is an open-weight AI model by Z.AI. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.
GLM 5.3 FlashX es un modelo de lenguaje de IA de Z.AI. Es open-weight: puede descargarlo y ejecutarlo en su propio hardware, de forma gratuita.
The weights are free and open — you can self-host GLM 5.3 FlashX and pay nothing per token. If you prefer a hosted API, list prices are $0.37 per million input tokens and $1.25 per million output tokens.
Yes. GLM 5.3 FlashX has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.