GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
Preise werden automatisch aktualisiert — täglich geprüft anhand der Listenpreise der Anbieter.
Artificial Analysis misst 289 der 425 Modelle, die hier verfolgt werden — dieses ist noch nicht abgedeckt. Die oben angegebene Preisposition wird über den gesamten Katalog gemessen und hängt nicht von Benchmarks ab.
GLM 5.3 FlashX is an open-weight AI model by Z.AI. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.
GLM 5.3 FlashX ist ein KI-Sprachmodell von Z.AI. Es ist open-weight: Sie können es herunterladen und auf Ihrer eigenen Hardware kostenlos ausführen.
The weights are free and open — you can self-host GLM 5.3 FlashX and pay nothing per token. If you prefer a hosted API, list prices are $0.37 per million input tokens and $1.25 per million output tokens.
Yes. GLM 5.3 FlashX has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.