GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration.
Les prix se mettent à jour automatiquement — vérifiés quotidiennement par rapport aux prix catalogue des fournisseurs.
Artificial Analysis mesure 289 des 425 modèles suivis ici — celui-ci n'est pas encore couvert. La position tarifaire ci-dessus est mesurée sur l'ensemble du catalogue et ne dépend pas des référentiels.
GLM 5.3 Prime is an open-weight AI model by Z.AI. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.
GLM 5.3 Prime est un modèle de langage IA de Z.AI. Il est open-weight : vous pouvez le télécharger et l'exécuter sur votre propre matériel, gratuitement.
The weights are free and open — you can self-host GLM 5.3 Prime and pay nothing per token. If you prefer a hosted API, list prices are $2.8 per million input tokens and $8.8 per million output tokens.
Yes. GLM 5.3 Prime has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.