Z.AI · alle Modelle ›

ZGLM 5.3 Flash (batch)OPEN

GLM-5.3-Flash ist ein natives multimodales Modell von Z.ai.

1MKontextfenster · Tokens
$0.06Eingabepreis · pro M Tokens
$0.2Ausgabepreis · pro M Tokens
Z.AIAnbieter

Preise werden automatisch aktualisiert — täglich geprüft anhand der Listenpreise der Anbieter.

Siehe Modellvergleiche → Vergleiche alle Modellpreise

Benchmarks & Leistung

Unabhängige Benchmark-Ergebnisse für GLM 5.3 Flash (batch), gemessen von Artificial Analysis. Höher ist besser.

25×
Mehr Intelligenz pro Dollar als Claude Opus 5.5.
For the same budget, GLM 5.3 Flash (batch) delivers 25 times more capability. Open weights also mean you can self-host it and pay nothing per token.
Intelligence index41.8
Coding index71.5
GPQA91.2%
Humanity's Last Exam39.9%
Long Context Reasoning80%
SciCode51.6%
τ-Bench Banking47.2%
Terminal-Bench84.3%
⚡ Speed46.9 tokens/sec
⏱ Latency3.16s to first token
💰 Blended price$0.237 / 1M tokens
📈 Value176.4 intelligence points per $
vs. Modelle, die hier gemessen werdenoberste 11%
Scores höher als 89% der 289 Modelle, die von Artificial Analysis gemessen und hier verfolgt werden.
Modelle auf diesem Niveau kosten $4.56 pro 1M Tokens (Median von 12) — dieses kostet $0.1.
Kein hier verfolgtes Modell ist sowohl günstiger als auch besser in diesem Index.
Benchmark-Daten von Artificial Analysis

Über dieses Modell

GLM 5.3 Flash (batch) ist ein Open-Weight-KI-Modell von Z.AI. Sie können es kostenlos herunterladen und selbst hosten; die unten aufgeführten Preise sind Listenpreise für gehostete APIs, die täglich aktualisiert werden, für den Fall, dass Sie Bequemlichkeit gegenüber Self-Hosting bevorzugen.

Häufig gestellte Fragen

Was ist GLM 5.3 Flash (batch)?

GLM 5.3 Flash (batch) ist ein KI-Sprachmodell von Z.AI. Es ist open-weight: Sie können es herunterladen und auf Ihrer eigenen Hardware kostenlos ausführen. Es erzielt 41.8 im Artificial Analysis Intelligenzindex.

Ist GLM 5.3 Flash (batch) kostenlos?

The weights are free and open — you can self-host GLM 5.3 Flash (batch) and pay nothing per token. If you prefer a hosted API, list prices are $0.06 per million input tokens and $0.2 per million output tokens.

What is GLM 5.3 Flash (batch) good at?

Unabhängige Benchmarks von Artificial Analysis geben ihm GPQA 91.2%, Humanity's Last Exam 39.9%, Long Context Reasoning 80%, SciCode 51.6%, τ-Bench Banking 47.2%, Terminal-Bench 84.3%. Es wird besonders verwendet für Code-Generierung.

How fast is GLM 5.3 Flash (batch)?

Es generiert etwa 46.9 Tokens pro Sekunde, mit einer Median 3.16s Verzögerung vor dem ersten Token. Unabhängig von Artificial Analysis gemessen.

Can I self-host GLM 5.3 Flash (batch)?

Yes. GLM 5.3 Flash (batch) has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.

Verwandte Modelle

GLM 5.3 FlashZ.AIGLM 5.3Z.AIGLM 5.2Z.AIGLM 5.3 FlashXZ.AIGLM 4.7Z.AIGLM 5Z.AI