z-ai

Z.ai: GLM 5.3 Flash (batch)

This AI model, GLM 5.3 Flash, is offered by vendor z-ai and can handle text, image, and video input modalities with a context length of up to 1 megabyte. It supports tools and reasoning capabilities, but does not provide structured output. GLM 5.3 Flash may be worth considering for those on a budget who require a model that can handle multiple input types with high context capacity. Its price point is $0.06 per million input tokens and $0.2 per million output tokens. With a blended benchmark score of 70.6 across one independent benchmark, it holds some value in terms of performance. Those seeking a cost-effective option within its capabilities may find this model suitable for their needs.

Quality Score
100/100
price + capability + benchmarks
Input Price
$0.06
per 1M tokens · checked 2026-09-27
Output Price
$0.20
per 1M tokens · checked 2026-09-27
Context Window
1,048,576
tokens

Benchmark results

Independent, published benchmarks. Blended score 70.6 across 1 benchmark, last refreshed 2026-09-27. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 69.0
Model ID
z-ai/glm-5.3-flash:batch
Vendor
z-ai
Released
August 2026
Tokenizer
Other
Input Modalities
text, image, video
Output Modalities
text
Max Output
131,072 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
✓ accepts images
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.06/M input and $0.20/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.04
A month of a busy support chatbot 5M in / 2M out $0.70

Price & spec history

Tracked daily by PicksByModel since 2026-08-29.

DateInput /MOutput /MContext
2026-09-23 $0.06 $0.20 1,048,576
2026-09-09 $0.07 $0.25 1,048,576
2026-08-29 $0.15 $0.50 1,048,575

Strong choice for

Category rankings

Where Z.ai: GLM 5.3 Flash (batch) places across the 21 categories it ranks in. How we rank →

#CategoryScore
#3 Image CaptioningVision · of 25 ranked 121
#8 Short-Form SummarizationWriting · of 25 ranked 127
#8 Email DraftingWriting · of 25 ranked 123
#8 Language LearningEducation · of 25 ranked 123
#8 Chat CompanionPersonal · of 25 ranked 127
#8 Trivia & General KnowledgePersonal · of 25 ranked 117
#10 Cheap Bulk InferenceCost · of 25 ranked 137
#11 Video SummarizationVideo · of 25 ranked 148
#12 Code CompletionCode · of 25 ranked 133
#12 Real-Time ChatLatency · of 25 ranked 118
#13 Sales / Cold EmailBusiness · of 25 ranked 119
#13 Job Application DraftingBusiness · of 25 ranked 119
#13 Video Auto-TaggingVideo · of 25 ranked 123
#13 Journaling HelperPersonal · of 25 ranked 119
#13 Recipe GenerationPersonal · of 25 ranked 119
#18 Self-Hosted / LocalCost · of 25 ranked 117
#19 Resume WritingBusiness · of 25 ranked 113
#20 Regex WritingCode · of 25 ranked 126
#20 Language TranslationWriting · of 25 ranked 126
#20 Tutoring for KidsPersonal · of 25 ranked 126
#24 Code DocumentationCode · of 25 ranked 138

Similar models

Quick answers

How much does Z.ai: GLM 5.3 Flash (batch) cost?
$0.06 per million input tokens and $0.20 per million output tokens.
What is Z.ai: GLM 5.3 Flash (batch)'s context window?
1,048,576 tokens, roughly 1,572 pages of text in a single request.
Does Z.ai: GLM 5.3 Flash (batch) support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can Z.ai: GLM 5.3 Flash (batch) process images?
Yes, it accepts image input.

The Model Movers Report

One email every Friday, built from this site's own rankings: the current top five by benchmark score, every model released in the last seven days, and one note worked out from that week's numbers. You can unsubscribe from any issue with one click.