GLM 4.5 Flash model ID, context window & pricing
glm-flash
Quick facts
Model ID glm-4-5-flash
Source EmpirioLabs AI
Context Window 200000
Pricing -
Capabilities tool calling, reasoning, structured output, temperature control
Model overview
GLM 4.5 Flash is an AI model from EmpirioLabs AI with 200000 token context window and text input support.
Public token pricing is not listed for this model in the current catalog source.
- Workloads that use text inputs with text outputs.
- Agent and tool workflows that need function calling.
- Reasoning-heavy prompts where stepwise problem solving matters.
Model ID glm-4-5-flash
Provider EmpirioLabs AI
Family glm-flash
Status -
Knowledge Cutoff 2025-04
Release Date 2025-07-28
Input Modalities text
Output Modalities text
Context Window 200000
Input Limit -
Output Limit 98304
Tool Calling Yes
Reasoning Yes
Structured Output Yes
Temperature Control Yes
Open Weights No
Input Cost / 1M tokens -
Output Cost / 1M tokens -
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens -
Cache Write Cost / 1M tokens -