whichllm — Browse and compare AI model specs and pricing

Nebius Token Factory

GLM-5.3-Flash model ID, context window & pricing

glm

Quick facts

Model ID zai-org/GLM-5.3-Flash
Source Nebius Token Factory
Context Window 1024000
Pricing $0.15 input / $0.50 output per 1M tokens
Capabilities tool calling, reasoning, structured output, temperature control, open weights

Model overview

GLM-5.3-Flash is an AI model from Nebius Token Factory with 1024000 token context window and text input support.

Published pricing is $0.15 input and $0.50 output per 1M tokens.

  • Workloads that use text inputs with text outputs.
  • Agent and tool workflows that need function calling.
  • Reasoning-heavy prompts where stepwise problem solving matters.
Model ID zai-org/GLM-5.3-Flash
Provider Nebius Token Factory
Family glm
Status -
Knowledge Cutoff -
Release Date 2026-08-26
Input Modalities text
Output Modalities text
Context Window 1024000
Input Limit -
Output Limit 1024000
Tool Calling Yes
Reasoning Yes
Structured Output Yes
Temperature Control Yes
Open Weights Yes
Input Cost / 1M tokens $0.15
Output Cost / 1M tokens $0.50
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens $0.15
Cache Write Cost / 1M tokens -