whichllm — Browse and compare AI model specs and pricing

LLM Gateway

GLM-4.7 Flash (EmberCloud) model ID, context window & pricing

glm-flash

Quick facts

Model ID embercloud/glm-4.7-flash
Source LLM Gateway
Context Window 200000
Pricing $0.06 input / $0.40 output per 1M tokens
Capabilities tool calling, temperature control, open weights

Model overview

GLM-4.7 Flash (EmberCloud) is an AI model from LLM Gateway with 200000 token context window and text input support.

Published pricing is $0.06 input and $0.40 output per 1M tokens.

  • Workloads that use text inputs with text outputs.
  • Agent and tool workflows that need function calling.
Model ID embercloud/glm-4.7-flash
Provider LLM Gateway
Family glm-flash
Status -
Knowledge Cutoff 2025-04
Release Date 2026-01-19
Input Modalities text
Output Modalities text
Context Window 200000
Input Limit -
Output Limit 131000
Tool Calling Yes
Reasoning No
Structured Output No
Temperature Control Yes
Open Weights Yes
Input Cost / 1M tokens $0.06
Output Cost / 1M tokens $0.40
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens $0.01
Cache Write Cost / 1M tokens -