This model always redirects to the latest model in the GLM Flash family.
Drop-in code to call this model. OpenRouter's API is OpenAI-compatible — most SDKs work by just swapping the base URL. The only thing that changes between models is the model slug below.
This model always redirects to the latest model in the GLM Flash family.
GLM Flash Latest costs $0.075/M input tokens and $0.25/M output tokens, with separate rates for Cache Read at $0.015/M tokens.
GLM Flash Latest has a 1,310,720 token context window. It supports up to 131,072 completion tokens.
Yes. GLM Flash Latest accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.
GLM Flash Latest accepts text, images and video as input and returns text.
GLM Flash Latest was released on August 27, 2026.