Skip to main content

GLM 4.7 Flash

Fast and lightTools

Z.ai · Released January 2026

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

It is part of Z.ai's GLM series, an open-weight line where the 5.x generation leads, Turbo servings cut latency and cost, and 5V adds vision.

Facts and pricing

Provider
Z.ai
Context window
203K tokens
Input price
0.66 credits / 1M tokens ($0.060 raw)
Output price
4.4 credits / 1M tokens ($0.40 raw)
Vision (image input)
No
Tool calling
Yes
Extended reasoning
No

Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.

Try GLM 4.7 Flash now

Ask GLM 4.7 Flash anything. Your prompt opens in the Zeplik app with this model selected.

GLM 4.7 Flash

Or open Zeplik with GLM 4.7 Flash already selected

What GLM 4.7 Flash is best for

Example prompts

Prompts that suit a fast and light model like GLM 4.7 Flash:

Z.ai family

ModelContextInput cr/MOutput cr/MReleased
GLM 5.3 Flash (batch)1.0M1.75.5August 2026
GLM 5.3 Flash1.0M0.832.8August 2026
GLM 5.31.0M15.448.4August 2026
GLM 5.21.0M10.633.4June 2026
GLM 5.2 (free)256K00June 2026
GLM 5.1200K10.633.4April 2026
GLM 5V Turbo203K13.244.0April 2026
GLM 5 Turbo203K13.244.0March 2026
GLM 5198K6.621.1February 2026
GLM 4.7 Flash203K0.664.4January 2026

All 16 Z.ai models

Frequently asked questions

What is GLM 4.7 Flash?
GLM 4.7 Flash is an AI model by Z.ai. As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
How much does GLM 4.7 Flash cost on Zeplik?
Input tokens cost 0.66 credits per million and output tokens 4.4 credits per million (1 credit = $0.10; the raw provider rates are $0.060 and $0.40 per million). New accounts start with free credits.
How long can a conversation with GLM 4.7 Flash be?
GLM 4.7 Flash has a 203K-token context window (202,752 tokens), which covers the conversation plus any documents you attach.
Does GLM 4.7 Flash support images and tools?
GLM 4.7 Flash is text-only and supports tool calling.

Related models

GLM 5.3 Flash (batch)Fast and light

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

1.0M ctx5.5 cr/M outAugust 2026
GLM 5.3 FlashFast and light

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

1.0M ctx2.8 cr/M outAugust 2026
GLM 5.3Open weight

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

1.0M ctx48.4 cr/M outAugust 2026
GLM 5.2Open weight

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

1.0M ctx33.4 cr/M outJune 2026
GLM 5.2 (free)Open weight

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

256K ctx0 cr/M outJune 2026
GLM 5.1Open weight

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

200K ctx33.4 cr/M outApril 2026

More you can do with GLM 4.7 Flash

GLM 4.7 Flash - Z.ai AI Model | Zeplik Chat