llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Zhipu AI model product

GLM 4.7 Flash

7 optimized for fast inference and lower latency.

Updated Aug 17, 2026. Default version: GLM-4.7-Flash

Compare
LLMBoard Score37.9GLM-4.7-Flash
Coverage80%6 benchmark families
Context window200KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

GLM 4.7 Flash Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

GLM-4.7-Flash LLMBoard score breakdown

GLM 4.7 Flash Benchmark Results

Benchmark scores for GLM-4.7-Flash.

6 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkTau-benchScore79.5%Rank04Participants6Percentile40.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2025Score91.6%Rank41Participants115Percentile64.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseCompScore42.8%Rank54Participants62Percentile13.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore14.4%Rank80Participants99Percentile19.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore59.2%Rank86Participants111Percentile22.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore75.2%Rank109Participants239Percentile54.6%EvidenceCEvaluatedAug 17, 2026

GLM 4.7 Flash Arena Results

Preference and agent-evaluation results for the default version.

3 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenatext factualityCategoryoverallRank146Rating / score1372.3Votes11,632ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank176Rating / score1367.4Votes11,760ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank179Rating / score1353.0Votes11,760ObservationsN/AResult dateAug 12, 2026

GLM 4.7 Flash Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.04 input, $0.30 output per 1M via CrofAI
Tracked offerings
24
23 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderHugging FaceProvider model IDzai-org/GLM-4.7-FlashRegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedAug 17, 2026
ProviderZ.AIProvider model IDglm-4.7-flashRegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedAug 17, 2026
ProviderZhipu AIProvider model IDglm-4.7-flashRegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedAug 17, 2026
ProviderEmpirioLabs AIProvider model IDglm-4-7-flashRegionglobalInput / 1MN/AOutput / 1MN/AContext200KUpdatedAug 17, 2026
ProviderCrofAIProvider model IDglm-4.7-flashRegionglobalInput / 1M$0.04Output / 1M$0.30Context202.8KUpdatedAug 17, 2026
ProviderDeep InfraProvider model IDzai-org/GLM-4.7-FlashRegionglobalInput / 1M$0.06Output / 1M$0.40Context202.8KUpdatedAug 17, 2026
ProviderOpenRouterProvider model IDz-ai/glm-4.7-flashRegionglobalInput / 1M$0.06Output / 1M$0.40Context202.8KUpdatedAug 17, 2026
ProviderEden AIProvider model IDdeepinfra/zai-org/GLM-4.7-FlashRegionglobalInput / 1M$0.06Output / 1M$0.40Context202.8KUpdatedAug 17, 2026
ProviderLLM GatewayProvider model IDglm-4.7-flashRegionglobalInput / 1M$0.06Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderKilo GatewayProvider model IDz-ai/glm-4.7-flashRegionglobalInput / 1M$0.06Output / 1M$0.40Context202.8KUpdatedAug 17, 2026
ProviderVenice AIProvider model IDzai-org-glm-4.7-flashRegionglobalInput / 1M$0.06Output / 1M$0.40Context128KUpdatedAug 17, 2026
ProviderCloudflare Workers AIProvider model ID@cf/zai-org/glm-4.7-flashRegionglobalInput / 1M$0.0605Output / 1M$0.40Context131.1KUpdatedAug 17, 2026
ProviderCloudflare AI GatewayProvider model IDworkers-ai/@cf/zai-org/glm-4.7-flashRegionglobalInput / 1M$0.0605Output / 1M$0.40Context131.1KUpdatedAug 17, 2026
ProviderEden AIProvider model IDcloudflare/@cf/zai-org/glm-4.7-flashRegionglobalInput / 1M$0.0605Output / 1M$0.40Context131.1KUpdatedAug 17, 2026
ProviderNanoGPTProvider model IDzai-org/glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderJiekou.AIProvider model IDzai-org/glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderAmazon BedrockProvider model IDzai.glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderMerge GatewayProvider model IDzai/glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderVercel AI GatewayProvider model IDzai/glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderEden AIProvider model IDamazon/zai.glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderNovitaAIProvider model IDzai-org/glm-4.7-flashRegionglobalInput / 1M$0.07Output / 1M$0.40Context200KUpdatedAug 17, 2026
ProviderCortecsProvider model IDglm-4.7-flashRegionglobalInput / 1M$0.08Output / 1M$0.478Context203KUpdatedAug 17, 2026
ProviderSyntheticProvider model IDhf:zai-org/GLM-4.7-FlashRegionglobalInput / 1M$0.10Output / 1M$0.50Context196.6KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

GLM 4.7 Flash Runtime Performance

Provider-specific output speed and catalog latency for GLM-4.7-Flash. Runtime does not affect the capability score.

1 rows
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderZAIOutput Speed50 tok/sCatalog Latency2 sMax Input128KMax Output16.4KUpdatedAug 17, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

GLM 4.7 Flash Specifications

Technical details for the model's default version.

Version
GLM-4.7-Flash
Released
Jan 19, 2026
Knowledge cutoff
Apr 1, 2025
Parameters
30B
Context window
200K
Max output
131.1K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

GLM 4.7 Flash Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionGLM-4.7-FlashReleasedJan 19, 2026LLMBoard37.9Parameters30BContext200KMax output131.1KOpen weightsYesLicenseMIT

GLM 4.7 Flash vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

GLM 4.7 FlashvsMercury 2GLM 4.7 FlashvsClaude Sonnet 4GLM 4.7 FlashvsGemini 2.5 FlashGLM 4.7 FlashvsClaude Sonnet 3.7GLM 4.7 FlashvsGemma 4 12BGLM 4.7 FlashvsNemotron 3.5 Lightning

Models similar to GLM 4.7 Flash

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#129+3.6
ZA

GLM 4.5 Air

Zhipu AI

41.5 LLMBoard

DetailsCompare
#103+11.2
ZA

GLM 4.5

Zhipu AI

49.1 LLMBoard

DetailsCompare
#85+15.7
ZA

GLM 4.6

Zhipu AI

53.6 LLMBoard

DetailsCompare
#78+18.3
ZA

GLM 5V Turbo

Zhipu AI

56.2 LLMBoard

DetailsCompare
#67+21.5
ZA

GLM 4.7

Zhipu AI

59.4 LLMBoard

DetailsCompare
#47+29.2
ZA

GLM 5

Zhipu AI

67.1 LLMBoard

DetailsCompare

What is GLM 4.7 Flash?

Key information about GLM 4.7 Flash and its available data.

7 optimized for fast inference and lower latency. 7 including thinking before acting, preserved reasoning across turns, and per-request thinking control for speed or accuracy trade-offs.

Ideal for applications requiring quick responses while maintaining strong performance on coding, agentic workflows, and general reasoning tasks.

Data as of 2026-08-17.

FAQ

Common questions about GLM 4.7 Flash.

When was GLM 4.7 Flash released?

GLM 4.7 Flash's default version was released on Jan 19, 2026.

How much does GLM 4.7 Flash cost?

No official standard PAYG price is currently available for GLM 4.7 Flash. The lowest tracked third-party offer starts at $0.04 input and $0.30 output via CrofAI.

Who created GLM 4.7 Flash?

GLM 4.7 Flash was created by Zhipu AI.

What is the context window for GLM 4.7 Flash?

The default version has a 200K token context window.

Is GLM 4.7 Flash open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer GLM 4.7 Flash?

24 provider offerings are linked to the default version.

What models should I compare GLM 4.7 Flash with?

Nearby ranked alternatives include Mercury 2, Claude Sonnet 4, Gemini 2.5 Flash.