Skip to content
  • Models
  • Rankings
  • Ori
ElevenLabs launch offer: every ElevenLabs model is 50% off through October 19, 2026. See ElevenLabs models
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Ori
  • Collections
  • Providers
  • Tools
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status
  • AI Site Map

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
  1. Home
  2. /
  3. Compare

DeepSeek V4 Flash 0731 vs GLM 5.3 Flash

Compare DeepSeek V4 Flash 0731 from DeepSeek and GLM 5.3 Flash from Z.ai on key metrics including benchmarks, price, context length, and other model features. Access both models and hundreds of others through the OpenRouter API.

Favicon for deepseek
DeepSeek V4 Flash 0731
Favicon for z-ai
GLM 5.3 Flash
Favicon for deepseek
DeepSeek V4 Flash 0731
Favicon for z-ai
GLM 5.3 Flash

Overview

DeepSeek V4 Flash 0731
Author
DeepSeek
GLM 5.3 Flash
Author
Z.ai
DeepSeek V4 Flash 0731
Context length
1.05Mtokens
GLM 5.3 Flash
Context length
1.05Mtokens
DeepSeek V4 Flash 0731
Reasoning
GLM 5.3 Flash
Reasoning
DeepSeek V4 Flash 0731
Input modalities
GLM 5.3 Flash
Input modalities
DeepSeek V4 Flash 0731
Output modalities
GLM 5.3 Flash
Output modalities

DeepSeek V4 Flash 0731 vs GLM 5.3 Flash: side-by-side summary

DeepSeek V4 Flash 0731 and GLM 5.3 Flash are available through the OpenRouter API, so switching between them takes a model slug change rather than a new integration.

DeepSeek V4 Flash 0731, from DeepSeek, has a 1,048,576-token context window and is priced at $0.0047/M tokens input, $1.28/M tokens output on OpenRouter.

GLM 5.3 Flash, from Z.ai, has a 1,048,576-token context window and is priced at $0.04/M tokens input, $0.50/M tokens output on OpenRouter.

Context window
  • DeepSeek V4 Flash 0731: 1,048,576 tokens
  • GLM 5.3 Flash: 1,048,576 tokens
Price
  • DeepSeek V4 Flash 0731: $0.0047/M tokens input, $1.28/M tokens output
  • GLM 5.3 Flash: $0.04/M tokens input, $0.50/M tokens output
Provided by
  • DeepSeek V4 Flash 0731: DeepSeek
  • GLM 5.3 Flash: Z.ai
  • DeepSeek V4 Flash 0731 model details, providers and pricing
  • model details, providers and pricing
GLM 5.3 Flash

DeepSeek V4 Flash 0731 vs GLM 5.3 Flash at a glance

DeepSeek V4 Flash 0731 vs GLM 5.3 Flash comparison
AttributeDeepSeek V4 Flash 0731GLM 5.3 Flash
Input price$0.0047/M tokens$0.04/M tokens
Output price$1.28/M tokens$0.50/M tokens
Context window1,048,576 tokens1,048,576 tokens
Intelligence Index34.341.8
Coding Index69.171.5
Agentic Index41.050.9
Latency (p50)0.48 s1.50 s
Throughput (p50)66.0 tok/s43.0 tok/s

Benchmark data updated Oct 5, 2026

Frequently asked questions

Is DeepSeek V4 Flash 0731 cheaper than GLM 5.3 Flash?

DeepSeek V4 Flash 0731 and GLM 5.3 Flash trade off on price: DeepSeek V4 Flash 0731 costs $0.0047/M tokens input and $1.28/M tokens output, while GLM 5.3 Flash costs $0.04/M tokens input and $0.50/M tokens output, so DeepSeek V4 Flash 0731 is cheaper for input tokens and GLM 5.3 Flash is cheaper for output tokens.

Which has the longer context window, DeepSeek V4 Flash 0731 or GLM 5.3 Flash?

DeepSeek V4 Flash 0731 and GLM 5.3 Flash have the same context window of 1,048,576 tokens.

Which is faster, DeepSeek V4 Flash 0731 or GLM 5.3 Flash?

DeepSeek V4 Flash 0731 is faster on OpenRouter, generating a median 66.0 tokens per second compared with 43.0 tokens per second for GLM 5.3 Flash.

Which scores higher on coding, DeepSeek V4 Flash 0731 or GLM 5.3 Flash?

GLM 5.3 Flash scores higher on coding, with an Artificial Analysis Coding Index of 71.5 compared with 69.1 for DeepSeek V4 Flash 0731.