ModelRushText

DeepSeek V4 Flash 0731 API

DeepSeek V4 Flash 0731 generates text from chat messages through the ModelRush API and returns generated text. Requests run synchronously and can stream tokens. It is callable in 6 execution regions, from $0.4664 per 1M input tokens.
Hybrid reasoningText generationOpenAI-compatible Chat Completions

Input

sync-or-stream
Estimated run$0.4664
Sign in to generate

Output preview

This is a model capability preview. Sign in to submit a live request and inspect the generated result here.
token5s

Model information

Capabilities and limits

Best for

Fast reasoningChat applicationsBatch text processing

Current customer pricing

Each row identifies the Operation, Region, price variant, and immutable price version used for execution.
ModelCapabilityExecution RegionPrice variantCustomer price
DeepSeek V4 Flash 0731 · modelrush/deepseek-v4-flash-0731Chat completionscn-beijingStandardInput $0.4664 · Output $1.3992 / 1M tokens
DeepSeek V4 Flash 0731 · modelrush/deepseek-v4-flash-0731Chat completionscn-hong-kongStandardInput $0.4664 · Output $1.3992 / 1M tokens
DeepSeek V4 Flash 0731 · modelrush/deepseek-v4-flash-0731Chat completionseu-frankfurtStandardInput $0.4664 · Output $1.3992 / 1M tokens
DeepSeek V4 Flash 0731 · modelrush/deepseek-v4-flash-0731Chat completionsjp-tokyoStandardInput $0.4664 · Output $1.3992 / 1M tokens
DeepSeek V4 Flash 0731 · modelrush/deepseek-v4-flash-0731Chat completionssg-singaporeStandardInput $0.484 · Output $1.452 / 1M tokens
DeepSeek V4 Flash 0731 · modelrush/deepseek-v4-flash-0731Chat completionsus-virginiaStandardInput $0.4664 · Output $1.3992 / 1M tokens

Other recommended models

LLM

DeepSeek V4 Pro

Reasoning model · released in 4 regions
Starting pricePricing pending
LLM

GLM 5.1

Reasoning model · released in 5 regions
Starting pricePricing pending
LLM

GLM 5.2 Us

Reasoning model · released in 1 region
Starting pricePricing pending
ModelRushOne integration, intelligent routing, transparent billing. Model infrastructure for developers and agents.
© 2026 ModelRushAll systems operational