deepseek-v4-flash-0731

by DeepSeek

DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced, and benchmark tests far surpass V4-Pro-Preview.

API Pricing

Input$0.15 / 1M tokens
Output$0.31 / 1M tokens
Cache read$0.0031 / 1M tokens

Specifications

Context window1,000,000 tokens
Modalitiestext
Featurestool calling, function calling, structured outputs, thinking
Endpointschat_completions, claude_api

FAQ

What is deepseek-v4-flash-0731?

DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced, and benchmark tests far surpass V4-Pro-Preview.

What is the context length of deepseek-v4-flash-0731?

deepseek-v4-flash-0731 has a 1,000,000 token context window.

How much does deepseek-v4-flash-0731 cost?

On AIHubMix, deepseek-v4-flash-0731 costs $0.15 per million input tokens and $0.31 per million output tokens. Cached input reads are billed at $0.0031 per million tokens.

What modalities does deepseek-v4-flash-0731 support?

deepseek-v4-flash-0731 accepts text input.

What features does deepseek-v4-flash-0731 support?

deepseek-v4-flash-0731 supports tool calling, function calling, structured outputs and thinking. Per-protocol parameter support is listed in the capability table on this page.

How do I call deepseek-v4-flash-0731 via API?

deepseek-v4-flash-0731 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-v4-flash-0731 — no other code changes needed.

Who develops deepseek-v4-flash-0731?

deepseek-v4-flash-0731 is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from DeepSeek

deepseek-v4-flash

by DeepSeek

DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…

$0.15/1M in · $0.31/1M out
1,000,000 tokens context

deepseek-v4-pro

by DeepSeek

DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…

$0.46/1M in · $0.93/1M out
1,000,000 tokens context

deepseek-v3.2

by DeepSeek

DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…

$0.3/1M in · $0.45/1M out
128,000 tokens context

deepseek-v3.2-think

by DeepSeek

DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…

$0.3/1M in · $0.45/1M out
128,000 tokens context

DeepSeek-V3.1-Terminus

by DeepSeek

DeepSeek-V3.1 non-thinking mode has now been updated to the DeepSeek-V3.1-Terminus…

$0.56/1M in · $1.68/1M out
160,000 tokens context

DeepSeek-V3.1-Think

by DeepSeek

Thinking mode of DeepSeek-V3.1; DeepSeek V3.1 is a text generation model provided by…

$0.56/1M in · $1.68/1M out
128,000 tokens context

Use deepseek-v4-flash-0731 via the AIHubMix unified API — one interface for every major LLM.