by DeepSeek
DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced, and benchmark tests far surpass V4-Pro-Preview.
DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced, and benchmark tests far surpass V4-Pro-Preview.
deepseek-v4-flash-0731 has a 1,000,000 token context window.
On AIHubMix, deepseek-v4-flash-0731 costs $0.15 per million input tokens and $0.31 per million output tokens. Cached input reads are billed at $0.0031 per million tokens.
deepseek-v4-flash-0731 accepts text input.
deepseek-v4-flash-0731 supports tool calling, function calling, structured outputs and thinking. Per-protocol parameter support is listed in the capability table on this page.
deepseek-v4-flash-0731 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-v4-flash-0731 — no other code changes needed.
deepseek-v4-flash-0731 is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.1 non-thinking mode has now been updated to the DeepSeek-V3.1-Terminus…
Thinking mode of DeepSeek-V3.1; DeepSeek V3.1 is a text generation model provided by…
Use deepseek-v4-flash-0731 via the AIHubMix unified API — one interface for every major LLM.