DeepSeek V4 Flash 0731 (batch)

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows. This is the GA release of DeepSeek V4 Flash.

ChatDeepSeek1M tokens$0.1526 / $0.3052 · 1M

deepseek/deepseek-v4-flash-0731:batch

Context
1M tokens
Completion cap
943,718
Tools
Yes
JSON
Yes
Released
2026-07-31

Call it from Geek Hub

Same OpenAI SDK. Change the base URL and the model id.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek/deepseek-v4-flash-0731:batch",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Get an API key

Frequently asked questions

What is DeepSeek V4 Flash 0731 (batch)?
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows. This is the GA release of DeepSeek V4 Flash. DeepSeek V4 Flash 0731 (batch) runs on the Geek Hub API (OpenAI-compatible). Model id: deepseek/deepseek-v4-flash-0731:batch.
Is DeepSeek V4 Flash 0731 (batch) free?
No. Input is $0.1526 per 1M tokens and output is $0.3052 per 1M tokens on Geek Hub (markup included).
What is the context length of DeepSeek V4 Flash 0731 (batch)?
DeepSeek V4 Flash 0731 (batch) has a 1M tokens context window. It supports up to 943,718 completion tokens.
Does DeepSeek V4 Flash 0731 (batch) support tool calling and structured outputs?
DeepSeek V4 Flash 0731 (batch) accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.
When was DeepSeek V4 Flash 0731 (batch) released?
DeepSeek V4 Flash 0731 (batch) was released on 2026-07-31.

More models from DeepSeek