DeepSeek

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

ReasoningTool callingStreamingJSON mode
1.31M token context Released Jul 31, 2026

Intelligence overview

0
Intelligence
0
Coding
0
Agentic

1.31M tok

Context window

943.7K tok

Max output

Capabilities

What DeepSeek: DeepSeek V4 Flash 0731 can do.

Reasoning

Extended reasoning

Works through problems step by step before answering, trading latency for accuracy on harder tasks.

Adjustable effort

Choose how much the model deliberates per request — trade speed for depth as the task demands.

Development

Tool calling

Calls external functions mid-response, then continues reasoning from the result.

Structured output

Returns responses constrained to a JSON schema, ready to parse without cleanup.

Streaming

Streams tokens as they're generated instead of waiting on the full response.

Performance

Intelligence34.5
Artificial Analysis
Coding69.1
Artificial Analysis
Agentic41.7
Artificial Analysis

Design Arena — ranked #24 across 8 categories

Pricing

Pay for what you use.

Every model in UltraGPT runs behind one flat subscription — this is the underlying per-token cost the catalog tracks.

INPUT

$0.06 / 1M tok

OUTPUT

$0.12 / 1M tok

Cache read$0.012 / 1M tok
Use cases

Built for serious work.

Build

Generate and refactor code, with the ability to run and verify tool calls and return structured, parseable output.

Think

Works through hard problems with adjustable reasoning effort — max, high, low.

Research

Holds up to 1.31M tokens of source material in context without losing the thread.

Automate

Chains tool calls across multi-step workflows, reasoning between each one.

Preview

Ask DeepSeek: DeepSeek V4 Flash 0731 anything.

A preview of the real UltraGPT composer — the model you see here is the model you get.

DeepSeek: DeepSeek V4 Flash 0731 Ready

Continues in the UltraGPT app — one login, every model.

Comparison

How DeepSeek: DeepSeek V4 Flash 0731 stacks up.

The closest models in the catalog by provider, capability and score — real entries, pulled from the same directory.

Specification

Technical details.

Specification

Context window
1.31M tokens
Max output
943.7K tokens
Input modalities
Text
Output modalities
Text
Release date
Jul 31, 2026
Reasoning effort
max, high, low
Streaming
Supported
Tool calling
Supported
JSON mode
Supported

Model identity

Provider
DeepSeek
Provider ID
deepseek

Model ID

Reasoning

Effort: max, high, low
Directory

More from DeepSeek, and everything else.

This model is one entry in the full UltraGPT catalog — browse the rest, compare pricing and context windows, or filter by capability.

DeepSeek: DeepSeek V4 Flash 0731, one subscription away.

UltraGPT puts this model — and every other one in the directory — behind a single login and one flat monthly price.