All Models
DeepSeek
DeepSeek

DeepSeek V4 Flash 0731

Agent-tuned revision of V4 Flash built for tool use, multi-step reasoning, and long workflows at exceptionally low cost.

reasoningfastefficientagentic
Provider
DeepSeek
Median cost per request
$0.018
Input price
$0.06 / 1M tokens
Output price
$0.12 / 1M tokens
Strengths
reasoning, fast, efficient, agentic
Modalities
text
Context window
1.0M tokens
Max output
66k tokens
Actions/tools
Supported
Popularity
Bottom 63% (75 of 118)

Build with DeepSeek V4 Flash 0731

Launch a project in a few clicks

1. Open the builder with DeepSeek V4 Flash 0731 preselected.
2. Pick a template or paste your prompt.
3. Ship to web, API, or embed.
CostDistribution

Price distribution shows typical request cost

Quartiles reveal what most runs cost; spikes usually mean long prompts or outputs

  • Median = typical cost per 100 requests for this model
  • Wide spread means costs vary a lot by prompt size or tools
  • Use this to estimate pricing before scaling usage

Price Distribution

Typical price per 100 requests across recent runs

Compare
SpeedDistribution

TTFT distribution shows how fast responses start

Lower time-to-first-token means snappier experiences

  • Median TTFT = typical wait before tokens start
  • Wide spread means response speed varies by prompt complexity
  • Aim for lower TTFT on interactive user flows

TTFT Distribution

Provider time-to-first-token distribution

Compare

Real builder experiences

Was this model actually good in Pickaxe?

Share wins, failures, and cost/performance tradeoffs with other builders. The more real-world runs, the better the guidance.

Related Models

DeepSeek V4 Flash 0423

A fast, 1M-context open model built for reasoning, tool use, and everyday AI workflows, offering strong performance at exceptionally low cost.

DeepSeek V4 Pro 0813

GA release of DeepSeek V4 Pro, built for advanced reasoning, coding, tool use, and long-context agent workflows.

DeepSeek V4.1 Flash

DeepSeek's fast, cost-efficient multimodal model for reasoning, coding, image understanding, tool use, and long-context agent workflows.