---
title: "DeepSeek: DeepSeek V4 Flash 0731 — unrated, priced in this catalogue · Undominated.ai"
canonical: https://undominated.ai/models/deepseek__deepseek-v4-flash-0731/
description: "DeepSeek: DeepSeek V4 Flash 0731: $0.060/M in, $0.180/M out. No independent quality score in this catalogue — unrated, not ranked, not scored zero."
---

# DeepSeek: DeepSeek V4 Flash 0731 — unrated, priced in this catalogue · Undominated.ai

> DeepSeek: DeepSeek V4 Flash 0731: $0.060/M in, $0.180/M out. No independent quality score in this catalogue — unrated, not ranked, not scored zero.

[Check base rates for this workload](/check/deepseek__deepseek-v4-flash-0731/) [Cheaper alternatives](/alternatives/) [Family](/families/deepseek/) [Compare models](https://undominated.ai/compare/?models=deepseek%2Fdeepseek-v4-flash-0731) Comparison context

Choose up to three standard models in Compare.

Requirements and prompt length apply in Compare. This page keeps its stated price basis and workload controls.

MODEL PROOF

# DeepSeek V4 Flash 0731

DeepSeek · 31 Jul 2026

 Balanced Summarise Chat Code gen Agentic

NOT INDEPENDENTLY RATED

## No independent LMArena score is published for this model.

 **$0.090** / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

Evidence as of 06 Oct 2026

 [openrouter/models](https://openrouter.ai/api/v1/models) Independent LMArena score Not independently rated
 Context window 1.05M
 Maximum output 943.72K
 Input modalities text
 Output modalities text
 Published input price $0.060 / 1M tokens
 Published output price $0.180 / 1M tokens
 Pricing kind fixed

Inspect [complete billing conditions](#billing-details) and endpoint terms below.

Price from DeepInfra: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (fp8), at its standard rate. The maker’s own precision is not known.

Cheaper right now: $0.066/M at StreamLake — promotion 90% It does not pass the like-for-like test, so it never sets a rank.

 Every price dimension · USD per million tokens

| Input | $0.060 /M |
| --- | --- |
| Output | $0.180 /M |
| Cached input | $0.015 /M |
| Cache write | Unknown /M |
| Cache write 1h | Unknown /M |
| Reasoning | Unknown /M |

### Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.060/M from DeepInfra.

 26 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

| Seller | Input | Output | Precision | Uptime (1d) |
| --- | --- | --- | --- | --- |
| Seller Relace relace/fp4 Endpoint terms Cached input /M $0.010 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.96% | Input $0.010 | Output $1.28 | Precision fp4 | Uptime (1d) 99.93% |
| Seller OpenInference open-inference/fp4 Endpoint terms Cached input /M $0.011 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.75% | Input $0.011 | Output $1.41 | Precision fp4 | Uptime (1d) 96.84% |
| Seller Sail Research sail-research/fp4 Endpoint terms Cached input /M $0.014 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.41% | Input $0.019 | Output $0.300 | Precision fp4 | Uptime (1d) 96.24% |
| Seller Sail Research sail-research/us Endpoint terms Cached input /M $0.012 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.53% | Input $0.019 | Output $0.420 | Precision fp4 | Uptime (1d) 94.46% |
| Seller Reka reka Endpoint terms Cached input /M $0.0056 Context limit 262,144 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.71% | Input $0.021 | Output $0.528 | Precision undeclared | Uptime (1d) 99.84% |
| Seller StreamLake streamlake/fp8 Endpoint terms Cached input /M $0.0014 Context limit 1,024,000 tokens Output limit 384,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount 90% · already included in these rates Uptime · last 30 minutes 99.77% | Input $0.044 | Output $0.132 | Precision fp8 | Uptime (1d) 99.77% |
| Seller Inceptron inceptron/fp4 Endpoint terms Cached input /M $0.027 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.50% | Input $0.050 | Output $0.650 | Precision fp4 | Uptime (1d) 99.87% |
| Seller DeepInfra deepinfra/fp8 Endpoint terms Cached input /M $0.015 Context limit 1,048,576 tokens Output limit 384,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.060 | Output $0.180 | Precision fp8 | Uptime (1d) 99.52% |
| Seller DigitalOcean digitalocean Endpoint terms Cached input /M $0.024 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.96% | Input $0.119 | Output $0.238 | Precision undeclared | Uptime (1d) 99.95% |
| Seller BaseTen baseten/fp8 Endpoint terms Cached input /M $0.028 Context limit 1,048,576 tokens Output limit 384,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.97% | Input $0.130 | Output $0.260 | Precision fp8 | Uptime (1d) 99.82% |
| Seller BaseTen baseten/fp8 Endpoint terms Cached input /M $0.028 Context limit 1,048,576 tokens Output limit 384,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.130 | Output $0.260 | Precision fp8 | Uptime (1d) 99.83% |
| Seller CoreWeave coreweave/fp8 Endpoint terms Cached input /M $0.070 Context limit 262,144 tokens Output limit 235,929 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.130 | Output $0.280 | Precision fp8 | Uptime (1d) 99.63% |
| Seller Parasail parasail/fp8 Endpoint terms Cached input /M $0.050 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.59% | Input $0.140 | Output $0.280 | Precision fp8 | Uptime (1d) 99.52% |
| Seller Cohere cohere Endpoint terms Cached input /M $0.070 Context limit 1,048,576 tokens Output limit 384,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.89% | Input $0.140 | Output $0.280 | Precision undeclared | Uptime (1d) 98.46% |
| Seller Together together Endpoint terms Cached input /M $0.030 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.39% | Input $0.140 | Output $0.280 | Precision undeclared | Uptime (1d) 99.61% |
| Seller Venice venice Endpoint terms Cached input /M $0.035 Context limit 1,000,000 tokens Output limit 32,768 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.61% | Input $0.175 | Output $0.350 | Precision undeclared | Uptime (1d) 97.99% |
| Seller Mancer 2 mancer/fp8 Endpoint terms Cached input /M Unknown Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.38% | Input $0.200 | Output $0.600 | Precision fp8 | Uptime (1d) 97.45% |
| Seller SiliconFlow siliconflow/fp8 Endpoint terms Cached input /M $0.028 Context limit 1,048,576 tokens Output limit 393,216 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.64% | Input $0.220 | Output $0.660 | Precision fp8 | Uptime (1d) 99.49% |
| Seller Wafer wafer/fast Endpoint terms Cached input /M $0.130 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.96% | Input $0.220 | Output $0.840 | Precision undeclared | Uptime (1d) 99.94% |
| Seller GMICloud gmicloud/fp8 Endpoint terms Cached input /M $0.0091 Context limit 1,048,575 tokens Output limit 943,717 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount 35% · already included in these rates Uptime · last 30 minutes 100.00% | Input $0.286 | Output $0.858 | Precision fp8 | Uptime (1d) 99.99% |
| Seller Phala phala Endpoint terms Cached input /M $0.020 Context limit 1,048,576 tokens Output limit 393,216 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount 30% · already included in these rates Uptime · last 30 minutes 100.00% | Input $0.308 | Output $0.924 | Precision undeclared | Uptime (1d) 99.81% |
| Seller Alibaba alibaba Endpoint terms Cached input /M $0.035 Context limit 1,000,000 tokens Output limit 393,216 tokens Tools Listed by endpoint Reasoning Listed by endpoint Time windows Standard-window rate shown. Discount 14:00–24:00 UTC every day: $0.176 in · $0.528 out. Promotional discount None reported Uptime · last 30 minutes 99.34% | Input $0.352 | Output $1.06 | Precision undeclared | Uptime (1d) 99.08% |
| Seller Novita novita/fp8 Endpoint terms Cached input /M $0.026 Context limit 1,048,576 tokens Output limit 393,216 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount 7% · already included in these rates Uptime · last 30 minutes 100.00% | Input $0.409 | Output $1.23 | Precision fp8 | Uptime (1d) 99.99% |
| Seller Baidu baidu/fp8 Endpoint terms Cached input /M $0.014 Context limit 1,048,576 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.97% | Input $0.440 | Output $1.32 | Precision fp8 | Uptime (1d) 99.94% |
| Seller AtlasCloud atlas-cloud/fp4 Endpoint terms Cached input /M $0.028 Context limit 1,048,576 tokens Output limit 393,216 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.440 | Output $1.32 | Precision fp4 | Uptime (1d) 99.86% |
| Seller Cloudflare cloudflare Endpoint terms Cached input /M $0.014 Context limit 1,048,576 tokens Output limit 943,718 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.440 | Output $1.32 | Precision undeclared | Uptime (1d) 99.98% |

### Independent scores

No independent benchmark has measured this model. Unrated is not a score of zero.

#### Best rankings by task

 - SVG #30 1186
- 3D #39 1216
- Websites #40 1245
- UI components #41 1242
- Code categories #42 1234
- Game development #43 1221
- Data visualisation #57 1188
- ASCII art #62 1100

Per-task Elo and rank from [Design Arena](https://designarena.ai), via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.

### Capability

| Context window | 1M |
| --- | --- |
| Max output | 943K |
| Input modes | text |
| Tool use | yes |
| Reasoning | optional |
| Open weights | Unknown |

### Provenance

| Price source | [openrouter.ai](https://openrouter.ai/api/v1/models) |
| --- | --- |
| Fetched | 2026-10-06 |
| Quality data | unrated |

## Our take

 editorial — not a measurement

This dated Flash identity must be checked separately from other DeepSeek variants. Use its own accepted output ceiling, prices and proof when evaluating large text-generation jobs. Do not transfer a peak/off-peak tariff, weight release or benchmark result from another DeepSeek model.

### Strengths

 - Text input and tool use are listed
- The record separates the context window from the output ceiling

### Weaknesses

 - A family name does not establish the same pricing schedule
- Verify the exact weight release and licence before planning self-hosting

### Reach for it when

 - Large text-generation evaluations with explicit output limits
- Comparisons that retain the dated model identity

### Avoid it if

 - You need multimodal input
- The plan depends on a tariff or weight release from another variant

Sources: [api-docs.deepseek.com](https://api-docs.deepseek.com/quick_start/pricing)

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

## Questions this page answers

 What does DeepSeek V4 Flash 0731 cost?

$0.060/M in, $0.180/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

 Does DeepSeek V4 Flash 0731 have an independent quality score?

DeepSeek V4 Flash 0731 has no independent quality score in this catalogue. Unrated is not a score of zero.

 What beats DeepSeek V4 Flash 0731?

DeepSeek V4 Flash 0731 has no independent quality score. Unrated is not a score of zero.

 Does this page use Artificial Analysis scores for DeepSeek V4 Flash 0731?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

 Does a missing score mean DeepSeek V4 Flash 0731 scored zero?

Unrated is not a score of zero.

 Is the cheapest DeepSeek V4 Flash 0731 endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.060/M from DeepInfra.

MODEL MONUMENT

## Save or share this proof

 Download proof card · 1200 × 630 Preview and more sharing options

DeepSeek: DeepSeek V4 Flash 0731. Balanced workload. Not independently rated. Evidence as of 06 Oct 2026.

 Copy model link Download square card · 1080 × 1080 Copy landscape image Share landscape card

## Related

No independent quality score, so there is no dominator to name. Unrated is not a clean bill of health.

 - Same vendor [DeepSeek V4.1 Flash](/models/deepseek__deepseek-v4.1-flash/)
- DeepSeek models [DeepSeek](/providers/deepseek/)
- Where this model appears [wire](/wire/)
- Same vendor [DeepSeek V4 Pro 0423](/models/deepseek__deepseek-v4-pro/)
- Where this model appears [spreads](/spreads/)
- Same vendor [DeepSeek V4 Flash 0423](/models/deepseek__deepseek-v4-flash/)
- Where this model appears [traps](/traps/)

## Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

 Markdown Copy
 [![DeepSeek V4 Flash 0731 — Undominated.ai dominance verdict](https://undominated.ai/badge/deepseek__deepseek-v4-flash-0731.svg)](https://undominated.ai/models/deepseek__deepseek-v4-flash-0731/)
 HTML Copy
 <a href="https://undominated.ai/models/deepseek__deepseek-v4-flash-0731/"><img src="https://undominated.ai/badge/deepseek__deepseek-v4-flash-0731.svg" alt="DeepSeek V4 Flash 0731 — Undominated.ai dominance verdict" height="36"></a>

Direct file: [https://undominated.ai/badge/deepseek__deepseek-v4-flash-0731.svg](https://undominated.ai/badge/deepseek__deepseek-v4-flash-0731.svg) — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

## Continue your investigation

 - [Explore model families](/families/)
- [Compare a shortlist](/compare/)
- [Check a model](/check/)
- [Understand the evidence](/methodology/)
