DeepSeek V4 Flash Latest
Decision snapshot from the global rankings; composite scoring matches the on-site model board.
Data updated:
Model
1 verified sources
Aggregator quote only — no official sync yet
Next steps on AI Hippo
About this model
DeepSeek V4 Flash Latest ranks #8 on our global model leaderboard. It is listed as a LLM with 1.0M ctx and a typical snapshot price around $0.14 per 1M tokens. Use Compare for side-by-side picks or the Token hub for verified multi-provider pricing from the same snapshot.
You can also explore more models from DeepSeek , and browse more options from 🇨🇳 China .
Capabilities & specs
- Modality
- text->text
- Input modalities
- text
- Supported API parameters
- frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_a, top_k, top_logprobs, top_p
Catalog description: This model always redirects to the latest model in the DeepSeek V4 Flash family.
Token pricing by provider
Compare per-provider token prices for this model across available platforms.
| Provider | Input / 1M tokens | Output / 1M tokens | Latency | Status |
|---|---|---|---|---|
| DeepSeek | $0.09 | $0.18 | — | Snapshot · 2026-08-02 |
| OpenRouter | $0.09 | $0.18 | — | Verified · 2026-08-02 |
Provider prices are sourced from the token comparison dataset and may change between snapshots.
Alternative picks
-
Meituan: LongCat 2.0
Compare now -
DeepSeek: DeepSeek V4 Flash 0731
Compare now -
Poolside: Laguna S 2.1
Compare now
Pick one or two more models on global rankings and use Compare to view them side by side.
FAQ
Why is the official price missing?
Official rows require a verified direct vendor or cloud API price. When only aggregators list a model, the Official column shows — and verified aggregator quotes appear in the token table.
How is context different from max output tokens?
Context window is how much input the model can accept in a single request. Max output tokens is the per-response generation cap from the primary listing—often much smaller than the context window.
Where do multi-provider prices come from?
Prices are crawled or synced from token providers (OpenRouter, Groq, Together, and others) and merged into the infrastructure comparison dataset. Verified rows show a fetch date in the status column.
Data source & methodology
- Source
- OpenRouter catalog, verified vendor crawls, and Hugging Face hub snapshots.
- Metrics
- Context window, blended 1M-token price, max output tokens, and capability fields from the OpenRouter models API.
- Update cadence
- Daily pipeline refresh; provider prices may change between snapshots.
- Fetched at
- 2026-08-02
- Method
- Composite rank uses on-site context × price weighting (not LMSYS Arena ELO).