DeepSeek V4 Flash API model profile
Evidence snapshot:
DeepSeek V4 Flash is a DeepSeek model profile with a 1,048,576-token context window. The dated catalog snapshot records text modalities and interfaces for coding, prompt caching, reasoning, reasoning controls, structured output.
DeepSeek V4 Flash is available under the Kendr alias kc-deepseek-v4-flash. Current account availability and customer credit quotes come from Kendr's live public model API and applicable account policy.
DeepSeek V4 Flash accepts text and returns text. The snapshot records coding, prompt caching, reasoning, reasoning controls, structured output, text generation, tool calling, and tools as capabilities or interfaces.
The dated reference snapshot lists $0.14 input and $0.28 output per million tokens. Cached input is $0.028 per million tokens. The separately published provider-route reference is Provider-specific (DeepSeek $0.14 / $0.28; QwenCloud $0.20 / $0.40 ($0.04/M cached)).
The 2026-08-12 snapshot includes 8 Design Arena categories. Publisher-reported launch results are labeled and are not presented as Kendr measurements; scores from different suites are not treated as interchangeable.
DeepSeek V4 Flash ranked #3 in the cited trailing-7-day third-party catalog usage dataset as of 2026-08-12. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share.
Profile data was reviewed for the 2026-08-12 snapshot. Claims retain their source dates, and unavailable fields are shown as unavailable instead of estimated.
Profile facts
- Profile type
- Kendr API model
- Kendr API alias
- kc-deepseek-v4-flash
- Model developer
- DeepSeek
- Kendr route provider
- Provider-routed DeepSeek
- Context window
- 1M
- Snapshot date
- 2026-08-12
- Reference catalog ID
- deepseek/deepseek-v4-flash
Context, modalities, and identity
DeepSeek V4 Flash accepts text and returns text. The snapshot records coding, prompt caching, reasoning, reasoning controls, structured output, text generation, tool calling, and tools as capabilities or interfaces.
- Model developer
- DeepSeek
- Kendr route provider
- Provider-routed DeepSeek
- Context window
- 1M
- Maximum output
- 393,216 tokens
- Input modalities
- text
- Output modalities
- text
- Knowledge cutoff
- Not disclosed
Capabilities and supported controls
Capabilities and parameters are reported from the dated catalog and source records; their presence does not guarantee identical behavior across every provider route.
- Supported parameters
- frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_a, top_k, top_logprobs, top_p
- coding
- prompt caching
- reasoning
- reasoning controls
- structured output
- text generation
- tool calling
- tools
Dated pricing snapshot
The dated reference snapshot lists $0.14 input and $0.28 output per million tokens. Cached input is $0.028 per million tokens. The separately published provider-route reference is Provider-specific (DeepSeek $0.14 / $0.28; QwenCloud $0.20 / $0.40 ($0.04/M cached)).
Reference figures are dated 2026-08-12; the live Kendr quote can differ by selected provider route, context tier, caching, tools, region, and current rate card.
- Price date
- 2026-08-12
- Input
- $0.14 per 1M tokens
- Cached input
- $0.028 per 1M tokens
- Output
- $0.28 per 1M tokens
Benchmark evidence and limitations
The 2026-08-12 snapshot includes 8 Design Arena categories. Publisher-reported launch results are labeled and are not presented as Kendr measurements; scores from different suites are not treated as interchangeable.
Third-party catalog fields normalized from the cited model dataset; benchmark methodology and coverage differ by source.
| Design Arena category | Elo | Win rate | Rank |
|---|---|---|---|
| 3d | 1242 | 49.3% | #37 |
| asciiart | 1148 | 42.8% | #45 |
| codecategories | 1231 | 48.9% | #39 |
| dataviz | 1151 | 40.5% | #72 |
| gamedev | 1237 | 50.2% | #35 |
| svg | 1199 | 48.4% | #28 |
| uicomponent | 1197 | 44.7% | #52 |
| website | 1228 | 49.1% | #41 |
Popularity and market context
DeepSeek V4 Flash ranked #3 in the cited trailing-7-day third-party catalog usage dataset as of 2026-08-12. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share.
- Third-party catalog rank
- #3
- Global market share
- Not inferred
- Observation date
- 2026-08-12
OpenAI-compatible API example
This example applies to the published Kendr alias on this hosted profile. Check the live public catalog before use.
curl https://api.kendr.org/v1/chat/completions \
-H "Authorization: Bearer $KENDR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"kc-deepseek-v4-flash","messages":[{"role":"user","content":"Hello"}]}'Frequently asked questions
What is DeepSeek V4 Flash?
DeepSeek V4 Flash is a DeepSeek model profile with a 1,048,576-token context window. The dated catalog snapshot records text modalities and interfaces for coding, prompt caching, reasoning, reasoning controls, structured output.
What context window does DeepSeek V4 Flash have?
The dated profile lists 1M of context and up to 393,216 output tokens. Provider routes, variants, and runtime configuration can impose lower effective limits.
What benchmark evidence is available for DeepSeek V4 Flash?
The 2026-08-12 snapshot includes 8 Design Arena categories. Publisher-reported launch results are labeled and are not presented as Kendr measurements; scores from different suites are not treated as interchangeable. Third-party catalog fields normalized from the cited model dataset; benchmark methodology and coverage differ by source.
What is the Kendr API alias for DeepSeek V4 Flash?
Use kc-deepseek-v4-flash as the model value. The live public model API and signed-in catalog remain authoritative for route availability and customer credit quotes.
How is DeepSeek V4 Flash priced on Kendr?
The dated reference snapshot lists $0.14 input and $0.28 output per million tokens. Cached input is $0.028 per million tokens. The separately published provider-route reference is Provider-specific (DeepSeek $0.14 / $0.28; QwenCloud $0.20 / $0.40 ($0.04/M cached)). Kendr applies one 5% markup to configured provider model cost; the live quote and settled routing receipt are authoritative for a request.
Sources and evidence dates
- DeepSeek V4 Flash third-party catalog record (catalog, checked 2026-08-12)
- Third-party model catalog methodology (methodology, checked 2026-08-12)
- DeepSeek official model documentation (primary, checked 2026-08-12)
- Design Arena leaderboard and methodology (benchmark, checked 2026-08-12)
- DeepSeek V4 Flash provider or route reference (primary, checked 2026-08-12)