deepseek-v4-flash API

DeepSeek deepseek-v4-flash for chat and coding; Flatkey routes this model through /v1/chat/completions. $0.176/$0.528 per 1M tokens 1,048,576-token context window.

Chat and codingLong contextTool workflows
Input /M
from $0.176/ 1M tokens-20%from $0.22/ 1M tokens
Cache /M
from $0.0056/ 1M tokens-20%from $0.007/ 1M tokens
Output /M
from $0.528/ 1M tokens-20%from $0.66/ 1M tokens
Provider
DeepSeek

deepseek-v4-flash API performance

deepseek-v4-flash API reliability and uptime

Live Flatkey request telemetry for deepseek-v4-flash appears here when enough production traffic is available.

Avg. provider uptime
99.0%
last 30 days
Latency
4.28s
last 30 days
Requests
1.45M
30-day window
Successful inference trend
#1 · 10.6T

deepseek-v4-flash API activity

deepseek-v4-flash API usage and request activity

Track request volume and successful inference activity for deepseek-v4-flash over the latest reporting window.

Requests
1.45M
30-day window
Latency
4.28s
last 30 days
Uptime
99.0%
last 30 days
Successful inference trendActivity

deepseek-v4-flash pricing

deepseek-v4-flash API pricing and billing

Prices below are calculated from Flatkey pricing data for this model and the visible groups currently returned by our pricing API.

Flatkey price

Input /M

Live catalog model
from $0.176/ 1M tokensfrom $0.22/ 1M tokens
Input /M
from $0.176/ 1M tokensfrom $0.22/ 1M tokens
Cache /M
from $0.0056/ 1M tokensfrom $0.007/ 1M tokens
Output /M
from $0.528/ 1M tokensfrom $0.66/ 1M tokens
View API
Shared balance

Add credits

Add credits

Use the same Flatkey balance and API key across image, video, audio, and text models.

Model catalog
Model Type
Text
API
/v1/chat/completions
Billing basis
1M tokens

deepseek-v4-flash chat and coding: core capabilities

DeepSeek's deepseek-v4-flash route supports text · file. Catalog categories: Programming, Roleplay, Marketing.

deepseek-v4-flash chat and coding API

Call deepseek-v4-flash through /v1/chat/completions for chat, coding or agent workflows supported by the model's listed modalities.

deepseek-v4-flash long-context work

deepseek-v4-flash is listed with 1,048,576-token; use that catalog value as an integration planning reference and validate limits for your account.

deepseek-v4-flash tool and agent workflows

Keep tool calls, structured output and streaming behavior aligned with the endpoint contract instead of assuming every provider feature is interchangeable.

deepseek-v4-flash production routing

Use one Flatkey key for deepseek-v4-flash, usage controls and model routing while retaining the exact model id in your SDK configuration.

Compare

deepseek-v4-flash API comparison: capabilities and access

Compare the catalog facts for deepseek-v4-flash with the previous-generation baseline before changing your integration.

CapabilityPrevious generationdeepseek-v4-flash
ProviderNot verified in this catalog snapshotDeepSeek
ModalitiesNot verified in this catalog snapshottext · file
ContextNot verified in this catalog snapshot1,048,576-token
EndpointNot verified in this catalog snapshot/v1/chat/completions
BillingNot verified in this catalog snapshot$0.176/$0.528 per 1M tokens

Why Flatkey

Why use Flatkey for deepseek-v4-flash?

OpenAI-compatible migration path

Chat Completions-style payloads reduce switching friction from existing model stacks.

Production routing

Keep usage, keys, quotas, and model routing in one Flatkey account.

Unified API access

Use one account and API key across image, video, audio, and language models.

Routing across upstream channels

Requests are spread across the channels serving this model, with 30-day uptime published above.

API

deepseek-v4-flash API

Four ways in, all on the same key and the same model catalog. Pick one to see a runnable example.

Call any model with an OpenAI-compatible API. Copy a ready-to-run example for your model and language.
Use the OpenAI SDK you already know — with Flatkey as the gateway.
Generate images and videos from your terminal. Let your AI assistant drive the workflow.
Connect your coding agent with one command, then use Flatkey from your existing projects.
curl https://router.flatkey.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer sk-***" \
  -d '{"model":"deepseek-v4-flash","messages":[{"role":"user","content":"Say hello in one sentence."}]}'
Docs

FAQ

deepseek-v4-flash API
pricing, features, and usage FAQs

Answers about deepseek-v4-flash pricing, capabilities, endpoint access and catalog limits.

What is deepseek-v4-flash used for?
deepseek-v4-flash is listed by DeepSeek for chat and coding; the catalog lists these modalities: text · file.
How is deepseek-v4-flash priced?
Use the pricing section above for current Flatkey prices from our pricing API.
Which API endpoint calls deepseek-v4-flash?
Flatkey routes this model through /v1/chat/completions. Set the model field to deepseek-v4-flash and follow the fields supported by that endpoint.
What context or input limits does deepseek-v4-flash have?
deepseek-v4-flash is listed with 1,048,576-token. Other limits depend on the route and current account availability, so verify them before production use.