Deepseek v4 Flash API

Run deepseek/deepseek-v4-flash through FastInfra's OpenAI-compatible API. Pay per token, no subscription, routed to the cheapest available provider.

Read the docs
Pricing

Deepseek v4 Flash pricing

Billed per token used. Prices sync automatically from wholesale providers.

Direction Price per 1M tokens
Input$0.09
Output$0.19
Availability

Served by 4 providers

Requests route to DeepInfra by default (cheapest). Pin a specific provider with the :provider suffix.

Provider Upstream model ID
openrouter deepseek/deepseek-v4-flash
fireworks accounts/fireworks/models/deepseek-v4-flash
deepinfra deepseek-ai/DeepSeek-V4-Flash
siliconflow deepseek-ai/DeepSeek-V4-Flash
Quickstart

Call Deepseek v4 Flash in 30 seconds

Works with any OpenAI SDK — change the base URL and API key only.

Python

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

response = client.chat.completions.create(
    model="deepseek/deepseek-v4-flash",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

curl

curl https://api.fastinfra.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek/deepseek-v4-flash",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
FAQ

Deepseek v4 Flash — common questions

How much does the deepseek/deepseek-v4-flash API cost?

On FastInfra, deepseek/deepseek-v4-flash costs $0.0945/1M input, $0.189/1M output tokens. Billing is per token used, with no subscription.

Is deepseek/deepseek-v4-flash compatible with the OpenAI SDK?

Yes. FastInfra exposes deepseek/deepseek-v4-flash through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.

Which providers serve deepseek/deepseek-v4-flash?

deepseek/deepseek-v4-flash is available from 4 provider(s): openrouter, fireworks, deepinfra, siliconflow. Requests route to DeepInfra by default (cheapest); append ":provider" to the model ID to pin one.