Models
Enterprise
Subscribe
Resource
Documentation
Console

deepseek/deepseek-v4.1-flash

DeepSeek.Color

DeepSeek: deepseek-v4.1-flash

Modetext → text
Input price
Output price
Context length
Weekly usage0 tokens
Listed at

DeepSeek

Latency
Throughput
Upload rate
Context length
Max output
Input priceper 1M tokens
Output priceper 1M tokens
Cache readper 1M tokens
Cache writeper 1M tokens

Performance

Avg TPS
Avg latency
Avg success rate
TPS
TTFT
Latency
Success rate

Success rate

Speed

Code Example

curl ${API_BASE_URL}/v1/chat/completions \
  -H "Authorization: Bearer <YOUR_API_KEY>" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4.1-flash",
    "messages": [
      {
        "role": "user",
        "content": "Hello"
      }
    ]
  }'

Replace <YOUR_API_KEY> with the API key generated on your token management page.

Authentication

All requests must include the Authorization: Bearer <TOKEN> header.

Endpoint Path

POST
openai

Supported Parameters

ParameterTypeDefaultDescription

Rate Limits

GroupRPMTPMRPD

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.

Pricing

ProviderBilling typePrice summary
DeepSeek0%
Per call
Input priceper 1M tokens
Completion priceper 1M tokens
Cache read priceper 1M tokens
Cache write priceper 1M tokens
DeepSeek0%
Per call
Input priceper 1M tokens
Completion priceper 1M tokens
Cache read priceper 1M tokens
Cache write priceper 1M tokens