DeepSeek

deepseek-v4-1-flash-260910

DeepSeekToken-basedDynamic Pricing
Create API Key

DeepSeek 的 deepseek-v4-1-flash-260910 模型,支持文本、图像输入和文本输出。支持推理。

textimagereasoningvisioncontext:1048576
Starting price
View full pricing
Context
1M
Maximum input window
Max output
393.2K
Maximum tokens per response
Modalities
→
Released
Sep 2026

Pricing by Supplier

official
官方接口直连
Time-based pricingBeijing time (UTC+8)

Peak hours

Monday–Friday · 09:00–12:00 / 14:00–18:00

Input$0.03/ 1M
Output$1.2/ 1M
Cache Read$0.006/ 1M

Off-peak hours

Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00

Saturday / Sunday · All day

Input$0.015/ 1M
Output$0.6/ 1M
Cache Read$0.003/ 1M

Time-based prices are determined at settlement. Tool fees are charged separately.

Volcengine
-8%
火山引擎官方接口
Time-based pricingBeijing time (UTC+8)

Peak hours

Monday–Friday · 09:00–12:00 / 14:00–18:00

Input$0.03$0.0276/ 1M
Output$1.2$1.104/ 1M
Cache Read$0.006$0.00552/ 1M

Off-peak hours

Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00

Saturday / Sunday · All day

Input$0.015$0.0138/ 1M
Output$0.6$0.552/ 1M
Cache Read$0.003$0.00276/ 1M

Time-based prices are determined at settlement. Tool fees are charged separately.

火山引擎特价
-55%
火山引擎官方接口,高并发,活动特价
Time-based pricingBeijing time (UTC+8)

Peak hours

Monday–Friday · 09:00–12:00 / 14:00–18:00

Input$0.03$0.0135/ 1M
Output$1.2$0.54/ 1M
Cache Read$0.006$0.0027/ 1M

Off-peak hours

Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00

Saturday / Sunday · All day

Input$0.015$0.00675/ 1M
Output$0.6$0.27/ 1M
Cache Read$0.003$0.00135/ 1M

Time-based prices are determined at settlement. Tool fees are charged separately.

Capabilities / Supported modalities

ReasoningVision
Input
Output

Provider & data privacy

Provider
DeepSeekDocs
Tokenizer
DeepSeek tokenizer (BPE)
License
DeepSeek LicenseOpen weights
Data retention79 daysNot used for upstream training by default

Performance

API access

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
official900362K18K
Volcengine360143K7.2K
火山引擎特价590236K12K

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply per token group.

API access questions

Everything you need to decide before creating an account.

Discover more models