DeepSeek

deepseek-v4-1-flash-260910

DeepSeekToken-basedDynamic Pricing
Create API Key

DeepSeek 的 deepseek-v4-1-flash-260910 模型,支持文本、图像输入和文本输出。支持推理。

textimagereasoningvisioncontext:1048576
Starting price
View full pricing
Context
1M
Maximum input window
Max output
393.2K
Maximum tokens per response
Modalities
→
Released
Sep 2026

Pricing by Supplier

official
官方接口直连
Time-based pricingBeijing time (UTC+8)

Peak hours

Monday–Friday · 09:00–12:00 / 14:00–18:00

Input$0.03/ 1M
Output$1.2/ 1M
Cache Read$0.006/ 1M

Off-peak hours

Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00

Saturday / Sunday · All day

Input$0.015/ 1M
Output$0.6/ 1M
Cache Read$0.003/ 1M

Time-based prices are determined at settlement. Tool fees are charged separately.

Volcengine
-2%
火山引擎官方接口
Time-based pricingBeijing time (UTC+8)

Peak hours

Monday–Friday · 09:00–12:00 / 14:00–18:00

Input$0.03$0.0294/ 1M
Output$1.2$1.176/ 1M
Cache Read$0.006$0.00588/ 1M

Off-peak hours

Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00

Saturday / Sunday · All day

Input$0.015$0.0147/ 1M
Output$0.6$0.588/ 1M
Cache Read$0.003$0.00294/ 1M

Time-based prices are determined at settlement. Tool fees are charged separately.

火山引擎特价
-50%
火山引擎官方接口,高并发,活动特价
Time-based pricingBeijing time (UTC+8)

Peak hours

Monday–Friday · 09:00–12:00 / 14:00–18:00

Input$0.03$0.015/ 1M
Output$1.2$0.6/ 1M
Cache Read$0.006$0.003/ 1M

Off-peak hours

Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00

Saturday / Sunday · All day

Input$0.015$0.0075/ 1M
Output$0.6$0.3/ 1M
Cache Read$0.003$0.0015/ 1M

Time-based prices are determined at settlement. Tool fees are charged separately.

Capabilities / Supported modalities

ReasoningVision
Input
Output

Provider & data privacy

Provider
DeepSeekDocs
Tokenizer
DeepSeek tokenizer (BPE)
License
DeepSeek LicenseOpen weights
Data retention79 daysNot used for upstream training by default

Performance

About deepseek-v4-1-flash-260910

DeepSeek 的 deepseek-v4-1-flash-260910 模型,支持文本、图像输入和文本输出。支持推理。

Use cases and prompting

Starting points for evaluation; supported inputs and options are listed in API access.

Use cases to explore

  • Draft and revise text with explicit audience, tone and format requirements.
  • Summarize supplied documents and compare answers against the original sources.

Practical tips

Separate instructions from source material, describe the desired output and provide an example. Test tool use or structured output only when listed as supported.

Example prompt

Summarize the following document in five bullet points. Separate confirmed facts from open questions, cite the relevant passages and do not invent missing information. Document: [paste your text]

API access

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
official900362K18K
Volcengine360143K7.2K
火山引擎特价590236K12K

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply per token group.

Frequently asked questions about deepseek-v4-1-flash-260910

What is deepseek-v4-1-flash-260910?

DeepSeek 的 deepseek-v4-1-flash-260910 模型,支持文本、图像输入和文本输出。支持推理。

How do I call deepseek-v4-1-flash-260910?

Create an API key with access to deepseek-v4-1-flash-260910, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is deepseek-v4-1-flash-260910 priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

What is the context window of deepseek-v4-1-flash-260910?

The model catalog lists a context window of 1048576 tokens. Check the selected endpoint for request limits.

What is the maximum output of deepseek-v4-1-flash-260910?

The model catalog lists a maximum output of 393216 tokens. Your request settings may set a lower limit.

How should I evaluate deepseek-v4-1-flash-260910 for my project?

Separate instructions from source material, describe the desired output and provide an example. Test tool use or structured output only when listed as supported.