BaiLian

qwen3-tts-instruct-flash-realtime

Alibaba按量收費

千问支持指令控制的实时文本转语音模型,可通过文字指令调整语音风格。

textaudiostreaming
起始價格
輸入 / 輸出 · 1M
上下文
—
最大輸入窗口
模態
→

Pricing by Supplier

official
官方接口直连
輸入$75/ 1M
輸出$75/ 1M
Alibaba
阿里巴巴百炼官方
輸入$82.5/ 1M
輸出$82.5/ 1M

能力 / 支援的模態

串流輸出
輸入
輸出

廠商與數據私隱

供應商
OpenAI文件
分詞器
cl100k_baseOlder GPT-3.5 family
許可證
Proprietary (commercial)商業閉源
數據保留30 日預設不會用於上游訓練

效能

About qwen3-tts-instruct-flash-realtime

千问支持指令控制的实时文本转语音模型,可通过文字指令调整语音风格。

Use cases and prompting

Starting points for evaluation; supported inputs and options are listed in API access.

Use cases to explore

  • Evaluate speech or audio workflows using representative recordings or scripts.
  • Compare intelligibility, pronunciation and noise handling for your target language.

Practical tips

Check the endpoint direction first: audio input, transcription or audio output. Confirm supported languages, file formats and length limits before preparing a batch.

API access

呼叫示例

請求POST/v1/chat/completions
請求範例
參數
參數類型預設值 / 範圍說明資訊
temperature
number
=10 ~ 2
採樣溫度;越低越穩定
top_p
number
=10 ~ 1
核採樣累積概率
max_tokens
integer>= 1回應中最大 token 數
frequency_penalty
number
=0-2 ~ 2
懲罰高頻 token 的重複出現
presence_penalty
number
=0-2 ~ 2
鼓勵引入新話題
stop
array—最多 4 個停止生成的字串
seed
integer—盡量保證可復現的採樣種子
n
integer
=1>= 1
生成的候選條數
stream
boolean
=false
透過 SSE 串流返回 token
response_format
object—強制輸出 JSON 物件或符合 Schema 的結果
tools
array—模型可呼叫的工具 / 函數聲明
tool_choice
string
autononerequired
工具選擇策略或具體工具名
logprobs
boolean
=false
返回每個 token 的對數概率
top_logprobs
integer0 ~ 20每個 token 返回的 top 概率數量
logit_bias
object—按 token 的 logit 偏置映射
user
string—用於風險審計的終端用戶標識

替換 <YOUR_API_KEY> 替換為令牌設定中的 API Key。

身份驗證

所有請求必須攜帶 Authorization: Bearer <TOKEN> 請求頭。Anthropic 格式的端點也接受 x-api-key 請求頭。

在「令牌」頁面生成 API Key,可以按模型、分組、IP、速率等維度精細化授權。

支援的參數

Generation parameters
參數類型預設值 / 範圍說明資訊
temperature
number
=10 ~ 2
採樣溫度;越低越穩定
top_p
number
=10 ~ 1
核採樣累積概率
max_tokens
integer>= 1回應中最大 token 數
frequency_penalty
number
=0-2 ~ 2
懲罰高頻 token 的重複出現
presence_penalty
number
=0-2 ~ 2
鼓勵引入新話題
stop
array—最多 4 個停止生成的字串
seed
integer—盡量保證可復現的採樣種子
n
integer
=1>= 1
生成的候選條數
stream
boolean
=false
透過 SSE 串流返回 token
response_format
object—強制輸出 JSON 物件或符合 Schema 的結果
tools
array—模型可呼叫的工具 / 函數聲明
tool_choice
string
autononerequired
工具選擇策略或具體工具名
logprobs
boolean
=false
返回每個 token 的對數概率
top_logprobs
integer0 ~ 20每個 token 返回的 top 概率數量
logit_bias
object—按 token 的 logit 偏置映射
user
string—用於風險審計的終端用戶標識

速率限制

供應商RPMTPMRPD
Alibaba480193K9.6K
official350140K7.0K

RPM = 每分鐘請求數,TPM = 每分鐘 token 數,RPD = 每日請求數。限制按令牌分組生效。

Frequently asked questions about qwen3-tts-instruct-flash-realtime

What is qwen3-tts-instruct-flash-realtime?

千问支持指令控制的实时文本转语音模型,可通过文字指令调整语音风格。

How do I call qwen3-tts-instruct-flash-realtime?

Create an API key with access to qwen3-tts-instruct-flash-realtime, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen3-tts-instruct-flash-realtime priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate qwen3-tts-instruct-flash-realtime for my project?

Check the endpoint direction first: audio input, transcription or audio output. Confirm supported languages, file formats and length limits before preparing a batch.