XiaomiMiMo

mimo-v2.5-tts-voicedesign

Xiaomi按量收費

无需参考音频,一句话描述即可生成全新音色。多维度语义理解,覆盖口音、年龄、气质与录音质感。

起始價格
輸入 / 輸出 · 1M
上下文
—
最大輸入窗口
模態
→
發佈於
Apr 2026

Pricing by Supplier

XiaomiMIMO
XiaomiMIMO 官方
可用

能力 / 支援的模態

函數呼叫工具JSON 模式結構化輸出
輸入
輸出

廠商與數據私隱

供應商
OpenAI文件
分詞器
cl100k_baseOlder GPT-3.5 family
許可證
Proprietary (commercial)商業閉源
數據保留30 日預設不會用於上游訓練

效能

About mimo-v2.5-tts-voicedesign

无需参考音频,一句话描述即可生成全新音色。多维度语义理解,覆盖口音、年龄、气质与录音质感。

Use cases and prompting

Starting points for evaluation; supported inputs and options are listed in API access.

Use cases to explore

  • Evaluate speech or audio workflows using representative recordings or scripts.
  • Compare intelligibility, pronunciation and noise handling for your target language.

Practical tips

Check the endpoint direction first: audio input, transcription or audio output. Confirm supported languages, file formats and length limits before preparing a batch.

API access

呼叫示例

請求POST/v1/chat/completions
請求範例
參數
參數類型預設值 / 範圍說明資訊
temperature
number
=10 ~ 2
採樣溫度;越低越穩定
top_p
number
=10 ~ 1
核採樣累積概率
max_tokens
integer>= 1回應中最大 token 數
frequency_penalty
number
=0-2 ~ 2
懲罰高頻 token 的重複出現
presence_penalty
number
=0-2 ~ 2
鼓勵引入新話題
stop
array—最多 4 個停止生成的字串
seed
integer—盡量保證可復現的採樣種子
n
integer
=1>= 1
生成的候選條數
stream
boolean
=false
透過 SSE 串流返回 token
response_format
object—強制輸出 JSON 物件或符合 Schema 的結果
tools
array—模型可呼叫的工具 / 函數聲明
tool_choice
string
autononerequired
工具選擇策略或具體工具名
logprobs
boolean
=false
返回每個 token 的對數概率
top_logprobs
integer0 ~ 20每個 token 返回的 top 概率數量
logit_bias
object—按 token 的 logit 偏置映射
user
string—用於風險審計的終端用戶標識

替換 <YOUR_API_KEY> 替換為令牌設定中的 API Key。

身份驗證

所有請求必須攜帶 Authorization: Bearer <TOKEN> 請求頭。Anthropic 格式的端點也接受 x-api-key 請求頭。

在「令牌」頁面生成 API Key,可以按模型、分組、IP、速率等維度精細化授權。

支援的參數

Generation parameters
參數類型預設值 / 範圍說明資訊
temperature
number
=10 ~ 2
採樣溫度;越低越穩定
top_p
number
=10 ~ 1
核採樣累積概率
max_tokens
integer>= 1回應中最大 token 數
frequency_penalty
number
=0-2 ~ 2
懲罰高頻 token 的重複出現
presence_penalty
number
=0-2 ~ 2
鼓勵引入新話題
stop
array—最多 4 個停止生成的字串
seed
integer—盡量保證可復現的採樣種子
n
integer
=1>= 1
生成的候選條數
stream
boolean
=false
透過 SSE 串流返回 token
response_format
object—強制輸出 JSON 物件或符合 Schema 的結果
tools
array—模型可呼叫的工具 / 函數聲明
tool_choice
string
autononerequired
工具選擇策略或具體工具名
logprobs
boolean
=false
返回每個 token 的對數概率
top_logprobs
integer0 ~ 20每個 token 返回的 top 概率數量
logit_bias
object—按 token 的 logit 偏置映射
user
string—用於風險審計的終端用戶標識

速率限制

供應商RPMTPMRPD
XiaomiMIMO400161K8.1K

RPM = 每分鐘請求數,TPM = 每分鐘 token 數,RPD = 每日請求數。限制按令牌分組生效。

Frequently asked questions about mimo-v2.5-tts-voicedesign

What is mimo-v2.5-tts-voicedesign?

无需参考音频,一句话描述即可生成全新音色。多维度语义理解,覆盖口音、年龄、气质与录音质感。

How do I call mimo-v2.5-tts-voicedesign?

Create an API key with access to mimo-v2.5-tts-voicedesign, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is mimo-v2.5-tts-voicedesign priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate mimo-v2.5-tts-voicedesign for my project?

Check the endpoint direction first: audio input, transcription or audio output. Confirm supported languages, file formats and length limits before preparing a batch.