BaiLian

qwen3-vl-embedding

Alibaba按量收費

千问多模态向量模型,支持文本、图像和视频输入,用于跨模态检索与相似度计算。

textimagevideoembeddingscontext:32000
起始價格
輸入 / 輸出 · 1M
上下文
32K
最大輸入窗口
模態
→

Pricing by Supplier

official
官方接口直连
輸入$75/ 1M
輸出$75/ 1M
Alibaba
阿里巴巴百炼官方
輸入$75/ 1M
輸出$75/ 1M

能力 / 支援的模態

嵌入
輸入
輸出

廠商與數據私隱

供應商
Alibaba (Qwen)文件
分詞器
Qwen tokenizer (tiktoken-compat)
許可證
Tongyi Qianwen License開放權重
數據保留86 日預設不會用於上游訓練

效能

About qwen3-vl-embedding

千问多模态向量模型,支持文本、图像和视频输入,用于跨模态检索与相似度计算。

Use cases and prompting

Starting points for evaluation; supported inputs and options are listed in API access.

Use cases to explore

  • Evaluate semantic search with representative queries and documents.
  • Compare retrieval quality on your own knowledge base before connecting a downstream assistant.

Practical tips

Check whether the endpoint returns embeddings or reranks documents. Keep indexing and query preprocessing consistent.

API access

呼叫示例

請求POST/v1/chat/completions
請求範例
參數
參數類型預設值 / 範圍說明資訊
input必填
string—需要向量化的文字或文字陣列
dimensions
integer>= 1將向量截斷到指定維度
encoding_format
enum
=float
向量傳輸的編碼格式
user
string—用於風險審計的終端用戶標識

替換 <YOUR_API_KEY> 替換為令牌設定中的 API Key。

身份驗證

所有請求必須攜帶 Authorization: Bearer <TOKEN> 請求頭。Anthropic 格式的端點也接受 x-api-key 請求頭。

在「令牌」頁面生成 API Key,可以按模型、分組、IP、速率等維度精細化授權。

支援的參數

Generation parameters
參數類型預設值 / 範圍說明資訊
input必填
string—需要向量化的文字或文字陣列
dimensions
integer>= 1將向量截斷到指定維度
encoding_format
enum
=float
向量傳輸的編碼格式
user
string—用於風險審計的終端用戶標識

速率限制

供應商RPMTPMRPD
Alibaba7.8K1.6M157K
official9.7K1.9M195K

RPM = 每分鐘請求數,TPM = 每分鐘 token 數,RPD = 每日請求數。限制按令牌分組生效。

Frequently asked questions about qwen3-vl-embedding

What is qwen3-vl-embedding?

千问多模态向量模型,支持文本、图像和视频输入,用于跨模态检索与相似度计算。

How do I call qwen3-vl-embedding?

Create an API key with access to qwen3-vl-embedding, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen3-vl-embedding priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

What is the context window of qwen3-vl-embedding?

The model catalog lists a context window of 32000 tokens. Check the selected endpoint for request limits.

How should I evaluate qwen3-vl-embedding for my project?

Check whether the endpoint returns embeddings or reranks documents. Keep indexing and query preprocessing consistent.