Qwen3-VL-Embedding
Multimodal embedding model for mixed text, image and video retrieval, with output dimensions from 256 to 2560.
Specs
- Context
- 32K
- Max output
- —
- Input price
- —
- Output price
- —
- Released
- 2026-01-21
Context
- Context window
- 32K
- Max input
- 32K
- Max output
- —
Pricing
- Not published
Modalities
- Input
- TextImageVideo
- Output
- Embedding
API
- API types
Reasoning
- Reasoning
- No
Info
- Status
- Active
- Released
- 2026-01-21
- Knowledge cutoff
- —
Features
- Not published
How to call
1 provider
alibabamodel = qwen3-vl-embedding
JSON
Standard format
{
"id": "qwen3-vl-embedding",
"object": "model",
"created": 1768953600,
"owned_by": "alibaba",
"name": "Qwen3-VL-Embedding",
"api": {
"types": [
"embeddings"
]
},
"limits": {
"context": 32000,
"input": 32000,
"output": null
},
"modalities": {
"input": [
"text",
"image",
"video"
],
"output": [
"embedding"
]
},
"reasoning": {
"supported": false,
"efforts": []
},
"pricing": null,
"features": [],
"info": {
"status": "active",
"release_date": "2026-01-21",
"knowledge_cutoff": null,
"description": "Multimodal embedding model for mixed text, image and video retrieval, with output dimensions from 256 to 2560.",
"docs": "https://help.aliyun.com/zh/model-studio/embedding-rerank-model",
"verified_at": "2026-10-11"
}
}Official sources
3