Qwen-VL-OCR
Qwen3-VL-architecture model for extracting text and structured data from scans, tables and receipts.
Specs
- Context
- —
- Max output
- 8.2K
- Input price
- $0.07/ 1M tokens
- Output price
- $0.16/ 1M tokens
- Released
- —
Context
- Context window
- —
- Max input
- —
- Max output
- 8.2K
Pricing · USD / 1M tokens
- Input
- $0.07
- Output
- $0.16
- Cache read
- —
- Cache write
- —
Modalities
- Input
- TextImage
- Output
- Text
API
- API types
Reasoning
- Reasoning
- No
Info
- Status
- Retired
- Released
- —
- Knowledge cutoff
- —
Features
- Confirmed
How to call
1 provider
alibabamodel = qwen-vl-ocr
chat
POSThttps://{WorkspaceId}.ap-southeast-1.maas.aliyuncs.com/compatible-mode/v1/chat/completions
JSON
Standard format
{
"id": "qwen-vl-ocr",
"object": "model",
"created": null,
"owned_by": "alibaba",
"name": "Qwen-VL-OCR",
"api": {
"types": [
"chat"
]
},
"limits": {
"context": null,
"input": null,
"output": 8192
},
"modalities": {
"input": [
"text",
"image"
],
"output": [
"text"
]
},
"reasoning": {
"supported": false,
"efforts": []
},
"pricing": {
"currency": "USD",
"unit": "1M_tokens",
"input": 0.07,
"output": 0.16,
"cache_read": null,
"cache_write": null
},
"features": [
"streaming"
],
"info": {
"status": "retired",
"release_date": null,
"knowledge_cutoff": null,
"description": "Qwen3-VL-architecture model for extracting text and structured data from scans, tables and receipts.",
"docs": "https://www.alibabacloud.com/help/en/model-studio/vision-model",
"verified_at": "2026-10-11"
}
}Official sources
4
Vision models | Alibaba Cloud Model Studiowww.alibabacloud.com2026-10-10Text generation models | Alibaba Cloud Model Studiowww.alibabacloud.com2026-10-10Model Studio model pricing | Alibaba Cloudwww.alibabacloud.com2026-10-10[Model Studio] Notice of Retirement for Selected Legacy Long-tail Models | Alibaba Cloudwww.alibabacloud.com2026-10-11