Qwen-Audio-3.1-ASR-Flash-Filetrans
Offline speech recognition model for long audio such as meetings and calls, with speaker separation and dialect support.
Specs
- Context
- 8.2K
- Max output
- 1K
- Input price
- $0.15/ 1M tokens
- Output price
- $0.47/ 1M tokens
- Released
- —
Context
- Context window
- 8.2K
- Max input
- 8.2K
- Max output
- 1K
Pricing · USD / 1M tokens
- Input
- $0.15
- Output
- $0.47
- Cache read
- —
- Cache write
- —
Modalities
- Input
- Audio
- Output
- Text
API
- API types
Reasoning
- Reasoning
- No
Info
- Status
- Active
- Released
- —
- Knowledge cutoff
- —
Features
- Not published
How to call
1 provider
alibabamodel = qwen-audio-3.1-asr-flash-filetrans
JSON
Standard format
{
"id": "qwen-audio-3.1-asr-flash-filetrans",
"object": "model",
"created": null,
"owned_by": "alibaba",
"name": "Qwen-Audio-3.1-ASR-Flash-Filetrans",
"api": {
"types": [
"audio"
]
},
"limits": {
"context": 8192,
"input": 8192,
"output": 1024
},
"modalities": {
"input": [
"audio"
],
"output": [
"text"
]
},
"reasoning": {
"supported": false,
"efforts": []
},
"pricing": {
"currency": "USD",
"unit": "1M_tokens",
"input": 0.15,
"output": 0.47,
"cache_read": null,
"cache_write": null
},
"features": [],
"info": {
"status": "active",
"release_date": null,
"knowledge_cutoff": null,
"description": "Offline speech recognition model for long audio such as meetings and calls, with speaker separation and dialect support.",
"docs": "https://docs.modelstudio.console.alibabacloud.com/en/model-studio/qwen-audio-3-1-asr-flash-filetrans",
"verified_at": "2026-10-11"
}
}Official sources
2