Ink Whisper
Cartesia affordable speech-to-text model with higher accuracy and lower latency than baseline Whisper.
Specs
- Context
- —
- Max output
- —
- Input price
- —
- Output price
- —
- Released
- 2025-06-04
Context
- Not published
Pricing
- Not published
Modalities
- Input
- Audio
- Output
- Text
API
- API types
Reasoning
- Not published
Info
- Status
- Active
- Released
- 2025-06-04
- Knowledge cutoff
- —
Features
- Not published
How to call
1 provider
cartesiamodel = ink-whisper
audio
POSThttps://api.cartesia.ai/stt
realtime
GEThttps://api.cartesia.ai/stt/websocket
JSON
Standard format
{
"id": "ink-whisper",
"object": "model",
"created": 1748995200,
"owned_by": "cartesia",
"name": "Ink Whisper",
"api": {
"types": [
"audio",
"realtime"
]
},
"limits": {
"context": null,
"input": null,
"output": null
},
"modalities": {
"input": [
"audio"
],
"output": [
"text"
]
},
"reasoning": {
"supported": null,
"efforts": []
},
"pricing": null,
"features": [],
"info": {
"status": "active",
"release_date": "2025-06-04",
"knowledge_cutoff": null,
"description": "Cartesia affordable speech-to-text model with higher accuracy and lower latency than baseline Whisper.",
"docs": "https://docs.cartesia.ai/build-with-cartesia/stt/older-models",
"verified_at": "2026-10-11"
}
}Official sources
1