GPT-5 Mini
Strong intelligence for cost sensitive, low latency, high volume workloads.
Specs
- Context
- 400K
- Max output
- 128K
- Input price
- $0.25/ 1M tokens
- Output price
- $2/ 1M tokens
- Released
- 2025-08-07
Context
- Context window
- 400K
- Max input
- 272K
- Max output
- 128K
Pricing · USD / 1M tokens
- Input
- $0.25
- Output
- $2
- Cache read
- $0.025
- Cache write
- —
Modalities
- Input
- TextImage
- Output
- Text
API
- API types
Reasoning
- Reasoning
- Yes
- Levels
- —
Info
- Status
- Active
- Released
- 2025-08-07
- Knowledge cutoff
- 2024-05-31
Features
- Confirmed
How to call
1 provider
openaimodel = gpt-5-minireasoning = reasoning.effort
chat
POSThttps://api.openai.com/v1/chat/completions
responses
POSThttps://api.openai.com/v1/responses
batch
POSThttps://api.openai.com/v1/batches
JSON
Standard format
{
"id": "gpt-5-mini",
"object": "model",
"created": 1754524800,
"owned_by": "openai",
"name": "GPT-5 Mini",
"api": {
"types": [
"chat",
"responses",
"batch"
]
},
"limits": {
"context": 400000,
"input": 272000,
"output": 128000
},
"modalities": {
"input": [
"text",
"image"
],
"output": [
"text"
]
},
"reasoning": {
"supported": true,
"efforts": []
},
"pricing": {
"currency": "USD",
"unit": "1M_tokens",
"input": 0.25,
"output": 2,
"cache_read": 0.025,
"cache_write": null
},
"features": [
"tools",
"structured_output",
"streaming",
"caching",
"batch"
],
"info": {
"status": "active",
"release_date": "2025-08-07",
"knowledge_cutoff": "2024-05-31",
"description": "Strong intelligence for cost sensitive, low latency, high volume workloads.",
"docs": "https://developers.openai.com/api/docs/models/gpt-5-mini",
"verified_at": "2026-10-10"
}
}Official sources
4