ModelInfo
English
All models

GPT-5 Mini

Strong intelligence for cost sensitive, low latency, high volume workloads.

OpenAIgpt-5-mini2025-08-07Report incorrect data

Specs

Context
400K
Max output
128K
Input price
$0.25/ 1M tokens
Output price
$2/ 1M tokens
Released
2025-08-07

Context

Context window
400K
Max input
272K
Max output
128K

Pricing · USD / 1M tokens

Input
$0.25
Output
$2
Cache read
$0.025
Cache write
—

Modalities

Input
TextImage
Output
Text

API

API types
chatresponsesbatch

Reasoning

Reasoning
Yes
Levels
—

Info

Status
Active
Released
2025-08-07
Knowledge cutoff
2024-05-31

Features

Confirmed
toolsstructured_outputstreamingcachingbatch
Source · official docsVerified 2026-10-10

How to call

1 provider

openaimodel = gpt-5-minireasoning = reasoning.effort

chat
POSThttps://api.openai.com/v1/chat/completions
responses
POSThttps://api.openai.com/v1/responses
batch
POSThttps://api.openai.com/v1/batches

JSON

Standard format
gpt-5-mini.json
{
  "id": "gpt-5-mini",
  "object": "model",
  "created": 1754524800,
  "owned_by": "openai",
  "name": "GPT-5 Mini",
  "api": {
    "types": [
      "chat",
      "responses",
      "batch"
    ]
  },
  "limits": {
    "context": 400000,
    "input": 272000,
    "output": 128000
  },
  "modalities": {
    "input": [
      "text",
      "image"
    ],
    "output": [
      "text"
    ]
  },
  "reasoning": {
    "supported": true,
    "efforts": []
  },
  "pricing": {
    "currency": "USD",
    "unit": "1M_tokens",
    "input": 0.25,
    "output": 2,
    "cache_read": 0.025,
    "cache_write": null
  },
  "features": [
    "tools",
    "structured_output",
    "streaming",
    "caching",
    "batch"
  ],
  "info": {
    "status": "active",
    "release_date": "2025-08-07",
    "knowledge_cutoff": "2024-05-31",
    "description": "Strong intelligence for cost sensitive, low latency, high volume workloads.",
    "docs": "https://developers.openai.com/api/docs/models/gpt-5-mini",
    "verified_at": "2026-10-10"
  }
}

Official sources

4