ModelInfo
English
All models

Gemini 1.5 Flash

Gemini 1.5 Flash is a fast and versatile multimodal model for scaling across diverse tasks.

Googlegemini-1.5-flash2024-05-24Report incorrect data

Specs

Context
1.05M
Max output
8.2K
Input price
$0.075/ 1M tokens
Output price
$0.3/ 1M tokens
Released
2024-05-24

Context

Context window
1.05M
Max input
1.05M
Max output
8.2K

Pricing · USD / 1M tokens

Input
$0.075
Output
$0.3
Cache read
$0.0187
Cache write
—

Modalities

Input
TextImageVideoAudio
Output
Text

API

API types
generate

Reasoning

Not published

Info

Status
Retired
Released
2024-05-24
Knowledge cutoff
—

Features

Confirmed
toolsstructured_outputcaching
Source · official docsVerified 2026-10-11

How to call

1 provider

googlemodel = gemini-1.5-flash

generate
POSThttps://generativelanguage.googleapis.com/v1beta/models/{model}:generateContent

JSON

Standard format
gemini-1.5-flash.json
{
  "id": "gemini-1.5-flash",
  "object": "model",
  "created": 1716508800,
  "owned_by": "google",
  "name": "Gemini 1.5 Flash",
  "api": {
    "types": [
      "generate"
    ]
  },
  "limits": {
    "context": 1048576,
    "input": 1048576,
    "output": 8192
  },
  "modalities": {
    "input": [
      "text",
      "image",
      "video",
      "audio"
    ],
    "output": [
      "text"
    ]
  },
  "reasoning": {
    "supported": null,
    "efforts": []
  },
  "pricing": {
    "currency": "USD",
    "unit": "1M_tokens",
    "input": 0.075,
    "output": 0.3,
    "cache_read": 0.01875,
    "cache_write": null
  },
  "features": [
    "tools",
    "structured_output",
    "caching"
  ],
  "info": {
    "status": "retired",
    "release_date": "2024-05-24",
    "knowledge_cutoff": null,
    "description": "Gemini 1.5 Flash is a fast and versatile multimodal model for scaling across diverse tasks.",
    "docs": "https://web.archive.org/web/20251112105521/https://cloud.google.com/vertex-ai/generative-ai/docs/learn/model-versions",
    "verified_at": "2026-10-11"
  }
}

Official sources

3