HomeModelsGoogle Gemini 3.6 Flash (1M)

Google Gemini 3.6 Flash (1M)

Google 1M context Released: 2026-07

The fast, cheap tier of Gemini 3.6. Flash keeps a 1M token window and multimodal input while running several times faster than Pro, which makes it the default choice for high-volume production traffic.

Visit Model Website

Pricing Information

Input Pricing

Standard:$0.5000
Per 1,000 tokens
Cached:$0.0500
Per 1,000 tokens (cached requests)

Output Pricing

Standard:$3.0000
Per 1,000 tokens

Example Costs

Short Conversation
1K input + 500 output tokens
$2.0000
Book Analysis
50K input + 2K output tokens
$31.00

Token Calculator

Tokens:0
Words:0
Characters:0
Input Cost:$0.00
Estimated based on current token count as input
Use our main calculator for more detailed estimates including input/output combinations.

Key Features

  • 1M token context window
  • Very high throughput
  • Multimodal input (text, image, audio, video)
  • Context caching
  • Thinking budget control

Common Use Cases

  • High-volume APIs
  • Real-time assistants
  • Media pipelines
  • RAG front ends

Ratings & Feedback

0.0 / 5 · 0 votes

Comments

Frequently Asked Questions

What are Google's main AI model families as of late 2024?

What is a key differentiating feature of Google's Gemini 1.5 Pro and newer Gemini models?

How does Google price its Gemini models, especially those with very large context windows?

Are Google's Gemini models primarily focused on multimodality?