Skip to content
Catalog

Google

Gemini 3.7 Flash

Speed-first Gemini Flash

ReportedAugust 13, 2026 (reported)
Main purpose
High tokens-per-second chat and cheap inference
Context window
1M tokens (reported)
API pricing
$0.75 input / $3.75 output per 1M tokens introductory, doubling January 1, 2027 (reported)
Access
Gemini app, Google AI Studio, Vertex AI

Capabilities

Key features

  • Very high output speed (reported)
  • 1M-token context (reported)
  • Introductory Flash pricing

Notes

  • Artificial Analysis summaries in mid-August 2026 listed Gemini 3.7 Flash first on output speed (~340 tokens/sec). Reported, not a Google quote.

Limitations

  • Introductory price is reported to rise in 2027.

Speed, price, and date from August 2026 trackers (Artificial Analysis / FelloAI).

Vendor source