Skip to main content
Product2026-08-135 min read

Gemini 3.7 Flash Is Now on Multos — The Fastest Intelligent Model Yet

Google's sharpest Flash yet launched today. 65.3% on DeepSWE, 340 tokens/second, and it's live on Multos Starter right now.

O
Oghenekaro Oboido

Google launched Gemini 3.7 Flash on August 13, 2026 — 23 days after Gemini 3.6 Flash — and it is the strongest Flash-tier model Google has ever shipped. Every benchmark moved: 65.3% on DeepSWE v1.1 (up from 49.0%), 30.4% on AutomationBench (up from 17.0%), and Artificial Analysis ranks it first out of 186 models on output speed at 340.1 tokens per second. It is live on Multos Starter right now.

The Numbers That Matter

FrontierCode 1.1 Main: 43.6% vs 34.4% for 3.6 Flash. DeepSWE v1.1: 65.3% vs 49.0% — a 16-point jump on long-horizon software engineering, the strongest gain any flash-tier model has posted on that test. WebDev Arena Elo: 1588 vs 1538. GDP.pdf (complex document comprehension): 34.0% vs 22.0%. AutomationBench (multi-step enterprise workflows): 30.4% vs 17.0%, nearly doubling. Independent scoring from Artificial Analysis: Intelligence Index 56 vs 52, placing it ahead of Claude Sonnet 5 (55) and just behind GPT-5.6 Terra (57).

Speed: First of 186 Models

Artificial Analysis ranked Gemini 3.7 Flash first out of 186 models on output speed at 340.1 tokens per second — compared to 210 t/s for 3.6 Flash. For streaming responses and long agentic runs you will feel the difference immediately. The blended price per million tokens is $0.58 at current introductory rates, half what 3.6 Flash cost at its own launch.

Spec Sheet

Model ID: gemini-3.7-flash. Context window: 1,048,576 tokens (1M). Output limit: 65,536 tokens (64k). Knowledge cutoff: March 2026. Input modalities: text, image, video, audio, PDF. Thinking levels: low, medium (default), high. Input price: $0.75 per 1M tokens introductory, rising to $1.50 on January 1 2027. Output price: $3.75 per 1M tokens introductory, rising to $7.50. Context caching: $0.075 per 1M tokens plus $0.50 per 1M tokens per hour storage.

Where It Sits on Multos

Gemini 3.7 Flash is a Starter tier model — available to every Starter, Pro, and Elite subscriber. It sits alongside Gemini 3.5 Flash-Lite and Gemini 3 Flash Preview at the standard tier. Pro and Elite subscribers also have access to Gemini 3.5 Flash, 3.6 Flash, and 3.1 Pro. Auto mode does not route to it by default yet while we complete internal evals. Pick it manually from the Google group in the model dropdown to use it today.

One Pricing Note

Google's introductory rate expires December 31, 2026. From January 1, 2027 the price doubles to $1.50 input / $7.50 output — exactly what 3.6 Flash cost at its launch. Until then both models are priced identically, so there is no reason to stay on 3.6 Flash. If you are choosing on cost alone, GPT-5.6 Luna ($0.20/$1.20) and DeepSeek V4 Flash ($0.14/$0.28) are cheaper — both available on Multos Lite.

Ready to Get Started?

The fastest model Artificial Analysis has ever benchmarked, with the strongest long-horizon coding score at the flash tier. Try Gemini 3.7 Flash on Multos — available on all Starter plans and above.

Start Building Free