Gemini 3.7 Flash Is Now on Multos — The Fastest Intelligent Model Yet
Google's sharpest Flash yet launched today. 65.3% on DeepSWE, 340 tokens/second, and it's live on Multos Starter right now.
Google launched Gemini 3.7 Flash on August 13, 2026 — 23 days after Gemini 3.6 Flash — and it is the strongest Flash-tier model Google has ever shipped. Every benchmark moved: 65.3% on DeepSWE v1.1 (up from 49.0%), 30.4% on AutomationBench (up from 17.0%), and Artificial Analysis ranks it first out of 186 models on output speed at 340.1 tokens per second. It is live on Multos Starter right now.
The Numbers That Matter
FrontierCode 1.1 Main: 43.6% vs 34.4% for 3.6 Flash. DeepSWE v1.1: 65.3% vs 49.0% — a 16-point jump on long-horizon software engineering, the strongest gain any flash-tier model has posted on that test. WebDev Arena Elo: 1588 vs 1538. GDP.pdf (complex document comprehension): 34.0% vs 22.0%. AutomationBench (multi-step enterprise workflows): 30.4% vs 17.0%, nearly doubling. Independent scoring from Artificial Analysis: Intelligence Index 56 vs 52, placing it ahead of Claude Sonnet 5 (55) and just behind GPT-5.6 Terra (57).
Speed: First of 186 Models
Artificial Analysis ranked Gemini 3.7 Flash first out of 186 models on output speed at 340.1 tokens per second — compared to 210 t/s for 3.6 Flash. For streaming responses and long agentic runs you will feel the difference immediately. The blended price per million tokens is $0.58 at current introductory rates, half what 3.6 Flash cost at its own launch.
Spec Sheet
Model ID: gemini-3.7-flash. Context window: 1,048,576 tokens (1M). Output limit: 65,536 tokens (64k). Knowledge cutoff: March 2026. Input modalities: text, image, video, audio, PDF. Thinking levels: low, medium (default), high. Input price: $0.75 per 1M tokens introductory, rising to $1.50 on January 1 2027. Output price: $3.75 per 1M tokens introductory, rising to $7.50. Context caching: $0.075 per 1M tokens plus $0.50 per 1M tokens per hour storage.
Where It Sits on Multos
Gemini 3.7 Flash is a Starter tier model — available to every Starter, Pro, and Elite subscriber. It sits alongside Gemini 3.5 Flash-Lite and Gemini 3 Flash Preview at the standard tier. Pro and Elite subscribers also have access to Gemini 3.5 Flash, 3.6 Flash, and 3.1 Pro. Auto mode does not route to it by default yet while we complete internal evals. Pick it manually from the Google group in the model dropdown to use it today.
One Pricing Note
Google's introductory rate expires December 31, 2026. From January 1, 2027 the price doubles to $1.50 input / $7.50 output — exactly what 3.6 Flash cost at its launch. Until then both models are priced identically, so there is no reason to stay on 3.6 Flash. If you are choosing on cost alone, GPT-5.6 Luna ($0.20/$1.20) and DeepSeek V4 Flash ($0.14/$0.28) are cheaper — both available on Multos Lite.
Ready to Get Started?
The fastest model Artificial Analysis has ever benchmarked, with the strongest long-horizon coding score at the flash tier. Try Gemini 3.7 Flash on Multos — available on all Starter plans and above.
Start Building FreeRelated Articles
GPT-5.6 Luna Price Cut 80% — Now Available on Free & Lite
OpenAI slashed GPT-5.6 Luna by 80% overnight — from $1/$6 to $0.20/$1.20 per million tokens. We're passing it straight through: Luna is now available on every plan including Free and Lite.
Read more →Claude Opus 5 & Sonnet 5 Now Available: Near-Frontier Intelligence at Every Tier
Anthropic's Claude Opus 5 and Claude Sonnet 5 are now live on Multos. Opus 5 delivers near-Fable-5 intelligence at half the price with 79.2% on SWE-bench Pro, while Sonnet 5 brings Opus-class agentic capabilities to Starter-tier users at $3/$15 per million tokens.
Read more →Kimi K3 Now Available: The World's Largest Open-Weight Model on Multos
Moonshot AI's 2.8-trillion-parameter Kimi K3 is now live on Multos. A frontier-class open model with a 1M token context window, ranked #1 on Frontend Code Arena and #4 overall on Artificial Analysis Intelligence Index.
Read more →