AI roundup
:Gemini 3.7 Flash: Three-Week Iteration Brings Coding Gains and 50% Price Cut
Sat 15 August 2026
What Shipped
Google released Gemini 3.7 Flash on August 13, 2026, three weeks after the 3.6 Flash debut. The model targets coding and agentic workflows with introductory pricing set at $0.75 per million input tokens and $3.75 per million output tokens through December 31, 2026, doubling to $1.50 and $7.50 respectively on January 1, 2027. Google did not disclose parameter count.
Architecture and Limits
The model supports a 1 million token context window and a 64,000 token output limit. Context caching costs $0.075 per million tokens during the introductory period, rising to $0.15 in 2027. Safety metrics compared to 3.6 Flash show Text to Text Safety at +1.17pp, Multilingual Safety at -0.48pp, and Unjustified-refusals at +0.84pp, with Google noting low unjustified refusals overall.
Coding and Agentic Benchmarks
FrontierCode 1.1 Main scores rose from 34.4% to 43.6%, exceeding Claude Sonnet 5 at 42.7% and GPT-5.6 Terra at 41.3%. DeepSWE v1.1 improved from 48.6% to 65.3%, remaining below GPT-5.6 Terra's 69.6%. Terminal-bench 3.0 increased from 5.4% to 14.9%, matching Claude Sonnet 5 at 14.6% but trailing GPT-5.6 Terra's 20.8%. OSWorld-2.0 hit 47.9% versus 3.6 Flash's 33.8%, while AutomationBench climbed from 17.0% to 30.4%, surpassing Claude Sonnet 5 at 10.7% and GPT-5.6 Terra at 23.6%.
Document and Multimodal Performance
GDP.pdf document comprehension increased from 22.0% to 34.0%, beating Claude Sonnet 5 at 28.0% and GPT-5.6 Terra at 24.7%. GDM-MRCR v2 long-context retrieval reached 97.0% against 3.6 Flash's 91.8% and GPT-5.6 Terra's 93.5%. LVBench video understanding ticked up to 85.4% from 84.2%. CharXiv chart reasoning without tools dipped slightly from 85.2% to 84.5%, while the Artificial Analysis Intelligence Index scored 56, up from 3.6 Flash's 52 but below GPT-5.6 Terra's 57.
Availability and Positioning
The model is live in the Gemini API, AI Studio, and Gemini Enterprise. Consumer access is limited to the Gemini Spark agent within the Gemini app for AI Pro or Ultra subscribers; the standard chatbot interface continues running 3.6 Flash. Pricing remains above OpenAI's GPT-5.6 Luna at $0.20 per million input tokens and $1.20 per million output tokens. The release follows the delayed Gemini 3.5 Pro, which missed its June 2026 launch window.
Source: Ars Technica AI