ChengRang

Gemini 3.8 Flash

AI Chatbots Freemium

Google's new-generation model in the Flash series released on 2026/9/2 (the third Flash version within six weeks), focusing on long-horizon coding agents and professional workflows: DeepSWE v1.1 73.7%, Terminal-Bench 2.1 89.4%, Vals Finance Agent v2 61.4%; 1M context/64K output, intro price $0.75/$3.75 per million tokens; also launched Flash Cyber for defensive side at the same event

GoogleLarge ModelCodingAgentCost-Effectiveness
Visit Gemini 3.8 Flash

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

Gemini 3.8 Flash is Google's new-generation model in the Flash series, released on September 2, 2026, and is the third Flash version launched within six weeks. This model has no Preview suffix, and its stable API model name is gemini-3.8-flash. It is positioned as a "workhorse" model for software engineering, agentic tasks, and multi-step reasoning in professional domains. The release also includes a specialized version, Gemini 3.8 Flash Cyber, designed for cybersecurity defenders, which shares the same base model but is only available to trusted institutions through the Fairwind Program. Gemini 3.8 Flash offers a 1M context window and a maximum output of 64K tokens, supports text, image, video, audio, and PDF inputs, and stands out among similar models with an inference speed of approximately 300 tokens per second.

Key Features

Use Cases

Pros

Pricing

Until December 31, 2026, the introductory price is $0.75 per million input tokens and $3.75 per million output tokens, the same as 3.7 Flash, with the free tier available; starting January 1, 2027, the standard price doubles to $1.50 for input and $7.50 for output per million tokens. The official note indicates that for complex tasks, the model may take additional reasoning steps and repeatedly call tools, so actual billing may be higher than 3.7 Flash. Media tests show an average cost increase of about 40% per task (approximately $0.58 compared to $0.41), with an average of about 30% more output tokens. Please refer to the official website for final pricing.

Summary

Gemini 3.8 Flash is Google's latest general-purpose model in the Flash series, released in September 2026, focusing on long-horizon software engineering and professional agentic workflows, with significant improvements over 3.7 Flash on multiple benchmarks. Its 1M context, 64K output, and fast inference capabilities suit complex tasks, and it also offers a Cyber version for cybersecurity defenders. Pricing during the introductory period matches 3.7 Flash, but actual costs for complex tasks may be higher. Overall, this model excels in coding, agentic tasks, and professional domain reasoning, making it suitable for applications requiring high throughput and deep processing.

Version History

Category
AI Chatbots
Pricing
Freemium
Tags
Google · Large Model · Coding
Website

Related Tools