Google has introduced Gemini 3.8 Flash, a new version of its fast, developer-focused AI model aimed at delivering stronger performance on complex tasks. The company says the model “works harder” than Gemini 3.7 Flash by using more reasoning steps and calling tools iteratively when needed.
This is a meaningful upgrade for developers building AI assistants, coding tools, workflow automations, and other applications that benefit from better problem-solving. More deliberate reasoning and tool use can help models complete multi-step tasks more reliably, especially when connected to external systems or data sources.
More capability with flexible cost control
Google is keeping the same introductory pricing as Gemini 3.7 Flash: $0.75 per million input tokens and $3.75 per million output tokens. However, because Gemini 3.8 Flash may use more tokens at higher effort levels to improve performance, actual costs could rise for some workloads.
- Best for performance: Gemini 3.8 Flash offers stronger reasoning for demanding tasks.
- Best for efficiency: Developers can continue using Gemini 3.7 Flash to reduce token usage.
- Best for builders: The choice gives teams more control over the balance between quality and cost.
Overall, Gemini 3.8 Flash reflects the rapid pace of AI model improvement: more capable systems are becoming available to developers quickly, with practical controls for deploying them in real-world products.