Skip to content
Breaking:

Google Launches Gemini 3.8 Flash Model With Advanced Reasoning and Higher Token Output

The updated model retains previous per-token rates but can consume more tokens per task, outperforming rivals on key software engineering benchmarks.

By The Company Wire3 min read
Share
Google — Google Launches Gemini 3.8 Flash Model With Advanced Reasoning and Higher Token Output
Google — Google Launches Gemini 3.8 Flash Model With Advanced Reasoning and Higher Token Output. Photo: The Verge.

Google has officially launched Gemini 3.8 Flash, expanding its flagship artificial intelligence model suite just weeks after debuting its predecessor, Gemini 3.7 Flash. The technology giant stated that the updated model is designed to perform more intensive processing by carrying out additional reasoning steps when handling intricate prompts and executing tool calls in an iterative sequence. As first reported by The Verge, the model's release brings enhanced computational depth while presenting new economic dynamics for developers and corporate customers.

The baseline pricing structure for Gemini 3.8 Flash remains identical to the previous generation, set at $0.75 per million input tokens and $3.75 per million output tokens. However, Google warned that overall expenses for customers could still increase during practical deployment. The company noted that the model may consume a larger volume of tokens to maximize performance outcomes, particularly when configured to function at higher effort thresholds. To assist cost-conscious organizations, Google confirmed that developers seeking to strictly limit token consumption can continue building on Gemini 3.7 Flash.

Initial benchmarking from tracking firm Artificial Analysis highlighted the performance-to-cost profile of the new release, characterizing Gemini 3.8 Flash as the least expensive model measured at its benchmarked tier of intelligence. Despite the static per-token rates, Artificial Analysis reported that overall task execution costs rose by approximately 40 percent compared to Gemini 3.7 Flash. The evaluation firm attributed this expenditure rise to a 30 percent increase in generated output tokens per task alongside an expanded number of evaluation turns during agentic performance assessments.

Industry leaders evaluated the model's technical capabilities relative to existing market competitors. John Ennis, chief executive officer of Aigora.ai, stated that Gemini 3.8 Flash provides coding quality comparable to Anthropic’s Opus 5 model, while operating at high execution speeds and at a fraction of the expense. Ennis specifically noted that the model’s capabilities would prove highly advantageous for specialized media automation tasks, including the creation of remotion videos.

According to performance metrics provided by Google, Gemini 3.8 Flash delivers substantial enhancements for autonomous AI agents and core software engineering applications. The model achieved top placement on the DeepSWE v1.1 software engineering benchmark, surpassing Gemini 3.7 Flash as well as competing frontier systems. Notably, it outperformed Anthropic's Fable 5, an alternative model that received an upgrade earlier in the week aimed at lowering user expenses through discounted rates on cached data processing.

In addition to general programming metrics, Gemini 3.8 Flash demonstrated high performance on domain-focused evaluation frameworks, securing leading results on both the Vals Finance Agent V2 benchmark and Harvey's Legal Agent benchmark. Alongside performance updates, Google highlighted integrated risk mitigations within the system, noting that the model ships with dedicated protections designed to prevent misuse across cyber offensive techniques and Chemical, Biological, Radiological, and Nuclear (CBRN) sectors.

Coinciding with the primary model rollout, Google unveiled Gemini 3.8 Flash Cyber alongside the inauguration of its Fairwind Program. The initiative is restricted to government entities and designated institutional partners, launching with a cohort of 650 participating members that includes endpoint security vendor CrowdStrike and the Center for Internet Security. The program is tailored to support security operations across public sector assets and national infrastructure.

Participants within the Fairwind Program receive targeted access to Gemini 3.8 Flash Cyber as well as CodeMender, a specialized Google AI agent created to autonomously discover and remediate software vulnerabilities. Google stated that these tools are intended to strengthen critical infrastructure, public services, and overall national security defenses. For broader market availability, the standard Gemini 3.8 Flash model is available immediately to developers, enterprise organizations, and individual consumers holding Google AI Pro or Google AI Ultra subscriptions.

Sources

  1. The Verge

Company: Google

Written by

The Company Wire

Newsroom · San Francisco

Inside the companies building what’s next. Reporting on startups, technology, funding and the people shaping them.