Output tokens / second
Artificial Analysis release snapshot. Gemini 3.8 Flash was the fastest model in its tracked set at publication.
Google's September 2 workhorse pairs a 59 Artificial Analysis Intelligence score with 304.6 output tokens per second, 1M multimodal context, and frontier-adjacent coding. Gemini 3.8 Flash is now available in SPIRITT Workspaces for agentic apps and workflows.
The SPIRITT availability update, official release, independent speed and intelligence evidence, launch reactions, and a useful warning about benchmark optimism.
Gemini 3.8 Flash is out at $0.75 per million input tokens and is available to all SPIRITT users with full agentic power. One interesting use case: video to UI.
— Tamir (@TamirSPIRITT) September 2, 2026
Introducing Gemini 3.8, our best reasoning and coding model yet. By leveraging long-running agentic loops, Google is building on 3.7 Flash with Gemini 3.8 Flash and Gemini 3.8 Flash Cyber.
— Google (@Google) September 2, 2026
3.8 Flash is the same price as 3.7, is about the same speed, and is available through the Gemini API, AI Studio, Antigravity, Gemini App, and more.
— Logan Kilpatrick (@OfficialLoganK) September 2, 2026
Google has released Gemini 3.8 Flash. With high reasoning, it scores 59 on the Artificial Analysis Intelligence Index and reaches the Intelligence versus Cost per Task Pareto frontier.
— Artificial Analysis (@ArtificialAnlys) September 2, 2026
Gemini 3.8 Flash is rolling out in Gemini, Google AI Studio, and APIs, with introductory pricing of $0.75 input and $3.75 output per million tokens.
— TestingCatalog News (@testingcatalog) September 2, 2026
Gemini 3.8 Flash is officially out at $0.75 input and $3.75 output per million tokens, with an 89.4% Terminal-Bench 2.1 result.
— Pankaj Kumar (@pankajkumar_dev) September 2, 2026
Gemini 3.8 Flash generally looks very good, but the Terminal-Bench 4.0 dip is a reason to stay skeptical. The only way to know is to try it out.
— Scott (@scottstts) September 2, 2026
Gemini 3.8 Flash is live, and the launch numbers are unusually strong for a Flash model, including 89.4% on Terminal-Bench 2.1.
— JAZII (@notjazii) September 2, 2026
Ahead of launch, early discussion focused on Gemini 3.8 Flash as a coding-oriented release tested inside Google's agentic development platform.
— Captain Insight (@CaptainInsightX) September 2, 2026
The independent speed result is exceptional. The current official table also shows both strong agent results and visible limits.
Google positions Gemini 3.8 Flash as its production workhorse for software engineering and agentic knowledge workflows. It builds on Gemini 3.7 Flash, accepts text, image, audio, and video, and can return up to 64K text tokens from a 1M-token input context.
Artificial Analysis measured the high-effort setting at 59 on its Intelligence Index and 304.6 output tokens per second. Google's current model card reports 73.7% on DeepSWE v1.1 and 89.4% on Terminal-Bench 2.1, while Terminal-Bench 4.0 and OSWorld show where stronger frontier systems still lead.
Introductory pricing is $0.75 per million input tokens and $3.75 per million output tokens through December 31, 2026. Higher effort can improve quality, but it uses more reasoning tokens and increases time to first token.
The exact Gemini 3.8 Flash route is now available in SPIRITT Workspaces and passed a live completion check on September 2. Select it in the model picker to build agentic apps and workflows with the new release.
Artificial Analysis release snapshot, September 2, 2026. Gemini 3.8 Flash high scores 59, up from Gemini 3.7 Flash high at 56, but behind the displayed frontier leaders.
Independent measurements establish overall speed, intelligence, and cost. Google's model card then supplies capability-specific comparisons, including the losses.
Artificial Analysis release snapshot. Gemini 3.8 Flash was the fastest model in its tracked set at publication.
Artificial Analysis estimates. High effort raises measured intelligence, cost, and latency relative to lower settings.
Google's current model card reports 73.7%, superseding the 71.0% in the launch-day graphic. Comparator provenance follows Google's methodology.
Google model card. Gemini results are self-computed; comparator provenance follows Google's methodology.
Google model card. Higher is better; this is not an independent rerun by SPIRITT.
Google model card. Differences among the leading displayed systems are narrow.
An important limitation: Gemini 3.8 improves over 3.7 but trails the strongest displayed comparators on the newer benchmark version.
Google model card. Stronger than Gemini 3.7 Flash, but not the leading displayed computer-use result.
Independent source: Artificial Analysis Gemini 3.8 Flash release and live model pages, observed September 2, 2026. Vendor source: Google's current Gemini 3.8 Flash model card and linked evaluation methodology. Scores, prices, and rosters can change. Run repeated evaluations on your own work before switching a production route.
Choose Gemini 3.8 Flash in the picker, then build the agentic workflow
Open a workspace and land in a fully equipped cloud computer: browser, files, terminal, integrations, and memory. No local setup. No thin chat box pretending to be an agent.

Open the model picker and select Gemini 3.8 Flash. The exact route is live in SPIRITT Workspaces and has passed a real completion check.

Tell it what to ship or what to run. Gemini 3.8 Flash can code, call tools, drive the browser, coordinate multi-step work, and keep going while you step away. The point is not another chat window. It is an agentic environment where Gemini 3.8 Flash actually does the job.

Open a SPIRITT Workspace, choose Gemini 3.8 Flash in the model picker, and turn the release into a working agentic app or workflow.