Google staff turn to Gemini 3.8 Flash as AI costs mount

Wait 5 sec.

Google employees have reportedly begun testing the Gemini 3.8 Flash Preview on its own coding platform, which is further confirmation that the leading AI companies are working on innovative and cost-effective models at a remarkable rate.Speed is important for companies observing the rising cost of AI. According to Gartner, global expenditure on AI platforms and models is projected to reach $64.25 billion in 2026, representing an annual increase of 63.4% over 2025, as companies are becoming increasingly picky about the value of their expenditures.Business Insider learned from images that the “Gemini 3.8 Flash Preview” model was posted on Google’s internal Jetski platform. An employee has reported that it seems to be better than version 3.7 Flash, but warned that this might not be the correct assessment. Google refused to make any comments. Hence, the version number and eventual public release remain unknown.A near-monthly cadence aimed at the low-cost tierGoogle has been on the go. Gemini 3.6 Flash was introduced on July 21, and 3.7 Flash became available on August 13, all taking a little over three weeks.That’s done on purpose. During Alphabet’s second-quarter earnings call, CEO Sundar Pichai stated that Google expected to release models “almost at a monthly cadence” while working on Gemini 4.Flash also represents the area where Google is having its greatest pricing effort. Google calls the series a “workhorse” for coding and agents – applications where repeat usages of the language model leads to significant token costs.Gemini 3.7 Flash launched at an introductory $0.75 per million input tokens and $3.75 per million output tokens through the end of 2026. That is exactly half the $1.50 input and $7.50 output launch pricing of Gemini 3.6 Flash.Racing on price because the frontier is out of reachGoogle’s approach is evident in the way its public product portfolio has gaps at the very top. On May 19, the company announced that Gemini 3.5 Pro was being used inside the company and would come out in June. However, by July 21, the model was still being tested with partners and Google announced that it would be rolled out to the public once it was ready.For now, that leaves Google competing aggressively on price and release speed while its flagship remains pending.Artificial Analysis gives Gemini 3.7 Flash at high reasoning effort an Intelligence Index score of 56, compared with 61 for GPT-5.6 Sol at maximum effort and 62 for Claude Fable 5.OpenAI’s Luna tier remains considerably cheaper than Google’s Flash offer at $0.20 per million input tokens and $1.20 per million output tokens. More broadly, the comparison shows how wide AI pricing has become, from cents to tens of dollars per million tokens depending on capability.Why buyers are the ones setting the termsThe market backdrop helps explain Google’s approach. Gartner expects AI model and platform spending to surge this year, but analyst Arunasree Cheparthi said enterprise budgets are facing “greater scrutiny,” with more attention on efficiency, cost control and measurable outcomes.Ramp’s spending data shows how that pressure can affect even top-performing models. In its first month, Claude Fable 5 accounted for just 6% of the tokens businesses bought from Anthropic and 11.4% of Anthropic model spending, despite launching at the top of Artificial Analysis’ intelligence ranking. Anthropic charges $10 per million input tokens and $50 per million output tokens for the model.Ramp economist Ara Kharazian said the pattern suggests there is an upper limit to what businesses will pay for raw capability. That favors models that are cheap, fast and good enough for the job — precisely the part of the market Google is trying to capture.The announcement of a public Gemini 3.8 Flash release is still uncertain. Currently, the information regarding its demonstration comes from the claims made by Business Insider, meaning that there has not yet been an official recognition of this fact by Google, which is of utmost importance given the frequently changing names of models and release dates in the market. If you're reading this, you’re already ahead. Stay there with our newsletter.