Gemini 3.8 Flash Upgrade in three weeks: Prices remain unchanged for now, but complex tasks may require additional Token
CoinMeta
09-04 09:57
Ai Focus
On September 2nd, Google released Gemini version 3.8, as well as a 3.8 version tailored for network defense. Just three weeks after the release of 3.7 Flash, this marks the third update in six weeks for the Flash series. Google positions 3.8 as the current strongest reasoning and coding Flash model. The standard version has been deployed in Gemini API, Google AI Studio, Android Studio, Gemini Enterprise, and some consumer products; the Cyber version is only available through the new Fairwind.
Helpful
No.Help

On September 2nd, Google released Gemini version 3.8, as well as a 3.8 version tailored for network defense. Just three weeks after the release of 3.7, this marks the third update in the Flash series within six weeks. Google positions 3.8 as the current strongest reasoning and coding Flash model. The standard version has been made available for Gemini API, Google AI Studio, Android Studio, Gemini Enterprise, and some consumer products; meanwhile, the Cyber version is only accessible on a priority basis to trusted government agencies, critical infrastructure operators, and software maintainers through the new Fairwind Program. Both versions share the same basic intelligence, but they differ in deployment permissions and security protections.

The introductory price for the regular 3.8 Flash is still $0.75 per million inputs and $3.75 per output, which is the same as 3.7. However, this price only applies until December 31, 2026; starting from January 1, 2027, the officially listed price will increase to $1.50 per input and $7.50 per output. What is more easily overlooked is that the model will perform more reasoning steps on complex tasks and repeatedly call tools. Google explicitly reminds that higher levels of effort may consume more Token. Although the unit price has not increased, it does not mean that the total cost for a task will remain the same.

Flash begins the competition for long-term tasks; beyond speed, completion rate is also a key factor to consider.

The evidence presented in Google covers long-cycle coding, expert analysis, and multi-step reasoning. 3.8 Flash outperformed most larger, cutting-edge models in long-cycle software engineering benchmarks of DeepSWE and v1.1; achieving 54.9% on HLE-Verified. The company also cited financial and legal agency benchmarks to illustrate its attempt to move Flash from a low-latency question-answering model to a working model capable of continuous planning, tool invocation, and delivery of complete results. Demonstrations such as hardware disassembly visualization, single-prompt map generation DOS, and 3D games emphasize that the model not only outputs code snippets but can continuously check and modify the final product in a loop.

However, the benchmark score cannot be directly converted into corporate productivity. The success rate of long tasks is often influenced by factors such as the size of the code repository, the dependency environment, the quality of testing, tool permissions, and the retry budget. While a "more diligent" strategy of 3.8 can increase the completion rate, it also increases the execution time and the usage of Token. Development teams need to keep track of the cost per call, the total Token for a task, the number of tool calls, the number of retries before success, and the time required for manual rework in order to determine whether the upgrade is truly more cost-effective. For workloads with latency or budget constraints, it is still recommended to reduce the level of effort to 3.7 or continue using version 3.7; the older versions are not immediately discontinued due to the release of new products.

The regular version is now available to developers and enterprises. Users with Google AI Pro and Ultra accounts can also use it within the Gemini application, as well as in searches for AI Mode, Gemini, in, and Sheets. The features, quotas, and regional availability may vary depending on the entry point, so it should not be assumed that "the model has been released" means all users or all products have access to exactly the same capabilities. This is especially true for workflows that involve autonomous tools; their successful implementation also depends on whether the application provides the corresponding tools, permissions, and a recoverable execution environment for the models.

The Cyber version is stronger but also narrower; 47.2% is not an automatic repair commitment.

Gemini Focusing on vulnerability discovery and patch generation. Google claims that it has a success rate of over 70% in internal vulnerability discovery evaluations for 20 programming languages; in external CWE-Bench patch evaluations, pass @1 has a rate of 47.2%, close to the 47.8% of a leading model, but at a lower cost. Chrome Internal usage by security teams shows that 3.8 Flash Cyber generates 2.6 times as many correct vulnerability patches as the best large commercial models. Wiz also reports an increase in the recall rate of its internal penetration testing benchmarks by 7.5 to 9.7 percentage points, with costs reduced by 2.3 to 5.2 times.

These numbers have reference value, but they also have their limitations. The internal benchmarks, hints, and manual judgment methods are not fully detailed in the announcement; 47.2% of pass @1 also means that a single generation cannot guarantee that more than half of the issues will be correctly fixed. Google prioritizes "fixing" over "utilizing," and adds protection against chemical, biological, radiological, nuclear, and cyber abuses to the standard 3.8. The Cyber version adopts looser cybersecurity restrictions to meet professional defense needs, therefore it is not publicly available and is only provided to trusted defenders within the Fairwind program.

What truly changes with this release is the approach to the Flash series: it no longer relies solely on low prices and speed to attract bulk requests, but instead pursues a higher completion rate for complex tasks through longer reasoning cycles. If companies simply replace version 3.7 with 3.8, the most common mistake is to compare only the pricing of the million Token units. A more cautious approach would be to conduct tiered testing on actual tasks: limit the effort level for simple requests, allow more cycles for longer tasks, and maintain manual approval and isolated environments for secure tasks. Version 3.8 of Flash is already available, and Cyber has also entered a limited-release phase; as for whether it can complete more production tasks at a lower total cost, that will have to be answered by each team's own end-to-end data analysis.

Tip
$0
Like
0
Save
0
Views 56
WalletJYS reminds readers to view blockchain rationally, stay aware of risks, and beware of virtual token issuance and speculation. All content on this site represents market information or related viewpoints only and does not constitute any form of investment advice. If you find sensitive content, please click“Report”,and we will handle it promptly。
Submit
Comment 0
Hot
Latest
No comments yet. Be the first!
Related
Solana Launches Payment Channel Solution: Millions of Small Payments Do Not Equal Millions of On-Chain Settlements
On September 3, the Solana Foundation introduced a payment channel solution for AI agents and high-frequency, small-value transactions, claiming that it can support up to 1 million payments per second. This figure is easily misinterpreted. The core of the payment channel is not to allow the Solana mainnet to record 1 million independent transactions per second, but rather to have users authorize a certain amount first, conduct a large number of off-chain measurements within the channel, and then settle the aggregated results on-chain. What it improves is the processing capacity of payment events, while reducing the actual number of on-chain settlements.
币界网
·2026-09-05 10:01:48
36
cirBTC Launched on Ethereum: Circle Putting Reserve Transparency at the Core of Bitcoin Packaging
On September 4th, Circle introduced its reserve scheme for its Bitcoin packaging product cirBTC. Official information indicates that cirBTC has already been launched on Ethereum, supported by native Bitcoins at a 1:1 ratio; when the mainnet of Arc goes live, it is planned to provide further native support, and in the future, it may also be extended to other blockchains. It is important to distinguish between these two states: the Ethereum version is already available, but the support for Arc is still planned to be launched when the mainnet goes live, and cannot be described as currently being fully covered.
币界网
·2026-09-05 10:00:42
36
Eurozone retail sales fell 0.6% month-on-month in July: Consumption has not slowed down, but there is a clear structural divergence
Preliminary estimates released by the European Statistical Office on September 4 showed that retail trade volume in the eurozone decreased by 0.6% month-on-month in July, after seasonal adjustment, while the overall EU saw a decline of 0.4%. This reverses the 0.2% growth seen in both the eurozone and the EU in June. Compared to the same period last year, retail sales in the eurozone still increased by 0.6%, and in the EU by 1.0%. Therefore, a more accurate assessment is that consumption has not yet entered a full recession, but rather that the monthly momentum has weakened, and the gap between different categories and member states has widened.
币百科
·2026-09-05 09:59:44
16
162,000 new jobs added in the U.S. in August: Employment is picking up, but the labor market is still in transition
The U.S. Bureau of Labor Statistics released employment data for August on September 4th, showing that non-farm employment increased by 162,000 jobs, with the unemployment rate remaining at 4.1%. Looking solely at the new jobs created, this is a significant increase compared to the average of 31,000 jobs per month over the previous twelve months; however, when considering industry structure, historical adjustments, and wages, this report appears more like a temporary rebound rather than a full-scale acceleration in the labor market.
币百科
·2026-09-05 09:58:44
16
Google Launches Teacher AI Training in India: Classroom Applications Start with Reducing the Burden of Lesson Preparation
On September 4th, Google India announced a training program for educators, aiming to help teachers utilize generative AI in lesson preparation, organizing materials, and stimulating classroom creativity through training activities conducted across multiple locations. The announcement was released around India's Teachers' Day, with the focus not on using AI to replace traditional teaching methods, but rather on integrating it into teachers' daily work: to reduce repetitive tasks, allowing teachers to devote more time to students, providing feedback, and enhancing classroom interactions.
CoinMeta
·2026-09-05 09:57:41
15
View More