Mark RadarMARK RADAR
About
EN
Sign in
Event File AI Google AI Models

Google Unveils Gemini 3.5 Flash for Autonomous Agents and Software Development

7 reports · First detected 2026-05-20 · Last active 2026-05-21

Google introduced Gemini 3.5 Flash at its 2026 Google I/O developer conference, extending its lineup of multimodal Gemini models and targeting AI applications that require frequent calls, long context windows and real-time responses. The model delivers performance exceeding that of the previous-generation Pro at Flash-tier cost while strengthening autonomous-agent and software-development capabilities.

The newly announced Gemini 3.5 Flash outperformed the previous-generation Pro in several agent benchmarks, with reasoning and processing speeds improving fourfold. Google also demonstrated that the model could build an operating system within 12 hours at a total cost of less than $1,000. The company reiterated plans to increase investment in AI infrastructure, saying it was seeing real market demand.

All Coverage

7 original reports

The Backstory

The history behind this event
Google Launches Gemini 3.8 Flash as Flagship Pro Model Stalls2026-09-03 · 5 reports · similarity 0.85

Google is accelerating its Gemini rollout with the Flash line, which is designed to balance speed, cost and reasoning performance as competition intensifies with OpenAI and Anthropic. The strategy has grown more important as Gemini 3.5 Pro, its flagship model, has reportedly been delayed three times because of an underlying architecture overhaul and reasoning bottlenecks, highlighting the engineering challenges involved in scaling frontier AI systems.

As of Sept. 3, 2026, Google has released Gemini 3.8 Flash after internal testing, touting broad gains in software engineering and reasoning while warning that the model’s heavier computation could raise costs. Google DeepMind also introduced Gemini 3.8 Flash Cyber, which uses the same core model under a more restricted access framework for vulnerability research. Early-access pricing was reported at about one-sixth the cost of Anthropic’s Opus 5, while Gemini 3.5 Pro still has no announced launch date.

Google Unveils Gemini 3.7 Flash in August AI Push2026-09-02 · 1 reports · similarity 0.87

Google is tightening the links between its Gemini models, Pixel hardware and productivity services as it seeks to turn advances in artificial intelligence into widely used products. Its August rollout emphasized lower-cost tools for developers, voice-driven workflows and on-device assistance, showing how the company is shifting from headline model performance toward practical applications across coding, communications, media creation and consumer devices.

In a Sept. 1, 2026 roundup, Google said Gemini 3.7 Flash arrived just three weeks after 3.6 Flash, with introductory pricing at half the earlier model’s original cost per million tokens. The company also introduced Gemini 3.5 Transcribe and the Pixel 11 lineup, powered by the Tensor G6 chip and the latest Gemini Nano. The Gemini app surpassed 1 billion monthly users, while Gemini Omni 1.1 Flash added scene extension, frame interpolation and 4K upscaling for video production.

Google Launches Gemini Omni 1.1 Flash With 40-Second Video Extension2026-08-30 · 3 reports · similarity 0.82

Google has released Gemini Omni 1.1 Flash, expanding its generative-video lineup as competition shifts from producing isolated clips to supporting longer, more controllable sequences. The model targets a persistent challenge in AI video: keeping characters and scenes consistent across separately generated segments. First- and last-frame controls, alongside tiered resolution pricing, are designed to give creators more influence over shot composition, output quality and production costs.

The upgrade adds a scene-extension feature that can build continuous, segmented videos lasting as long as 40 seconds. It also increases the preceding context used for each continuation to 10 seconds, aimed at improving character consistency between clips. Gemini Omni 1.1 Flash supports specified opening and closing frames, four resolution-based pricing tiers and 4K upscaling. The supplied reports did not specify a launch date or disclose prices for the four tiers.

Google Launches Gemini 3.7 Flash With Half-Price API Offer2026-08-17 · 9 reports · similarity 0.82

Google’s Flash lineup is positioned as a workhorse tier that balances speed, cost and reasoning capability for high-volume coding and AI-agent workloads. Rapid model cycles and falling inference costs are becoming central to competition among Google, OpenAI and other providers seeking to lock developers and enterprises into their platforms. Gemini 3.7 Flash first surfaced in leak reports and a briefly visible Python SDK pull request, signaling that another release was imminent.

Google formally released Gemini 3.7 Flash on Aug. 13, 2026, just over three weeks after unveiling Gemini 3.6 Flash on July 21. Standard Gemini API pricing is set at an introductory $0.75 per million input tokens and $3.75 per million output tokens through Dec. 31, half the regular rates. Prices rise to $1.50 and $7.50, respectively, on Jan. 1, 2027. Google describes the model as its most capable Flash offering for agentic workflows and multimodal reasoning.

Google Rolls Out Faster, Cheaper Gemini Models2026-08-05 · 1 reports · similarity 0.88

Google is positioning Gemini as the foundation for AI agents that can plan, call tools and complete multi-step tasks across software and, increasingly, robotics. Its Flash family targets the middle ground between frontier-grade intelligence and production economics, where latency, token use and reliability determine whether companies can deploy agents at scale. The strategy also links Google DeepMind’s models with consumer gateways including Search and the Gemini app, intensifying competition for developers and everyday AI users.

On July 21, Google DeepMind released Gemini 3.6 Flash and Gemini 3.5 Flash-Lite. The company said 3.6 Flash uses 17% fewer output tokens than 3.5 Flash and costs $1.50 per million input tokens and $7.50 per million output tokens. Flash-Lite generates 350 output tokens a second and is priced at $0.30 for input and $2.50 for output. Both are available through the Gemini API and Google AI Studio, while Flash-Lite is rolling out in Google Search.

Google Launches Three Gemini Models for AI Agents2026-07-29 · 9 reports · similarity 0.89

The generative AI race is shifting from chatbots toward agents that can call tools and complete multi-step tasks, making inference speed, token consumption and cybersecurity increasingly important to enterprise buyers. Google is expanding the lightweight Flash tier to lower the cost of running such workloads at scale, while strengthening the Gemini API ecosystem as competition intensifies among model providers.

As of July 29, 2026, Google has unveiled three models: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber. The releases emphasize faster computation and greater token efficiency for agentic workloads, while the cybersecurity-tuned Flash Cyber has entered a pilot focused on vulnerability remediation. Google said Gemini 3.5 Pro remains in testing, leaving a gap in the lineup, and that pretraining for the next-generation Gemini 4 has begun.

Google Unveils Lower-Cost Gemini Cybersecurity Model2026-07-22 · 2 reports · similarity 0.82

Artificial intelligence systems are becoming faster at discovering, validating and repairing software vulnerabilities, but the same capabilities can help attackers develop exploits. That dual-use risk has pushed developers to balance performance, deployment cost and tight access controls. Google built Gemini 3.5 Flash Cyber on its lightweight Gemini 3.5 Flash model, positioning it as a lower-cost, per-token alternative to larger cybersecurity systems such as Anthropic’s Mythos.

Google announced Gemini 3.5 Flash Cyber on July 21, 2026, pairing the specialized model with its CodeMender security agent. CodeMender runs multiple Flash Cyber agents and combines their findings into a single report covering vulnerability detection and remediation. Google said the system delivered competitive frontier-level performance on the CyberGym benchmark, without disclosing a numerical score. Access will begin soon through a limited pilot for governments and trusted partners; the company did not provide specific pricing or a date for broader availability.

Google Launches Gemini 3.1 Flash Live, Expands Search Live to More Than 200 Countries2026-03-27 · 2 reports · similarity 0.80

Google first launched Search Live in the United States in September 2025, extending generative AI beyond text search to continuous voice- and camera-based follow-up questions. The service combines Google’s search index with AI Mode, breaking down questions during conversations and providing links to webpages. The move reflects Google’s effort to reshape its core search gateway with Gemini. The company did not disclose the investment behind the latest rollout.

On March 26, 2026, Google released Gemini 3.1 Flash Live, a real-time audio model designed to improve the speed, naturalness and stability of multilingual responses. It also expanded Search Live to more than 200 countries and territories where AI Mode is available, with support for 98 languages. Users can tap Live in the Google app on Android or iOS, or use Google Lens to conduct continuous searches by combining the camera with voice queries.

Google Launches Gemini 3.1 Flash-Lite for Low-Cost, Large-Scale Workloads2026-03-04 · 3 reports · similarity 0.87

Google DeepMind’s Gemini 3 lineup comprises Pro, Flash and Flash-Lite, designed respectively for advanced reasoning, speed and cost efficiency. Flash-Lite targets high-volume tasks such as translation, content moderation and data labeling, lowering the barrier for businesses deploying generative AI at scale.

Google released a preview of Gemini 3.1 Flash-Lite on March 3, 2026, making it available through Google AI Studio and Vertex AI. The API charges $0.25 per 1 million input tokens and $1.50 per 1 million output tokens, about one-eighth the price of the Pro version. Google said output speed increased by 45%.

Mark Radar|MARK RADAR

If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →

All times are in Taipei time (GMT+8)