Google Unveils Lower-Cost Gemini Cybersecurity Model
Artificial intelligence systems are becoming faster at discovering, validating and repairing software vulnerabilities, but the same capabilities can help attackers develop exploits. That dual-use risk has pushed developers to balance performance, deployment cost and tight access controls. Google built Gemini 3.5 Flash Cyber on its lightweight Gemini 3.5 Flash model, positioning it as a lower-cost, per-token alternative to larger cybersecurity systems such as Anthropic’s Mythos.
Google announced Gemini 3.5 Flash Cyber on July 21, 2026, pairing the specialized model with its CodeMender security agent. CodeMender runs multiple Flash Cyber agents and combines their findings into a single report covering vulnerability detection and remediation. Google said the system delivered competitive frontier-level performance on the CyberGym benchmark, without disclosing a numerical score. Access will begin soon through a limited pilot for governments and trusted partners; the company did not provide specific pricing or a date for broader availability.
All Coverage
2 original reportsThe Backstory
The history behind this eventGoogle Launches Three Gemini Models for AI Agents
The generative AI race is shifting from chatbots toward agents that can call tools and complete multi-step tasks, making inference speed, token consumption and cybersecurity increasingly important to enterprise buyers. Google is expanding the lightweight Flash tier to lower the cost of running such workloads at scale, while strengthening the Gemini API ecosystem as competition intensifies among model providers.
As of July 29, 2026, Google has unveiled three models: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber. The releases emphasize faster computation and greater token efficiency for agentic workloads, while the cybersecurity-tuned Flash Cyber has entered a pilot focused on vulnerability remediation. Google said Gemini 3.5 Pro remains in testing, leaving a gap in the lineup, and that pretraining for the next-generation Gemini 4 has begun.
Google Unveils Gemini 3.5 Flash for Autonomous Agents and Software Development
Google introduced Gemini 3.5 Flash at its 2026 Google I/O developer conference, extending its lineup of multimodal Gemini models and targeting AI applications that require frequent calls, long context windows and real-time responses. The model delivers performance exceeding that of the previous-generation Pro at Flash-tier cost while strengthening autonomous-agent and software-development capabilities.
The newly announced Gemini 3.5 Flash outperformed the previous-generation Pro in several agent benchmarks, with reasoning and processing speeds improving fourfold. Google also demonstrated that the model could build an operating system within 12 hours at a total cost of less than $1,000. The company reiterated plans to increase investment in AI infrastructure, saying it was seeing real market demand.
Google Launches Gemini 3.1 Flash-Lite for Low-Cost, Large-Scale Workloads
Google DeepMind’s Gemini 3 lineup comprises Pro, Flash and Flash-Lite, designed respectively for advanced reasoning, speed and cost efficiency. Flash-Lite targets high-volume tasks such as translation, content moderation and data labeling, lowering the barrier for businesses deploying generative AI at scale.
Google released a preview of Gemini 3.1 Flash-Lite on March 3, 2026, making it available through Google AI Studio and Vertex AI. The API charges $0.25 per 1 million input tokens and $1.50 per 1 million output tokens, about one-eighth the price of the Pro version. Google said output speed increased by 45%.
Subscribe to Mark Radar Weekly
Every Friday, the week's strongest signals in your inbox. Unsubscribe anytime.
If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →