Connect with us

Hi, what are you looking for?

SecurityWeekSecurityWeek

Artificial Intelligence

Google Launches Gemini 4 Argon With Guardrail-Free Access for Vetted Defenders

The company says its new frontier AI model found a critical vulnerability in software used by hospitals worldwide.

Gemini

Google on Wednesday announced Gemini 4 Argon, its new frontier AI model, which it is initially rolling out to a select group of trusted cyber defenders through its Fairwind Program.

Google’s internal teams are also using the model. The company says it will expand access gradually as it gathers feedback from early testers.

The tech giant says Argon is designed for complex workflows in software engineering, enterprise knowledge work such as legal and finance, and cybersecurity defense. On the security side, the company says it trained the model to be highly capable at cyber defense, and that it can autonomously find, validate, and patch critical software vulnerabilities.

“For trusted defenders and our own internal teams at Google, we’ll be releasing Argon without cyber guardrails so they can leverage its full frontier-level cybersecurity defense capabilities,” the company said.

Google launched Fairwind in early September as a limited access program for governments, Google Cloud customers, and cybersecurity partners. It initially combined the Gemini 3.8 Flash Cyber model with Google’s CodeMender harness, which finds, verifies, and fixes vulnerabilities. At launch, the program had more than 650 participating partners.

Google says Argon found a flaw in hospital software

Wiz, which Google acquired earlier this year, is using Argon in its Scan for Good initiative, which finds and remediates high-risk exposures in critical public infrastructure for free.

Advertisement. Scroll to continue reading.

“In an early demonstration of its impact, the model uncovered a critical vulnerability exposing sensitive personal information across healthcare software used by hospitals worldwide, identifying a severe risk that previous frontier models had missed,” Google said.

The announcement does not name the affected software or say whether the issue has been addressed.

On CWE-bench v1, a vulnerability remediation benchmark developed by Collinear AI, Argon tied for first place with a score of 68%, alongside OpenAI’s GPT-6 Astra and xAI’s Grok 4.7.

The company also claims improvements in vulnerability discovery compared to Gemini 3.8 Flash Cyber. On Google’s internal vulnerability benchmark, Argon found a wide range of exposures across complex codebases written in 20 programming languages.

On Wiz’s internal black-box penetration testing benchmark, which targets live web systems without access to source code, Argon outperformed Gemini 3.8 Flash Cyber. Google says it was better at discovering the attack surface, identifying vulnerabilities, and producing proof-of-concept evidence.

Google is strengthening safeguards before a wider release

“Safely releasing frontier capabilities at this level requires a phased approach,” Google said, noting that it is taking part in the US government’s voluntary process for pre-release model access.

For the broader rollout, Google says the model is designed to refuse requests that could enable cyber or chemical, biological, radiological, and nuclear (CBRN) attacks, while still supporting legitimate dual-use scientific research.

These safeguards include improved techniques for monitoring the model’s internal activations to spot misuse. Google also describes Argon as its most resilient model yet against indirect prompt injection.

“In order to prevent Argon from stepping out of bounds to try to accomplish a task in a way that goes beyond the user’s intentions, we are deploying misalignment mitigations that monitor Argon’s chain-of-thought and actions and stop execution when necessary,” the company explained.

The company is also isolating and sealing its sandboxed environments before high-risk training or evaluations begin.

Related: Google: AI Is Changing the Pace and Profile of Vulnerability Discovery

Related: Anthropic Flags AI Agent Liability Risks as OpenAI Faces Hacking Lawsuit

Related: Trump Says Top Tech Firms Have Signed Accord to ‘Self-Police’ AI Development

Written By

Eduard Kovacs (@EduardKovacs) is senior managing editor at SecurityWeek. He worked as a high school IT teacher before starting a career in journalism in 2011. Eduard holds a bachelor’s degree in industrial informatics and a master’s degree in computer techniques applied in electrical engineering.

Daily Briefing Newsletter

Subscribe to the SecurityWeek Email Briefing for the latest cybersecurity threats, trends, and expert insights.

Trending

Daily Briefing Newsletter

Subscribe to the SecurityWeek Email Briefing to stay informed on the latest threats, trends, and technology, along with insightful columns from industry experts.

Learn how to address potential risks and not restrict AI adoption in your organization. See what a centralized AI gateway is and how it works in practice.

Register

Join as we decipher the world of zero trust and share war stories on securing an organization by eliminating implicit trust and continuously validating every stage of a digital interaction.

Register

People on the Move

David Cass has joined Grayscale Investments as Chief Risk Officer.

Thomas Dager has been appointed Vice President and Chief Information Security Officer at The Goodyear Tire & Rubber Company.

Alex Stamos has become Chief Information Security Officer at Cognition.

More People On The Move

Expert Insights

Daily Briefing Newsletter

Subscribe to the SecurityWeek Email Briefing to stay informed on the latest cybersecurity news, threats, and expert insights. Unsubscribe at any time.