Google Rolls Out Gemini 4 Argon to Trusted Cyber Defenders, Plans Guardrail-Free Version
At a glance
- Severity
- Low
- Used in attacks
- No flaws named
- Vendors and products
- Reported by
- 1 outlet
Google on Wednesday announced its latest frontier artificial intelligence (AI) model, Gemini 4 Argon, that it said is being rolled out to a set of trusted cyber defenders through its Fairwind Program.
"It delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense," Koray Kavukcuoglu, senior vice president of Google DeepMind and Chief AI Architect at Google, said.
The development comes nearly a month after the tech giant unveiled Gemini 3.8 Flash Cyber, which it described as the most capable cybersecurity model.
Like similar models from rivals Anthropic and OpenAI, Argon is assessed to be highly capable at autonomously finding, validating, and patching critical software vulnerabilities.
This includes a previously unknown critical vulnerability exposing sensitive personal information across healthcare software used by hospitals worldwide. Google did not reveal which software was affected by the security flaw.
Argon, according to Google, demonstrates "impressive leaps" in vulnerability discovery over 3.8 Flash Cyber, and outperforms the model when it comes to discovering the attack surface and generating proof-of-concepts (PoCs) to validate the findings.
Google said it plans to release a version of Argon without cyber guardrails to trusted defenders and its internal teams so that they can take advantage of its full capabilities.
Ahead of a broader rollout, the company said it's working to strengthen safeguards to rein in misalignment, prevent model misuse by bad actors, and make it resilient to indirect prompt injections (IPIs). According to a model evaluation released by Google, Argon outperforms other models to take the top spot in the Gray Swan's IPI benchmark.
"We are deploying misalignment mitigations that monitor Argon’s chain-of-thought and actions and stop execution when necessary," Google said. "We strongly encourage the rest of the industry to preserve reasoning transparency in these pivotal moments of increased capabilities while navigating alignment risks, so that model thoughts remain helpful in identifying and diagnosing misalignment."
Originally published by The Hacker News. © The Hacker News. Written by info@thehackernews.com (The Hacker News).
Fastnexa security experts
Dealing with this in your own company?
If this story touches software, suppliers or systems you use, a Fastnexa security expert can tell you what it means for you and what to do first.
Think you’ve already been hit? Don’t wait on a form: call or WhatsApp +1 (732) 454 2616. We reply within 1 hour, 24/7. Emergency help →
Coverage
One outlet has carried this so far.
2026-10-01 07:49 UTC
Related stories
- Irony alert: OpenAI whines that Chinese model stole its special IP that it stole from everybody else
The Register · 2026-09-30
- Attackers Abuse ChatGPT Custom GPTs to Deliver RAT via ClickFix Lures
The Hacker News · 2026-09-30
- Attackers Combine ChatGPT Feature Abuse With ClickFix to Deliver Trojan Malware
Infosecurity Magazine · 2026-09-30
- Add one more AI worry to the nightmare scenario: self-replicating prompt injections
The Register · 2026-09-29
- Custom ChatGPTs push ClickFix attacks to deploy RAT malware
BleepingComputer · 2026-09-29