Companies

Google Launches Gemini 4 Argon, Its Most Advanced AI Model Yet

Google's Gemini 4 Argon sets a real-world software engineering record and rolls out first to cybersecurity partners amid U.S. government safety evaluations.

Google rolls out Gemini 4 Argon, its most advanced AI model
Google rolls out Gemini 4 Argon, its most advanced AI modelAI-generated
By Sophie Lindqvist4 min read

Updated

Why it matters

  • Alphabet unveiled Gemini 4 Argon on Wednesday, its most advanced AI model, setting a real-world software engineering record.
  • Argon ties with OpenAI's GPT-6 Astra and Grok 4.7 on cybersecurity benchmarks and leads GPT-6 Astra and Anthropic's Fable 5.1 on the Vals Index.
  • Google is rolling the model out in phases, starting with trusted cybersecurity partners, while working with the U.S. government on pre-release safety evaluations — a day after CEO Sundar Pichai signed a voluntary AI safety agreement with President Trump.

Alphabet unveiled Gemini 4 Argon on Wednesday, its most advanced artificial intelligence model yet, with major improvements in coding, cybersecurity, and complex professional work.

Google said the model sets a new record in real-world software engineering. It ties for first place in cybersecurity and leads another benchmark that measures performance across finance, legal, and other professional tasks. Argon also ties with OpenAI's GPT-6 Astra and Grok 4.7 on cybersecurity evaluation benchmarks, and it finishes ahead of GPT-6 Astra and Anthropic's Fable 5.1 on the Vals Index.

The stakes are straightforward. Frontier model leadership determines who sets the terms of the AI market — enterprise contracts, developer mindshare, and, increasingly, government relationships. Gemini 3 put Google back at the forefront of the AI model race nearly a year ago, and Argon represents the company's next major push at the frontier after a recent pivot toward scaling faster, lower-cost flash models.

An internal workhorse before a public release

Argon is already working inside Google. The company said it uses the model to optimize memory at its data centers, freeing up hundreds of terabytes of memory without buying additional hardware. Quantum computing researchers have also used the model.

Those internal deployments give Google concrete evidence of the model's value in production infrastructure — a signal that matters to enterprise buyers weighing frontier models against cheaper alternatives.

Cybersecurity defenders get it first

Google plans to launch the new model in phases. The first phase starts with trusted cybersecurity partners, while the company works with the U.S. government on pre-release safety evaluations.

"Starting this rollout in this way gives us more confidence, but also enables us to put a model that is trained and strong in cyber defense in the hands of defenders as soon as possible," Tulsee Doshi, Google's Gemini model product lead, told CNBC.

The phased approach marks a departure from the standard practice of wide public release on day one. It reflects the mounting pressure on AI developers to demonstrate safety rigor before shipping frontier capabilities — particularly models with strong offensive and defensive security skills.

The timing is not incidental. The model release comes a day after CEO Sundar Pichai signed a voluntary agreement with President Donald Trump and major tech executives at the White House, following a luncheon to address rising AI safety concerns. Launching Argon through a government-evaluated, defender-first pipeline gives Google a concrete example of that commitment in practice.

A step up on security benchmarks

Doshi described Gemini 4 as an "incredibly well-rounded" model that excels at running long, multi-step tasks — the kind of sustained, agentic work that has become the central test for frontier models.

The new model represents a meaningful cybersecurity step up from Google's 3.8 Flash Cyber model, which the company released earlier this month. Argon outperforms it in vulnerability discovery, according to Google.

"We really believe that a model of this caliber and this level of frontier performance is meaningfully important for defenders," Doshi said.

Safeguards before the public launch

Google said it is looking to scale safeguards in four key areas, including misuse and prompt injection, before launching Argon publicly. Prompt injection — attacks that manipulate models through crafted inputs — has emerged as one of the hardest problems in deploying AI agents that interact with untrusted data.

The company did not specify a date for the public launch.

Why it matters

Argon arrives at a moment when the AI industry's competitive and regulatory tracks are converging. Google is simultaneously racing OpenAI, xAI, and Anthropic on benchmarks while signing voluntary safety commitments at the White House. The defender-first rollout lets the company do both: it puts frontier capability in the hands of cybersecurity partners immediately while building a safety record with the U.S. government before general availability.

The benchmark spread tells its own story. Argon ties GPT-6 Astra and Grok 4.7 on cybersecurity but pulls ahead of GPT-6 Astra and Anthropic's Fable 5.1 on the Vals Index, which spans professional domains such as finance and legal. No single model dominates across the board — and that compression at the top means deployment strategy, safety posture, and internal proof points like Google's data center memory gains may decide more customers than raw scores do.

The open question is timing. Google has committed to public availability only after it scales safeguards in its four identified areas, and the pace of that work will determine whether Argon's frontier lead holds against rivals shipping broadly today.

Original: blog.google

Share this article:

More from Sophie Lindqvist

Sophie Lindqvist

Show full bio

Staff writer covering marketplaces and e-commerce at AI In Context.

151 articles

Related articles

  1. Google Announces Gemini 4 Argon, But No One Outside Can Use It
  2. Google Announces Gemini 4 Argon, Its Next Frontier Model
  3. Google Ships Gemini 3.5 Flash, Promises Pro Model Next Month
  4. Google Upgrades Gemini 3 Deep Think With Record Benchmark Runs
  5. Google Ships Gemini 3.8 Flash and a Cybersecurity-Only Variant

« Previous article