Google Releases Gemini 4 Argon, Claiming Benchmark Edge Over OpenAI and Anthropic
Google has released Gemini 4 Argon, a flagship model it says beats top OpenAI and Anthropic offerings on some benchmarks, with a phased rollout starting with trusted cybersecurity partners and U.S. safety reviews.
The launch follows a turbulent year for Google's AI ambitions. Gemini 3, released at the end of 2025, put Google in the frontier AI conversation, but Anthropic and OpenAI have been at the forefront of advances for much of 2026. When Demis Hassabis stepped down as chief executive of DeepMind in August, his successor, Koray Kavukcuoglu, inherited a race to catch up with the companies led by Sam Altman and Dario Amodei.
On the Artificial Analysis Intelligence Index, a composite benchmark score, Gemini 4 Argon trails only Anthropic's Claude Opus 5.5 and Claude Sonnet 5.5 on its leaderboard.
"Google Gemini 4 Argon shows advanced reasoning on critical tasks, according to benchmarks, particularly legal reasoning, finance and other aspects of enterprise knowledge work, including long-running tasks," Tim Law, director of research for AI at IDC, told CNBC.
Lian Jye Su, chief analyst at Omdia, said Google had been "late to the cybersecurity and coding game," but that Gemini 4 took the company to the frontier in AI, particularly in cybersecurity. Su said the numerous security incidents caused by rogue AI systems made by OpenAI and Anthropic had created an opportunity for Google to "stake its claims as the trusted model provider for safe and secure AI operations."
Google has emphasized that positioning. The company plans to launch the model in phases, beginning with trusted cybersecurity partners, while working with the U.S. government on pre-release safety evaluations.
"Starting this rollout in this way gives us more confidence, but also enables us to put a model that is trained and strong in cyber defense in the hands of defenders as soon as possible," Tulsee Doshi, Google's Gemini model product lead, told CNBC.
Not all analysts see the release as a change in leadership. Nick Patience, AI lead at the Futurum Group, said that while Argon has made Google competitive again in AI, the model does not make it the "leader."
"At a minimum, though, Gemini 4 Argon opens up another competitive front for enterprise knowledge work with other leading frontier models," Law said.
Google said in its release that the new model is already being used internally to optimize memory at its data centers, freeing up hundreds of terabytes of memory without buying additional hardware. Quantum computing researchers have also used the model, according to the company.
The wider test will come with broader availability. Google has given no indication of when that might be.
"While Google touts the success of internal use cases, the final proof will be in enterprise production environments once the model is fully released," Law said.
Editor's Summary
Google has released Gemini 4 Argon, a flagship model it says outperforms top OpenAI and Anthropic offerings on some benchmarks, with a phased rollout starting with trusted cybersecurity partners and U.S. government safety evaluations. Analysts at IDC and Omdia describe it as a return to the frontier, particularly in cybersecurity and enterprise reasoning, while the Futurum Group says it does not make Google the leader. Google has not said when the model will be more widely available, which analysts call the decisive test.