CISOOnline

Google makes Gemini 4 AI model available to a trusted few

The company is also aiming to close the gap with Anthropic and OpenAI in cyberdefence by enabling Argon to autonomously find, validate, and patch critical software vulnerabilities. On CWE-bench v1, a benchmark for vulnerability remediation, Google reports a 68% score, tied for the top position with Grok 4.7, GPT-6 Astra and Claude Opus 5.5.

While Google has published a broad set of benchmark results to support its claims, Jain cautioned that the scores should be treated as hints, not proof. “Argon looks to be good and beating rivals at using automated tools to complete multi-step tasks without getting confused, and it sticks closely to real facts. Its everyday coding abilities are basically average and tied with others. It still trails others when it comes to creative writing, nuanced explanations, and running command-line computer terminals,” he said.

Jain noted Argon shows that Google has regained much of the capability gap, but it has lost some momentum and developer mindshare. It is now competitive at the frontier, but not clearly ahead across all areas.



Source link