Close Menu
    Facebook X (Twitter) Instagram
    • Privacy Policy
    • Terms Of Service
    • Social Media Disclaimer
    • DMCA Compliance
    • Anti-Spam Policy
    Facebook X (Twitter) Instagram
    Fintech Fetch
    • Home
    • Crypto News
      • Bitcoin
      • Ethereum
      • Altcoins
      • Blockchain
      • DeFi
    • AI News
    • Stock News
    • Learn
      • AI for Beginners
      • AI Tips
      • Make Money with AI
    • Reviews
    • Tools
      • Best AI Tools
      • Crypto Market Cap List
      • Stock Market Overview
      • Market Heatmap
    • Contact
    Fintech Fetch
    Home»Crypto News»Blockchain»Gemini 4 Launches: Google’s Leading AI Model Outperforms Others in Cybersecurity
    Myriad: How low will Nvidia go? Click to make your prediction.
    Blockchain

    Gemini 4 Launches: Google’s Leading AI Model Outperforms Others in Cybersecurity

    October 1, 20264 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email
    aistudios

    In brief

    • Google unveiled Gemini 4 Argon on Wednesday, scoring 77.9% on DeepSWE v1.1 and leading 12 of 18 benchmarks in its own comparison table.
    • It posted a 0.7% attack success rate on Gray Swan’s prompt injection test, ahead of Claude Opus 5.5 and Claude Fable 5.1, which both scored 1.0%.
    • Argon goes first to vetted cyber defenders through the Fairwind Program, without cyber guardrails, before reaching paid API customers and Google AI Ultra subscribers.

    Gemini 4 is finally here, one week after the release of Claude Opus 5.5 and one day after GPT 6.1 Sol, proving American labs are very much committed to slowing down AI development. Please excuse our sarcasm.

    Google unveiled Gemini 4 Argon on Wednesday, calling it its frontier model, meaning its most capable, for coding, office work and cyber defense.

    Myriad: How low will Nvidia go? Click to make your prediction.

    On DeepSWE v1.1, a test of whether an AI can finish long, messy, real-world software engineering jobs, scored as a percentage, Argon hit 77.9%. Claude Opus 5.5 got 74.2%, GPT-6 Astra 74.1% and Claude Fable 5.1 67.4%.

    For scale, Gemini 3.6 Flash managed 49% on the same test in July. Argon can also write up to 1 million tokens in one reply, up from 64,000. A token is a chunk of text, roughly three-quarters of a word, so that is about 750,000 words versus about 48,000.

    Take these numbers with a grain of salt, though. Google computed its own DeepSWE score, while rivals’ numbers came from a public leaderboard and company reports. Its table also concedes ground. Argon leads on 12 of 18 benchmarks, ties one and trails on five, a mix of coding, science and computer-control tests.

    changelly

    But the model’s flashy feature is cyber capabilities. Hide a secret instruction inside an email, wait for an AI assistant to read it, and see if the AI obeys the stranger instead of you. That is indirect prompt injection, and it is the nightmare for anyone who wants to hand an AI their inbox or shopping cart.

    On Gray Swan’s Indirect Prompt Injection benchmark, which hides malicious instructions in content agents read and scores how often the attacks work within 15 tries, Argon landed at 0.7%. Lower is better. Claude Opus 5.5 and Claude Fable 5.1 both scored 1.0%.

    GPT-6 Astra came in at 8.5%. Grok 4.6 and Kimi K3 got tricked just over half the time, at 51.8% and 52.7%.

    Argon goes to vetted security teams through the Fairwind Program, Google’s limited-access cyber defense initiative, which launched September 2 with more than 650 partners including governments and critical infrastructure operators. And it ships “without cyber guardrails,” the built-in refusals that normally stop a model from helping with hacking.

    The logic behind such a move is that defenders need a model that can think like an attacker to patch holes before criminals find them. The catch is that the same skill cuts both ways, so Google says a phased rollout is the only safe path. It is also taking part in the U.S. government’s voluntary process for pre-release model access.

    Google isn’t the first to put a cyber model behind a velvet rope. An early version of Anthropic’s Claude Mythos helped find 271 vulnerabilities in Firefox, meaning 271 security holes Mozilla then patched. OpenAI has taken a similar route with its Trusted Access for Cyber program.

    Argon’s cyber scores jump over Gemini 3.8 Flash Cyber, the restricted model Google launched with Fairwind. On the Wiz Penetration Test Benchmark, an internal Google test that asks an AI to write working exploits against real web-application flaws without seeing the code, and scores the share solved on the first try, Argon hit 70.9% against 58.2%.

    Google also says Argon helped security firm Wiz find a critical flaw in healthcare software used by hospitals worldwide, one earlier frontier models had missed.

    The launch follows a rough summer for Google. In July it shipped smaller Flash models but skipped the promised Gemini 3.5 Pro, and Alphabet shares fell about 4.4%. Argon also landed the same day President Trump unveiled a voluntary, penalty-free AI accord that Google’s leadership signed.

    Google says wider release comes as soon as possible, starting with paid API customers and Google AI Ultra subscribers.

    Introductory pricing is $2 per million input tokens and $10 per million output tokens. Google hasn’t said when that period ends, only that standard rates are $4 and $20.

    livechat
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Fintech Fetch Editorial Team
    • Website

    Related Posts

    crypto

    Soneium and DayOneDream Aim to Tokenize K-Pop Revenue on the Blockchain

    September 30, 2026

    Chainlink CCIP 2.0 Reveals Bridge Vulnerabilities, While Issuer Restrictions Cause Delays

    September 29, 2026

    Zano Reverses Blockchain Transactions by One Month Following Exploit

    September 28, 2026
    Bitcoin Price Eyes $100K as Fidelity's Timmer Flips From His 'Year Off' Call

    Solana’s Alpenglow Reaches Second Testnet, Approaching 100ms Finality

    September 27, 2026
    Add A Comment

    Comments are closed.

    Join our email newsletter and get news & updates into your inbox for free.


    Privacy Policy

    Thanks! We sent confirmation message to your inbox.

    notion
    Latest Posts
    Laziest Way to Make Money with AI (If You're Broke)

    Laziest Way to Make Money with AI (If You’re Broke)

    October 1, 2026
    OpenAI's GPT Escaped Again, and it Proves How Dangerous AI Really Is

    OpenAI’s GPT Escaped Again, and it Proves How Dangerous AI Really Is

    October 1, 2026
    Cointelegraph

    Bitcoin ETFs Continue $3.1B Trend While Ether ETF Inflows Decline

    October 1, 2026
    Ripple No Longer With BIS Task Force; Questions Over XRPL Paper Linger

    FCA Welcomes Crypto to the UK Ahead of Upcoming Regulations

    September 30, 2026
    Bitcoin could put the average ETF buyer back in losses this week

    Bitcoin May Lead Average ETF Investors into the Red This Week

    September 30, 2026
    notion
    LEGAL INFORMATION
    • Privacy Policy
    • Terms Of Service
    • Social Media Disclaimer
    • DMCA Compliance
    • Anti-Spam Policy
    Top Insights
    ZeroDev Secures $6.7 Million Funding To Scale ERC-4337 Smart Accounts

    ZeroDev Raises $6.7 Million in Funding to Expand ERC-4337 Smart Accounts

    October 1, 2026
    Myriad: How low will Nvidia go? Click to make your prediction.

    Gemini 4 Launches: Google’s Leading AI Model Outperforms Others in Cybersecurity

    October 1, 2026
    Customgpt
    Facebook X (Twitter) Instagram Pinterest
    © 2026 FintechFetch.com - All rights reserved.

    Type above and press Enter to search. Press Esc to cancel.