Skip to content
AI ConnectPowered by VELENTIS
AI-generated2 min

Google DeepMind Announces Gemini 4 Argon with 1M Token Output Limit

Google DeepMind has introduced Gemini 4 Argon, featuring a 1-million-token output window designed for deep reasoning, alongside severe access restrictions that spark user pushback.

This article was AI-generated and published automatically. Context, labelling and all sources at the end of the article.

(KI-generiertes Symbolbild: Gemini / AI Connect)

Google DeepMind has officially unveiled Gemini 4 Argon, marking the debut of its fourth-generation model family. The company is placing a massive expansion of output capacity at the center of the release, introducing an unprecedented output limit of one million tokens per single response. Typical frontier models have previously capped generation at 64,000 or 128,000 tokens, with competing systems such as GPT-6 Astra reaching approximately 262,000 tokens. This expanded generation ceiling is specifically designed to support prolonged reasoning workflows and extensive test-time compute.

Google backs the architectural changes with concrete benchmark results, particularly in autonomous software engineering. On the rigorous DeepSWE benchmark, Gemini 4 Argon achieved a score of 77.9 percent, representing a substantial gain over the approximately 65 percent recorded by Gemini 3.7 Flash. On AutomationBench, which evaluates multi-step automation routines, the new model reached 51.3 percent. These results underscore an industry-wide pivot toward models capable of tackling end-to-end coding tasks without human guidance.

A core application area for Argon is cyber defense and automated code remediation. Testing on CWE-bench v1 shows the model achieving a 68 percent resolution rate, tying with GPT-6 Astra. Google is using this foundation to deploy CodeMender, an automated vulnerability patching system designed to independently identify and remediate security flaws across more than twenty programming languages.

Despite its technical milestones, Google faces widespread pushback over its distribution strategy. Access to Gemini 4 Argon is strictly restricted, initially available only through the Fairwind safety evaluation initiative and the newly introduced Google AI Ultra subscription tier. This new enterprise tier is priced between 100 and 200 dollars per month, presenting a significant barrier to entry for individual developers and smaller teams.

The pricing structure has sparked intense frustration among existing Google Pro subscribers, who find themselves entirely shut out from Argon despite their ongoing monthly commitments. The exclusive rollout highlights the economic realities of operating cutting-edge reasoning models, as tech companies increasingly segment frontier capabilities into premium price tiers to offset escalating compute expenses.

What this means for you

Gemini 4 Argon demonstrates how expanding token generation windows transforms autonomous code repair and reasoning. However, confining these capabilities to high-cost enterprise tiers illustrates the growing financial wall separating standard users from frontier AI capabilities.

Evidence

Solidly sourced
59/100
  • Google DeepMind introduced Gemini 4 Argon featuring an output limit of one million tokens per response.

    single source
  • Argon scores 77.9 percent on DeepSWE and 51.3 percent on AutomationBench.

    single source
  • On CWE-bench v1, Argon reaches 68 percent and serves as the foundation for the CodeMender automated vulnerability patching system across more than 20 languages.

    verified
  • Access is restricted to the Fairwind safety evaluation program and the Google AI Ultra tier priced between 100 and 200 dollars per month.

    single source

The evidence score is computed, not hand-set: from confidence, the number of sources and the share of verified statements.

Source & transparency

As of: October 03, 2026

AI-generatedAI-generated: produced automatically from vetted sources with technical quality checks (source, quote and figure verification); no human sign-off of each item before publication

Sources
2
Verified statements
1 / 4
Evidence score
59Solidly sourced

Want to put this into practice?

We connect you with suitable AI providers from the DACH region, free of charge and without obligation.

What's next?