CheckNet.NETWORK DIAGNOSTICS / TOOLKIT
WorkspaceAI Model TrackerNETWORK TOOLKIT
Model trackerAll models and earlier releases

Google DeepMind · launched 30 Sept 2026

Gemini 4 Argon

Limited access

Google's newest frontier model for long-running coding, knowledge work and cybersecurity defense, initially available to trusted testers.

Read the official announcement
Model details
Vendor
Google DeepMind
Launch date
30 Sept 2026
Input price
$2.00 / 1M tokens
Output price
$10 / 1M tokens
Cached input
$0.10 / 1M tokens
Batch discount
Not published
Context window
Not published
Max output
1M tokens
Input types
Text, Image, Video, File
Availability
Limited access

Announced introductory rates for the upcoming broader launch. Regular rates will be $4 input and $20 output per million tokens; the introductory end date and batch discount are not published.

Official pricing source · checked 8 Oct 2026

Official model specifications

Price compared with other models

ModelLaunchedInput / 1MOutput / 1MCached / 1MBlended / 1Mvs Gemini 4 Argon
Claude Haiku 5.5AnthropicListed prices and comparisons apply to prompts up to 100K tokens. Above 100K, input costs $0.50, output $2.50 and cache reads $0.05 per million tokens.7 Oct 2026$0.10$0.50$0.01$0.2095% cheaper
GPT-6 LunaOpenAI22 Sept 2026$0.10$0.50$0.01$0.2095% cheaper
GPT-5.6 LunaOpenAI9 Jul 2026$0.20$1.20$0.02$0.4589% cheaper
Gemini 3.5 Flash-LiteGoogle DeepMindText, image, video and audio inputs share the listed rate. Cache storage and grounding are billed separately.21 Jul 2026$0.30$2.50$0.03$0.8579% cheaper
Gemini 3.8 FlashGoogle DeepMindIntroductory rates through 31 December 2026. From 1 January 2027, input is $1.50, output $7.50 and cached input $0.15 per million tokens. Cache storage and grounding are billed separately.2 Sept 2026$0.75$3.75$0.075$1.5063% cheaper
Claude Haiku 4.5Anthropic15 Oct 2025$1.00$5.00$0.10$2.0050% cheaper
Grok 4.7SpaceXAIBase rates apply up to 200K context tokens; higher-context requests use different rates. Batch API is not supported. The separate fast variant costs twice as much.21 Sept 2026$2.00$6.00$0.50$3.0025% cheaper
Claude Sonnet 5Anthropic30 Jun 2026$2.00$10$0.20$4.00Same price
Claude Sonnet 5.5AnthropicCache reads were reduced from $0.20 to $0.10 per million tokens on 7 October 2026.28 Sept 2026$2.00$10$0.10$4.00Same price
Gemini 4 ArgonGoogle DeepMind · Limited accessAnnounced introductory rates for the upcoming broader launch. Regular rates will be $4 input and $20 output per million tokens; the introductory end date and batch discount are not published.30 Sept 2026$2.00$10$0.10$4.00Baseline
GPT-6 SolOpenAIListed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.40 and output $15 per million tokens for the full request.22 Sept 2026$2.00$10$0.20$4.00Same price
GPT-6.1 SolOpenAIListed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.20 and output $15 per million tokens for the full request.29 Sept 2026$2.00$10$0.10$4.00Same price
Gemini 3.1 ProGoogle DeepMind · PreviewRates apply to prompts up to 200K tokens. Above 200K, input is $4, output $18 and cached input $0.40 per million tokens. Cache storage and grounding are billed separately.19 Feb 2026$2.00$12$0.20$4.5013% more
Claude Opus 5.5Anthropic22 Sept 2026$4.00$20$0.20$8.002× the price
GPT-5.6 SolOpenAIPrice before the GPT-6 launch, from OpenAI's announcement. OpenRouter now lists $2 / $10.9 Jul 2026$4.00$20–$8.002× the price
Claude Opus 5Anthropic24 Jul 2026$5.00$25$0.50$102.5× the price
Claude Fable 5Anthropic9 Jun 2026$10$50$1.00$205× the price
Claude Fable 5.1Anthropic1 Sept 2026$10$50$0.25$205× the price
GPT-6 AstraOpenAI4 Sept 2026$10$50$1.00$205× the price

GPT-6.1 Sol and GPT-6 Sol rates were checked on 8 October 2026 against OpenAI model documentation. USD per million tokens, checked 2 Oct 2026 via OpenRouter. Comparison uses the blended price (3 input tokens for every output token). Haiku 5.5 and Sonnet 5.5 prices were updated from Anthropic's 7 October 2026 announcement. Gemini and Grok prices were checked against official Google and xAI documentation on 8 October 2026.

Official benchmark results

Gemini 4 Argon launch results

Reported by Google DeepMind on 30 Sept 2026 · exact values

Exact published scores · Higher is better. Unreported results are omitted.

BenchmarkGemini 4 ArgonGPT-6 AstraClaude Fable 5.1Claude Opus 5.5
Vals IndexKnowledge work68.9%63.1%65.8%67%
AutomationBenchKnowledge work · Score51.3%41.4%31.4%42.5%
Vals Finance Agent v2Knowledge work65.4%53.5%58.9%58.6%
Harvey's Legal Agent BenchmarkKnowledge work19.6%5.4%6.7%3.8%
DeepSWE v1.1Agentic coding77.9%74.1%67.4%74.2%
FrontierSWE v2Agentic coding55%65.5%56.3%62.3%
Vibe Code BenchAgentic coding91.9%89.6%90.3%90.3%
Terminal-bench 4.0Agentic coding57.4%58.2%57.9%66.4%
PostTrainBenchML engineering45.3%44.3%40.2%49.3%
Terminal-Bench Science 0.1Science and math57.6%68.1%52.6%63.3%
LABBench 2Science and math88.8%85.4%68.6%73.1%
RiemannBenchScience and math76%72%65.6%69.6%
GraphWalksLong context · Up to 128k, BFS (F1)99.7%98.7%91.4%90.6%
GraphWalksLong context · 256k to 1M, BFS (F1)84.2%71.8%65%66.8%
Agent's Last ExamComputer use · Pass rate39.5%34.2%–38.2%
OSWorld-2.0Computer use · Offline subset Partial score69.2%72.6%––
ChartographyMultimodal understanding71.6%71%46.2%66.3%
LVBenchMultimodal understanding91.7%87.5%79.7%83.7%
CWE-bench v1Cybersecurity68%68%58%67%
View Google DeepMind's original charts
Google DeepMind: Gemini 4 Argon DeepSWE v1.1 scores compared with GPT-6 Astra, Claude Fable 5.1 and Claude Opus 5.5.
Original chart published by Google DeepMind, shown for reference.
Google DeepMind: Gemini 4 Argon Vals Index scores compared with GPT-6 Astra, Claude Fable 5.1 and Claude Opus 5.5.
Original chart published by Google DeepMind, shown for reference.
Google DeepMind: Gemini 4 Argon Vals Finance Agent v2 scores compared with GPT-6 Astra, Claude Fable 5.1 and Claude Opus 5.5.
Original chart published by Google DeepMind, shown for reference.
Google DeepMind: Gemini 4 Argon Harvey's Legal Agent Benchmark scores compared with GPT-6 Astra, Claude Fable 5.1 and Claude Opus 5.5.
Original chart published by Google DeepMind, shown for reference.
Google DeepMind: Gemini 4 Argon AutomationBench scores compared with GPT-6 Astra, Claude Fable 5.1 and Claude Opus 5.5.
Original chart published by Google DeepMind, shown for reference.
Google DeepMind: Gemini 4 Argon CWE-bench v1 scores.
Original chart published by Google DeepMind, shown for reference.
  • Bold marks the best reported score in each row.
  • Exact scores from Google DeepMind's performance table. Argon is initially available to trusted testers.
  • OSWorld-2.0 uses the offline subset and partial scoring. GraphWalks reports BFS F1, with separate context ranges.
  • Scores reflect Google's published evaluation setup. Benchmark versions and harnesses can differ from other labs' launch results. A dash means no result was published.
Source: Gemini 4 Argon performance and methodology (Google DeepMind)