Google DeepMind · launched 30 Sept 2026
Gemini 4 Argon
Limited access
Google's newest frontier model for long-running coding, knowledge work and cybersecurity defense, initially available to trusted testers.
Read the official announcement- Vendor
- Google DeepMind
- Launch date
- 30 Sept 2026
- Input price
- $2.00 / 1M tokens
- Output price
- $10 / 1M tokens
- Cached input
- $0.10 / 1M tokens
- Batch discount
- Not published
- Context window
- Not published
- Max output
- 1M tokens
- Input types
- Text, Image, Video, File
- Availability
- Limited access
Announced introductory rates for the upcoming broader launch. Regular rates will be $4 input and $20 output per million tokens; the introductory end date and batch discount are not published.
Official pricing source · checked 8 Oct 2026
Price compared with other models
| Model | Launched | Input / 1M | Output / 1M | Cached / 1M | Blended / 1M | vs Gemini 4 Argon |
|---|---|---|---|---|---|---|
| Claude Haiku 5.5AnthropicListed prices and comparisons apply to prompts up to 100K tokens. Above 100K, input costs $0.50, output $2.50 and cache reads $0.05 per million tokens. | 7 Oct 2026 | $0.10 | $0.50 | $0.01 | $0.20 | 95% cheaper |
| GPT-6 LunaOpenAI | 22 Sept 2026 | $0.10 | $0.50 | $0.01 | $0.20 | 95% cheaper |
| GPT-5.6 LunaOpenAI | 9 Jul 2026 | $0.20 | $1.20 | $0.02 | $0.45 | 89% cheaper |
| Gemini 3.5 Flash-LiteGoogle DeepMindText, image, video and audio inputs share the listed rate. Cache storage and grounding are billed separately. | 21 Jul 2026 | $0.30 | $2.50 | $0.03 | $0.85 | 79% cheaper |
| Gemini 3.8 FlashGoogle DeepMindIntroductory rates through 31 December 2026. From 1 January 2027, input is $1.50, output $7.50 and cached input $0.15 per million tokens. Cache storage and grounding are billed separately. | 2 Sept 2026 | $0.75 | $3.75 | $0.075 | $1.50 | 63% cheaper |
| Claude Haiku 4.5Anthropic | 15 Oct 2025 | $1.00 | $5.00 | $0.10 | $2.00 | 50% cheaper |
| Grok 4.7SpaceXAIBase rates apply up to 200K context tokens; higher-context requests use different rates. Batch API is not supported. The separate fast variant costs twice as much. | 21 Sept 2026 | $2.00 | $6.00 | $0.50 | $3.00 | 25% cheaper |
| Claude Sonnet 5Anthropic | 30 Jun 2026 | $2.00 | $10 | $0.20 | $4.00 | Same price |
| Claude Sonnet 5.5AnthropicCache reads were reduced from $0.20 to $0.10 per million tokens on 7 October 2026. | 28 Sept 2026 | $2.00 | $10 | $0.10 | $4.00 | Same price |
| Gemini 4 ArgonGoogle DeepMind · Limited accessAnnounced introductory rates for the upcoming broader launch. Regular rates will be $4 input and $20 output per million tokens; the introductory end date and batch discount are not published. | 30 Sept 2026 | $2.00 | $10 | $0.10 | $4.00 | Baseline |
| GPT-6 SolOpenAIListed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.40 and output $15 per million tokens for the full request. | 22 Sept 2026 | $2.00 | $10 | $0.20 | $4.00 | Same price |
| GPT-6.1 SolOpenAIListed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.20 and output $15 per million tokens for the full request. | 29 Sept 2026 | $2.00 | $10 | $0.10 | $4.00 | Same price |
| Gemini 3.1 ProGoogle DeepMind · PreviewRates apply to prompts up to 200K tokens. Above 200K, input is $4, output $18 and cached input $0.40 per million tokens. Cache storage and grounding are billed separately. | 19 Feb 2026 | $2.00 | $12 | $0.20 | $4.50 | 13% more |
| Claude Opus 5.5Anthropic | 22 Sept 2026 | $4.00 | $20 | $0.20 | $8.00 | 2× the price |
| GPT-5.6 SolOpenAIPrice before the GPT-6 launch, from OpenAI's announcement. OpenRouter now lists $2 / $10. | 9 Jul 2026 | $4.00 | $20 | – | $8.00 | 2× the price |
| Claude Opus 5Anthropic | 24 Jul 2026 | $5.00 | $25 | $0.50 | $10 | 2.5× the price |
| Claude Fable 5Anthropic | 9 Jun 2026 | $10 | $50 | $1.00 | $20 | 5× the price |
| Claude Fable 5.1Anthropic | 1 Sept 2026 | $10 | $50 | $0.25 | $20 | 5× the price |
| GPT-6 AstraOpenAI | 4 Sept 2026 | $10 | $50 | $1.00 | $20 | 5× the price |
GPT-6.1 Sol and GPT-6 Sol rates were checked on 8 October 2026 against OpenAI model documentation. USD per million tokens, checked 2 Oct 2026 via OpenRouter. Comparison uses the blended price (3 input tokens for every output token). Haiku 5.5 and Sonnet 5.5 prices were updated from Anthropic's 7 October 2026 announcement. Gemini and Grok prices were checked against official Google and xAI documentation on 8 October 2026.
Official benchmark results
Reported by Google DeepMind on 30 Sept 2026 · exact values
Exact published scores · Higher is better. Unreported results are omitted.
| Benchmark | Gemini 4 Argon | GPT-6 Astra | Claude Fable 5.1 | Claude Opus 5.5 |
|---|---|---|---|---|
| Vals IndexKnowledge work | 68.9% | 63.1% | 65.8% | 67% |
| AutomationBenchKnowledge work · Score | 51.3% | 41.4% | 31.4% | 42.5% |
| Vals Finance Agent v2Knowledge work | 65.4% | 53.5% | 58.9% | 58.6% |
| Harvey's Legal Agent BenchmarkKnowledge work | 19.6% | 5.4% | 6.7% | 3.8% |
| DeepSWE v1.1Agentic coding | 77.9% | 74.1% | 67.4% | 74.2% |
| FrontierSWE v2Agentic coding | 55% | 65.5% | 56.3% | 62.3% |
| Vibe Code BenchAgentic coding | 91.9% | 89.6% | 90.3% | 90.3% |
| Terminal-bench 4.0Agentic coding | 57.4% | 58.2% | 57.9% | 66.4% |
| PostTrainBenchML engineering | 45.3% | 44.3% | 40.2% | 49.3% |
| Terminal-Bench Science 0.1Science and math | 57.6% | 68.1% | 52.6% | 63.3% |
| LABBench 2Science and math | 88.8% | 85.4% | 68.6% | 73.1% |
| RiemannBenchScience and math | 76% | 72% | 65.6% | 69.6% |
| GraphWalksLong context · Up to 128k, BFS (F1) | 99.7% | 98.7% | 91.4% | 90.6% |
| GraphWalksLong context · 256k to 1M, BFS (F1) | 84.2% | 71.8% | 65% | 66.8% |
| Agent's Last ExamComputer use · Pass rate | 39.5% | 34.2% | – | 38.2% |
| OSWorld-2.0Computer use · Offline subset Partial score | 69.2% | 72.6% | – | – |
| ChartographyMultimodal understanding | 71.6% | 71% | 46.2% | 66.3% |
| LVBenchMultimodal understanding | 91.7% | 87.5% | 79.7% | 83.7% |
| CWE-bench v1Cybersecurity | 68% | 68% | 58% | 67% |
View Google DeepMind's original charts






- Bold marks the best reported score in each row.
- Exact scores from Google DeepMind's performance table. Argon is initially available to trusted testers.
- OSWorld-2.0 uses the offline subset and partial scoring. GraphWalks reports BFS F1, with separate context ranges.
- Scores reflect Google's published evaluation setup. Benchmark versions and harnesses can differ from other labs' launch results. A dash means no result was published.