- Home
- Gemini 4 Argon
Gemini 4 Argon: API Pricing, Context Window and Specs
Gemini 4 Argon is Google’s new frontier model, announced on September 30, 2026. Google aims it at long, multi-step work in software engineering, enterprise knowledge work such as legal and finance, and cybersecurity defense, and raised its output limit to 1 million tokens, up from 64K.
Argon is not yet available in the API. Google is giving access first to trusted cyber defenders and testers, then to paid API customers and Google AI Ultra subscribers. It will launch at an introductory price of $2 per million input tokens and $10 per million output tokens, rising to $4 and $20 once the introductory period ends.
- Input
- $2
- Output
- $10
- Cached input
- $0.10
- Cache write
- Not announced
- Batch
- Not announced
- Long context
- Not announced
USD per 1M tokens at the introductory price, which later rises to $4 / $20. Source: Google announcement, checked Oct 2, 2026.
Quick estimate
Gemini 4 Argon pricing details
- Introductory price: $2 input and $10 output per 1M tokens. Google has not said how long the introductory period lasts.
- After the introductory period: $4 input and $20 output per 1M tokens.
- Cached input is 95% cheaper than regular input: $0.10 per 1M tokens at the introductory price. Google has not published the cached rate that applies after the introductory period.
- Batch pricing, cache storage fees, long-context pricing and the context window have not been announced.
What Gemini 4 Argon costs in practice
Standard-tier cost with no caching, for a few common request shapes.
| Workload | Tokens in / out | Per request | Per 10,000 requests |
|---|---|---|---|
| Chatbot | 500 / 300 | $0.004 | $40.00 |
| RAG Q&A | 4,000 / 500 | $0.013 | $130 |
| Coding agent | 30,000 / 2,000 | $0.08 | $800 |
| Summarization | 8,000 / 600 | $0.022 | $220 |
Gemini 4 Argon specs
- Context window
- Not announced
- Max output tokens
- 1M
- Released
- Not yet
- Status
- Announced Sep 30, 2026; not yet in the API
- API model ID
- Not announced
- Input types
- Text, images, video and documents
- First access
- Trusted testers, then paid API customers and Google AI Ultra subscribers
From the official announcement.
Gemini 4 Argon vs similar models
| Model | Provider | Input $/1M | Output $/1M | Cached $/1M | Context |
|---|---|---|---|---|---|
| Gemini 4 Argon Google | $2 | $10 | $0.10 | — | |
| Claude Opus 5.5 Anthropic | Anthropic | $4 | $20 | $0.20 | 1M |
| Gemini 3.1 Pro Preview Google | $2 | $12 | $0.20 | 1.05M | |
| GPT-6.1 Sol OpenAI | OpenAI | $2 | $10 | $0.10 | 1.05M |
| GPT-6 Astra OpenAI | OpenAI | $10 | $50 | $1 | 1.05M |
| Claude Sonnet 5.5 Anthropic | Anthropic | $2 | $10 | $0.20 | 1M |
| Grok 4.7 xAI | xAI | $2 | $6 | $0.50 | 500K |
Gemini 4 Argon FAQ
Gemini 4 Argon will launch at an introductory price of $2 per million input tokens and $10 per million output tokens, with cached input at $0.10 per million. After the introductory period, the price rises to $4 input and $20 output per million tokens.
Google has not given a date. Argon is rolling out first to trusted cyber defenders and testers. Google says paid API customers and Google AI Ultra subscribers will be first in line when access widens.
Google has not said. The announcement only states that $4 per million input tokens and $20 per million output tokens apply once the introductory period ends.
Up to 1 million output tokens in a single response, up from 64K on earlier Gemini models. Google has not announced the context window yet.
Related tools
- LLM API pricing tableEvery model's input, output and cached rates in one sortable table.
- LLM Price ComparisonPut two to four models side by side: per-token rates, context window, discounts and what the same job costs on each.
- Cheapest LLM APIsRanked lists of the lowest-priced models overall, among flagships, with 1M-token context, with image input and with open weights.
- Token Cost CalculatorEnter a token count and see what that many input and output tokens cost on every model.
- LLM Cost CalculatorPick a workload and request volume, and compare the bill across models with caching and batch discounts.