Unofficial guide. Not affiliated with Meta.

Last updated:

Gemini 4 Argon: price, access and benchmarks

Gemini 4 Argon is Google DeepMind's new frontier model and the first of the Gemini 4 generation. Google announced it on September 30, 2026, but only a small group of security partners can use it so far.

Short answer On Google's own table Argon leads on 12 of 18 benchmarks and ties on 1, with big gaps on legal, finance and long-context work. For terminal coding agents, Claude Opus 5.5 still scores higher. The launch price of $2 / $10 is half of Opus 5.5, but most developers can't call Argon yet.

Key facts

MakerGoogle DeepMind
AnnouncedSeptember 30, 2026
Who can use it nowSecurity teams in Google's Fairwind Program
Next to get itPaid Gemini API customers and Google AI Ultra subscribers, "as soon as possible" per Google (no date)
Launch price per 1M$2 input, $10 output, $0.10 cached input
Standard price per 1M$4 input, $20 output, after the launch period
Max output1M tokens per response (Gemini models were capped at 64K before)
Strong atAgentic coding, legal and finance work, long context, charts and long video, finding and patching security bugs

Source: Google: Gemini 4 Argon announcement

Gemini 4 Argon API pricing

PriceInputCached inputOutput
Launch price$2$0.10$10
Standard price$4$0.20$20

Source: Google: Gemini 4 Argon announcement

How to get Gemini 4 Argon

Benchmarks (Google's numbers)

Google's table from the Gemini model page. Google ran every model at its highest thinking setting. The best score in each row is in bold.

BenchmarkGemini 4 ArgonOpus 5.5GPT-6 AstraFable 5.1
Vals Index68.9%67.0%63.1%65.8%
AutomationBench51.3%42.5%41.4%31.4%
Vals Finance Agent v265.4%58.6%53.5%58.9%
Harvey's Legal Agent Benchmark19.6%3.8%5.4%6.7%
DeepSWE v1.177.9%74.2%74.1%67.4%
FrontierSWE v255.0%62.3%65.5%56.3%
Vibe Code Bench91.9%90.3%89.6%90.3%
Terminal-Bench 4.057.4%66.4%58.2%57.9%
PostTrainBench45.3%49.3%44.3%40.2%
Terminal-Bench Science 0.157.6%63.3%68.1%52.6%
LABBench 288.8%73.1%85.4%68.6%
RiemannBench76.0%69.6%72.0%65.6%
GraphWalks BFS (up to 128K)99.7%90.6%98.7%91.4%
GraphWalks BFS (256K-1M)84.2%66.8%71.8%65.0%
Agent's Last Exam39.5%38.2%34.2%—
OSWorld-2.069.2%—72.6%—
Chartography71.6%66.3%71.0%46.2%
LVBench91.7%83.7%87.5%79.7%
CWE-bench v168.0%67.0%68.0%58.0%

Biggest leads: Harvey's legal agent benchmark (19.6% vs 3.8% for Opus 5.5), AutomationBench (51.3% vs 42.5%), long-context GraphWalks from 256K to 1M (84.2% vs 66.8%) and DeepSWE (77.9% vs 74.2%).

Where it trails: Terminal-Bench 4.0 (57.4% vs 66.4% for Opus 5.5), FrontierSWE v2 (55.0% vs 65.5% for GPT-6 Astra), Terminal-Bench Science and OSWorld-2.0.

Source: Google DeepMind: Gemini model page

Source: Google DeepMind: evaluation methodology (PDF)

Cost per coding task

List prices with our calculator's assumptions, per agent task:

Codebase sizeArgon (launch)Argon (standard)Opus 5.5GPT-6 Astra
Small (< 10k lines)$0.112$0.223$0.223$0.576
Medium (10k-100k lines)$0.249$0.498$0.498$1.290
Large / monorepo (> 100k lines)$0.594$1.188$1.188$3.090

At 173 tasks a month on a medium codebase: about $43 on Argon at the launch price ($86 at the standard price), $86 on Claude Opus 5.5 and $223 on GPT-6 Astra. These use the same token counts for every model; real usage differs by model.

Gemini 4 Argon vs Opus 5.5 Gemini 4 Argon vs GPT-6 Astra All model prices

What the press says

FAQ

Is Gemini 4 Argon out?

It was announced on September 30, 2026, but only members of Google's Fairwind cyber-defense program can use it. Paid API customers and Google AI Ultra subscribers are next, with no date given.

How much does Gemini 4 Argon cost?

$2 per 1M input tokens and $10 per 1M output tokens at launch, with cached input at $0.10. After the launch period it goes to $4 / $20.

Can I use Gemini 4 Argon for free?

No. Google has only named paid routes: the paid Gemini API and the Google AI Ultra subscription, and neither is open yet.

Is Gemini 4 Argon better than Claude Opus 5.5?

On Google's table it beats Opus 5.5 on most benchmarks, including DeepSWE (77.9% vs 74.2%). Opus 5.5 is ahead on Terminal-Bench 4.0 (66.4% vs 57.4%). See our full comparison.

What is the Fairwind Program?

Google's early-access scheme for trusted cyber defenders, such as governments, critical infrastructure and security labs. They get Argon without its cyber guardrails for defensive security work.

Independent page. Not affiliated with Google. Benchmarks are Google's own; we have not run our own tests.