Gemini 4 Argon API
Coming soon, tracking availabilityGoogle · text
Google DeepMind's new frontier model for long-horizon software engineering, legal and finance work, and cybersecurity defence, with a 1M-token output limit. Trusted testers only for now.
Overview
About Gemini 4 Argon
Gemini 4 Argon is Google DeepMind's new frontier model, built for long-horizon software engineering, enterprise knowledge work like legal and finance, and cybersecurity defence. Google announced it on 30 September 2026 and is rolling it out first to trusted cyber defenders in its Fairwind Program; developers get access later, starting with paid API customers. Its output limit is 1M tokens, up from 64K, so one call can reason and write at very long length. The announced introductory price is $2 input / $10 output per million tokens, with cached input 95% off. This page tracks availability, and the model goes live on the same CompanyFabric key the day public API access ships. Until then, Gemini 3.1 Pro is the closest Google model you can call. Every model on CompanyFabric shares the same API key, billing balance, and rate-limit envelope, one integration covers the entire library.
Facts
At a glance
Availability
Announced
API pending, this page tracks it
Pricing
TBA
Estimated from the provider's current line
Capabilities
tools, vision
Commercial use
See license
Proprietary. Access is limited to trusted cyber defenders and testers in Google's Fairwind Program; no public API terms yet.
Variants
All endpoints
Gemini 4 Argon
Coming soon$2/Mtok inGoogle DeepMind's new frontier model for long-horizon software engineering, legal and finance work, and cybersecurity defence, with a 1M-token output limit. Trusted testers only for now.
Chat Completions
Gemini 3.7 Flash
$0.75/Mtok inGoogle's current fast model, million-token context.
Chat Completions
Gemini 3.5 Flash
$1.5/Mtok inThe previous stable Flash, still current for most work.
Chat Completions
Gemini 2.5 Pro
Retired$1.25/Mtok inGoogle's stable Pro tier, the non-preview reasoning model.
Chat Completions
Gemini 2.5 Flash
Retired$0.3/Mtok inLong-context workhorse on the stable 2.5 line.
Chat Completions
Gemini 3.1 Pro (preview)
Preview$2/Mtok inGoogle's newest Pro tier, shipped as a preview.
Chat Completions
Gemini 3 Flash (preview)
Preview$0.5/Mtok inThe Gemini 3 Flash preview line.
Chat Completions
Gemini 3.8 Flash
$0.75/Mtok inGoogle's most capable Flash model, built for long-horizon software engineering and autonomous agents.
Chat Completions
Gemini 3.6 Flash
$0.75/Mtok inThe previous-generation Flash, balancing speed and multimodal work across everyday agentic tasks.
Chat Completions
Gemini 3.5 Flash-Lite
$0.3/Mtok inCost-efficient Gemini for high-volume agentic work, translation and simple data processing.
Chat Completions
Gemini 3.1 Flash-Lite
$0.25/Mtok inGoogle's cheapest text model here, for high-volume classification and extraction.
Chat Completions
Quickstart
How to use the Gemini 4 Argon API
- 1
Create a free CompanyFabric account, starter credits are included, no card required.
- 2
Copy your API key from the dashboard. One key covers every model in the library.
- 3
Point any OpenAI SDK at api.companyfabric.com/v1 and set model to "google/gemini-4-argon".
- 4
Read the exact metered cost of the call from the response, the same number shown in the playground.
Use cases
What teams build with it
Agents & copilots
Drive agent loops, tool calls, and user-facing assistants over the OpenAI-compatible endpoint.
agentstool useassistants
Content pipelines
Summarization, rewriting, and structured generation at metered per-token cost.
contentsummarization
Extraction & classification
Turn unstructured input into typed JSON with schema-constrained output.
extractionjson mode
Get better output
Prompting tips
- State the output format before the task, models follow contracts declared early.
- Few-shot beats description: two worked examples outperform a paragraph of rules.
- Cap output length explicitly when cost matters; output tokens dominate most bills.
Exact, metered, prepaid
Gemini 4 Argon pricing
Gemini 4 Argon pricing is per token: $2 / M input · $10 / M output · $0.1 / M cached input. Output tokens cost 5× input tokens, so long generations, not long prompts, are what actually drive most bills. A typical call with a 2k-token prompt and a 1k-token answer costs about $0.0140.
Every call is metered exactly and attributed to your key, the dashboard, the API usage response, and your invoice all read from the same ledger row. Cap any key's monthly spend with a budget cap, and bring your own provider key (BYOK) at 0% platform fee if you already have a direct contract.
| Endpoint | Type | Price |
|---|---|---|
| Gemini 4 Argon | Chat | $2 / M input · $10 / M output · $0.1 / M cached input |
| Gemini 3.7 Flash | Chat | $0.75 / M input · $3.75 / M output |
| Gemini 3.5 Flash | Chat | $1.5 / M input · $9 / M output |
| Gemini 2.5 Pro | Chat | $1.25 / M input · $10 / M output |
| Gemini 2.5 Flash | Chat | $0.3 / M input · $2.5 / M output |
| Gemini 3.1 Pro (preview) | Chat | $2 / M input · $12 / M output |
| Gemini 3 Flash (preview) | Chat | $0.5 / M input · $3 / M output |
| Gemini 3.8 Flash | Chat | $0.75 / M input · $3.75 / M output |
| Gemini 3.6 Flash | Chat | $0.75 / M input · $3.75 / M output |
| Gemini 3.5 Flash-Lite | Chat | $0.3 / M input · $2.5 / M output |
| Gemini 3.1 Flash-Lite | Chat | $0.25 / M input · $1.5 / M output |
Prices are estimates pending the provider's official announcement.
Comparisons
Gemini 4 Argon vs the alternatives
Gemini 4 Argon vs Claude Opus 5.5
Gemini 4 Argon and Claude Opus 5.5 compete for the same text workloads. Both run behind the same CompanyFabric key at exact metered prices, A/B them on your real prompts and let the outputs decide, then switch with a one-line change.
Full comparison: Gemini 4 Argon vs Claude Opus 5.5 →Gemini 4 Argon vs GPT-6 Astra
Gemini 4 Argon and GPT-6 Astra compete for the same text workloads. Both run behind the same CompanyFabric key at exact metered prices, A/B them on your real prompts and let the outputs decide, then switch with a one-line change.
Full comparison: Gemini 4 Argon vs GPT-6 Astra →Gemini 4 Argon vs Claude Fable 5.1
Gemini 4 Argon and Claude Fable 5.1 compete for the same text workloads. Both run behind the same CompanyFabric key at exact metered prices, A/B them on your real prompts and let the outputs decide, then switch with a one-line change.
Full comparison: Gemini 4 Argon vs Claude Fable 5.1 →Compare
Gemini 4 Argon head to head
FAQ
Frequently asked questions
Is there a Gemini 4 Argon API?
Not yet, Gemini 4 Argon is announced but API access is not public. This page tracks availability and updates the moment access ships. Until then, Gemini 3.7 Flash is the closest live model and runs today with the same key.
When will the Gemini 4 Argon API launch?
Google has not committed to a public date. Create a free key now and you can call Gemini 4 Argon on day one, no waitlist, same endpoint.
Why use Gemini 4 Argon on CompanyFabric instead of going direct?
One key, one prepaid balance, and one rate-limit envelope across the whole library, at pricing at parity with or below the direct API. Model switching is a one-line change, and BYOK routing is 0% fee.
Is the API OpenAI-compatible?
Yes, point any OpenAI SDK at api.companyfabric.com/v1 and set model to "google/gemini-4-argon".
Can I use the output commercially?
Proprietary. Access is limited to trusted cyber defenders and testers in Google's Fairwind Program; no public API terms yet.
How does billing work?
Prepaid credits via card, or BYOK with your own provider keys at 0% platform fee. No subscription required; balances never expire.
How much will the Gemini 4 Argon API cost?
Google announced an introductory price of $2 per million input tokens and $10 per million output tokens, with cached input tokens 95% off ($0.10 per million). That is the same headline rate as Claude Sonnet 5.5 and half of Claude Opus 5.5. It is not billable anywhere yet: there is no public API.
How does Gemini 4 Argon compare with Claude Opus 5.5 and GPT-6 Astra?
On Google DeepMind's published table it leads on knowledge work (Vals Index 68.9% against 67.0% for Opus 5.5 and 63.1% for GPT-6 Astra), long context (GraphWalks 256K to 1M, 84.2% against 66.8% and 71.8%), DeepSWE v1.1 (77.9%) and long video (LVBench 91.7%). It trails on Terminal-bench 4.0 (57.4%, where Opus 5.5 scores 66.4%), FrontierSWE v2 (55.0% against 65.5% for GPT-6 Astra) and OSWorld-2.0 (69.2% against 72.6%). These are Google's numbers, not independent results.
Who can use Gemini 4 Argon today?
Trusted cyber defenders and testers in Google's Fairwind Program. Google says it will widen access to developers, enterprises and consumers once its safeguards are hardened, starting with paid API customers and Google AI Ultra subscribers. No date is public.
Provider
About Google
Google publishes the gemini family. Also in this family: Gemini 3.7 Flash, Gemini 3.5 Flash, Gemini 2.5 Pro, Gemini 2.5 Flash, Gemini 3.1 Pro (preview), Gemini 3 Flash (preview), Gemini 3.8 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.1 Flash-Lite. License: Proprietary. Access is limited to trusted cyber defenders and testers in Google's Fairwind Program; no public API terms yet.
Explore
Related model APIs
Claude Sonnet 4.5
RetiringAnthropic
Frontier-quality text and reasoning through one OpenAI-compatible endpoint.
text
Claude Haiku 4.5
$1/Mtok inAnthropic
Fast, cheap frontier-family model for high-volume tasks.
text
Claude Opus 4.6
$5/Mtok inAnthropic
The frontier tier for the hardest reasoning and agentic work.
text
GPT-5.5
$5/Mtok inOpenAI
OpenAI's flagship, one key away.
text
GPT-6 Astra
$10/Mtok inOpenAI
OpenAI's GPT-6 flagship, a million-token window one key away.
text
Claude Fable 5
$10/Mtok inAnthropic
Anthropic's first Mythos-class model, the new flagship tier above Opus, tuned for the hardest reasoning and agentic work.
text
One key. Every model. Exact prices.
Get a key now and call Gemini 4 Argon the day it ships, Gemini 3.7 Flash runs today. Free starter credits included, no subscription required.