$ diff gemini-4-argon inkling
Gemini 4 Argon vs Inkling
Values come from Gemini 4 Argon and Inkling, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.
Google DeepMindGemini 4 Argon
- context
- 2M
- weights
- closed
- $/M in
- $2
- $/M out
- $10
- context
- 1M
- weights
- open
- $/M in
- —
- $/M out
- —
What actually differs
- Context window
- Gemini 4 Argon takes 2M against 1M — 2× more room in a single request.
- Weights
- Inkling publishes weights (Apache 2.0 (open-weight)); Gemini 4 Argon is API-only. That decides self-hosting, air-gapped deployment and fine-tuning before any capability question does.
- Recency
- Gemini 4 Argon shipped 77 days after Inkling (2026-09-30 vs 2026-07-15).
Full spec
| Attribute | Gemini 4 Argon | Inkling |
|---|---|---|
| Developer | Google DeepMind | Thinking Machines |
| Released | 2026-09-30 | 2026-07-15 |
| Context window | 2M tokens | 1M |
| Pricing | $2/M input · $10/M output · cached input $0.10/M (introductory; stated to double to $4/$20) | not recorded |
| License | proprietary | Apache 2.0 (open-weight) |
| Availability | Fairwind Program (trusted cyber defenders) only; paid API and Google AI Ultra stated as next, no date | Hugging Face (BF16 and NVFP4), Tinker (fine-tuning), SGLang, vLLM, llama.cpp, Unsloth |