$ diff gemini-3-7-flash grok-4-6
Gemini 3.7 Flash vs Grok 4.6
Values come from Gemini 3.7 Flash and Grok 4.6, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.
Google DeepMindGemini 3.7 Flash
- context
- 1M
- weights
- closed
- $/M in
- $0.75
- $/M out
- $3.75
- context
- 500K
- weights
- closed
- $/M in
- $2
- $/M out
- $6
What actually differs
- Context window
- Gemini 3.7 Flash takes 1M against 500K — 2× more room in a single request.
- Input price
- Gemini 3.7 Flash at $0.75/M against $2/M — 2.7× cheaper to feed. Standard rates: one of these labs also quotes a lower cached-input tier, which applies only when a prefix is reused — see the full spec below.
- Output price
- Gemini 3.7 Flash at $3.75/M against $6/M — 1.6× cheaper to generate. Output dominates the bill on most agentic workloads, where the model writes far more than it reads.
- Recency
- Gemini 3.7 Flash shipped 6 days after Grok 4.6 (2026-08-13 vs 2026-08-07).
Full spec
| Attribute | Gemini 3.7 Flash | Grok 4.6 |
|---|---|---|
| Developer | Google DeepMind | xAI |
| Released | 2026-08-13 | 2026-08-07 |
| Context window | 1,000,000 tokens (64,000 max output) | 500,000 tokens |
| Pricing | $0.75/M input · $3.75/M output (introductory, through 2026-12-31) · $1.50/M · $7.50/M from 2027-01-01 | $2/M input · $6/M output; a request reaching 200K tokens re-prices in full at $4.00/M input · $1.00/M cached input · $12.00/M output |
| License | proprietary | proprietary |
| Availability | Gemini API, Google AI Studio, Android Studio, Google Antigravity, Gemini Enterprise Agent Platform, Gemini Enterprise app, Spark in the Gemini app (AI Pro and Ultra) | API |