AI Trend Notifier
EN

$ diff gemini-3-7-flash kimi-k3

Gemini 3.7 Flash vs Kimi K3

Values come from Gemini 3.7 Flash and Kimi K3, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.

What actually differs

Context window
Identical — both accept 1M tokens, so context length is not a reason to pick either.
Input price
Gemini 3.7 Flash at $0.75/M against $3/M — 4.0× cheaper to feed. Standard rates: one of these labs also quotes a lower cached-input tier, which applies only when a prefix is reused — see the full spec below.
Output price
Gemini 3.7 Flash at $3.75/M against $15/M — 4.0× cheaper to generate. Output dominates the bill on most agentic workloads, where the model writes far more than it reads.
Weights
Kimi K3 publishes weights (Modified MIT (commercial use permitted; weights publicly downloadable)); Gemini 3.7 Flash is API-only. That decides self-hosting, air-gapped deployment and fine-tuning before any capability question does.
Recency
Gemini 3.7 Flash shipped 28 days after Kimi K3 (2026-08-13 vs 2026-07-16).

Full spec

AttributeGemini 3.7 FlashKimi K3
DeveloperGoogle DeepMindMoonshot AI
Released2026-08-132026-07-16
Context window1,000,000 tokens (64,000 max output)1M tokens
Pricing$0.75/M input · $3.75/M output (introductory, through 2026-12-31) · $1.50/M · $7.50/M from 2027-01-01$0.30/M cache-hit input · $3.00/M cache-miss input · $15.00/M output
LicenseproprietaryModified MIT (commercial use permitted; weights publicly downloadable)
AvailabilityGemini API, Google AI Studio, Android Studio, Google Antigravity, Gemini Enterprise Agent Platform, Gemini Enterprise app, Spark in the Gemini app (AI Pro and Ultra)Kimi Code, Kimi app, Kimi API; weights on Hugging Face (`moonshot-ai/kimi-k3`, released 2026-07-27)

How these pages are produced

← all comparisons