$ diff hy4-preview gemini-3-7-flash
Hy4 preview vs Gemini 3.7 Flash
Values come from Hy4 preview and Gemini 3.7 Flash, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.
TencentHy4 preview
- context
- 1M
- weights
- open
- $/M in
- $0.834
- $/M out
- $2.501
- context
- 1M
- weights
- closed
- $/M in
- $0.75
- $/M out
- $3.75
What actually differs
- Context window
- Identical — both accept 1M tokens, so context length is not a reason to pick either.
- Input price
- Gemini 3.7 Flash at $0.75/M against $0.834/M — 1.1× cheaper to feed.
- Output price
- Hy4 preview at $2.501/M against $3.75/M — 1.5× cheaper to generate. Output dominates the bill on most agentic workloads, where the model writes far more than it reads.
- Weights
- Hy4 preview publishes weights (Apache-2.0); Gemini 3.7 Flash is API-only. That decides self-hosting, air-gapped deployment and fine-tuning before any capability question does.
- Recency
- Hy4 preview shipped 15 days after Gemini 3.7 Flash (2026-08-28 vs 2026-08-13).
Full spec
| Attribute | Hy4 preview | Gemini 3.7 Flash |
|---|---|---|
| Developer | Tencent | Google DeepMind |
| Released | 2026-08-28 | 2026-08-13 |
| Context window | over 1,000,000 | 1,000,000 tokens (64,000 max output) |
| Pricing | $0.834/M input · $2.501/M output | $0.75/M input · $3.75/M output (introductory, through 2026-12-31) · $1.50/M · $7.50/M from 2027-01-01 |
| License | Apache-2.0 | proprietary |
| Availability | Hugging Face (BF16 + FP8), ModelScope, GitCode, CNB, Tencent Cloud TokenHub, OpenRouter, WorkBuddy, CodeBuddy, Yuanbao, ima | Gemini API, Google AI Studio, Android Studio, Google Antigravity, Gemini Enterprise Agent Platform, Gemini Enterprise app, Spark in the Gemini app (AI Pro and Ultra) |