AI Trend Notifier
EN

$ diff gemini-3-8-flash hy4-preview

Gemini 3.8 Flash vs Hy4 preview

Values come from Gemini 3.8 Flash and Hy4 preview, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.

What actually differs

Context window
Gemini 3.8 Flash takes 1.0M against 1M — modestly more room in a single request.
Input price
Gemini 3.8 Flash at $0.75/M against $0.834/M — 1.1× cheaper to feed.
Output price
Hy4 preview at $2.501/M against $3.75/M — 1.5× cheaper to generate. Output dominates the bill on most agentic workloads, where the model writes far more than it reads.
Weights
Hy4 preview publishes weights (Apache-2.0); Gemini 3.8 Flash is API-only. That decides self-hosting, air-gapped deployment and fine-tuning before any capability question does.
Recency
Gemini 3.8 Flash shipped 5 days after Hy4 preview (2026-09-02 vs 2026-08-28).

Full spec

AttributeGemini 3.8 FlashHy4 preview
DeveloperGoogle DeepMindTencent
Released2026-09-022026-08-28
Context window1,048,576 tokens (65,536 max output)over 1,000,000
Pricing$0.75/M input · $3.75/M output (introductory, through 2026-12-31) · $1.50/M · $7.50/M from 2027-01-01$0.834/M input · $2.501/M output
LicenseproprietaryApache-2.0
AvailabilityGemini API, Google AI Studio, Antigravity, Android Studio, Gemini Enterprise / Gemini Enterprise Agent PlatformHugging Face (BF16 + FP8), ModelScope, GitCode, CNB, Tencent Cloud TokenHub, OpenRouter, WorkBuddy, CodeBuddy, Yuanbao, ima

How these pages are produced

← all comparisons