$ diff fugu-max gemini-3-8-flash
Fugu Max vs Gemini 3.8 Flash
Values come from Fugu Max and Gemini 3.8 Flash, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.
Sakana AIFugu Max
- context
- 1M
- weights
- —
- $/M in
- $2
- $/M out
- $6
- context
- 1.0M
- weights
- closed
- $/M in
- $0.75
- $/M out
- $3.75
What actually differs
- Context window
- Gemini 3.8 Flash takes 1.0M against 1M — modestly more room in a single request.
- Input price
- Gemini 3.8 Flash at $0.75/M against $2/M — 2.7× cheaper to feed.
- Output price
- Gemini 3.8 Flash at $3.75/M against $6/M — 1.6× cheaper to generate. Output dominates the bill on most agentic workloads, where the model writes far more than it reads.
- Recency
- Fugu Max shipped 9 days after Gemini 3.8 Flash (2026-09-11 vs 2026-09-02).
Full spec
| Attribute | Fugu Max | Gemini 3.8 Flash |
|---|---|---|
| Developer | Sakana AI | Google DeepMind |
| Released | 2026-09-11 | 2026-09-02 |
| Context window | 1,000,000 | 1,048,576 tokens (65,536 max output) |
| Pricing | $2/M input · $6/M output | $0.75/M input · $3.75/M output (introductory, through 2026-12-31) · $1.50/M · $7.50/M from 2027-01-01 |
| License | not recorded | proprietary |
| Availability | Sakana API (OpenAI-compatible), OpenRouter, NanoGPT, Kilo Gateway, LLM Gateway | Gemini API, Google AI Studio, Antigravity, Android Studio, Gemini Enterprise / Gemini Enterprise Agent Platform |