Custom Comparison
Compare 2 AI Models
DeepSeek V4.1-Flash vs Gemini 3.8 Flash
DeepSeek V4.1-FlashDeepSeek
Gemini 3.8 FlashGoogle
Up to 3 models. Every selection is in the address bar, so the comparison is a link you can send.
Interactive Comparison
Loading chart...
$0.30 / $1.20 per 1M tokens (peak, cache miss); $0.15 / $0.60 off-peak | $0.75 / $3.75 per 1M tokens (introductory, to 31 Dec 2026) | |
|---|---|---|
| Identity | ||
| Developer | DeepSeek | |
| Released | Sep 2026 | Sep 2026 |
| Status | GA | GA |
| Licence | MIT | Proprietary |
| Self-hostable | Yes - open weights on Hugging Face | No |
| Cost | ||
| Blended $/1M tokensinput × 0.75 + output × 0.25 | $0.525 / 1M tokens | $1.500 / 1M tokens |
| Input price | $0.300 / 1M tokens | $0.750 / 1M tokens |
| Output price | $1.200 / 1M tokens | $3.750 / 1M tokens |
| Cached input | $0.006 / 1M tokens (peak); $0.003 off-peak | — |
| Batch discount | Off-peak billing: 50% off all rates | — |
| Capacity | ||
| Context window | 1M tokens | 1M tokens |
| Max output | 384k tokens | 64k tokens |
| Long-context surcharge | None (flat rate across the 1M window) | — |
| Capability | ||
| Vision in | Yes | — |
| Audio in | No | — |
| Function calling | Yes | — |
| Structured output | Yes - JSON output | — |
| Extended reasoning | Yes - thinking (default) and non-thinking modes | Tunable thinking: low, medium (default), high. The "minimal" level is not supported. |
| Web search | No | — |
| Code execution | No | — |
| Access | ||
| API | Yes | Yes |
| Chat app | Yes | Yes |
| Cloud marketplaces | Not offered | — |
| Fine-tuning | Not offered (weights are open) | — |
| Measured quality | ||
| LMArena Elo | 1473 (checked Oct 2026) | 1494 (checked Oct 2026) |
| LMArena rank | Rank 40 | Rank 11 |
| GPQA Diamond | — | 95 |
| Elo per dollarLMArena Elo ÷ blended $/1M tokens | 2806 | 996 |