Skip to content
Custom Comparison

Compare 2 AI Models

Qwen3.8 vs Gemini 3.6 Flash

Qwen3.8Alibaba
Gemini 3.6 FlashGoogle

Up to 3 models. Every selection is in the address bar, so the comparison is a link you can send.

Interactive Comparison

Loading chart...
$2 / $6 per million tokens
$0.75 / $3.75 per million tokens
Identity
DeveloperAlibabaGoogle
ReleasedAug 2026Jul 2026
StatusGA—
LicenceOpen weights under a licence named "qwen3.8-max" - NOT Apache 2.0 (the Qwen3.8-27B sibling is Apache 2.0, which is a common source of confusion)Proprietary
Self-hostableYes - Qwen/Qwen3.8-2.4T-A95B and an FP8 checkpoint published 12 Aug 2026No
Cost
Blended $/1M tokensinput × 0.75 + output × 0.25$3.000 / 1M tokens$1.500 / 1M tokens
Input price$2.000 / 1M tokens$0.750 / 1M tokens
Output price$6.000 / 1M tokens$3.750 / 1M tokens
Cached inputNot documented—
Batch discountNot documented—
Free tierNot documented—
Capacity
Context window262k native (extensible to ~1.01M)1M tokens
Max output128k tokens—
Long-context surchargeNot documented—
Capability
Vision inNo - text-only. Alibaba’s model card states "Multimodal inputs are not supported". The multimodal sibling is Qwen3.8-27B (Apache 2.0), which is tagged Image-Text-to-Text and does accept images and video.Yes
Audio inNo—
Function callingYes—
Structured outputYes—
Extended reasoningYes - thinking mode is required for all interactions, not optional—
Web searchNot documented—
Code executionNot documented—
Access
APIYes—
Chat appYes—
Cloud marketplacesAlibaba Cloud Model Studio—
Fine-tuningYes - open weights permit it—
Measured quality
LMArena Elo1483 (checked Oct 2026)1482 (checked Oct 2026)
LMArena rankRank 22Rank 23
GPQA Diamond—94
Elo per dollarLMArena Elo ÷ blended $/1M tokens494988