Head to head
Google: Gemini 3.8 Flash vs Qwen: Qwen3.8 Max
Live prices, context windows, capabilities, and measured quality — updated continuously.
Side by side
| Metric | Google: Gemini 3.8 Flash Google Ai Studio | Qwen: Qwen3.8 Max Alibaba |
|---|---|---|
| Input $/M tokens | $0.750 | $2.00 |
| Output $/M tokens | $3.75 | $6.00 |
| Context window | 1.0M | 1.0M |
| Max output | 66K | 131K |
| WM Score | 88 | 87 |
| Benchmark quality | 86.9 | 88.9 |
| Hosts serving it | 2 | 1 |
| Throughput (tok/s) | 149 | 41 |
| Latency (ms) | 2460 | 1504 |
| Vision | ||
| Tools | ||
| Reasoning | ||
| JSON mode |
Which should you pick?
Google: Gemini 3.8 Flash costs less per million input tokens — about 2.7× cheaper than Qwen: Qwen3.8 Max. Qwen: Qwen3.8 Max scores higher on measured quality. Google: Gemini 3.8 Flash takes a larger context window, which matters for long documents or big codebases. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.
