Google's Gemini 3.8 Flash and Alibaba's Qwen3.8-Max compared on specs, pricing, and capabilities — as configured in AI Crucible ensembles.
| Specification | Gemini 3.8 Flash | Qwen3.8-Max |
|---|---|---|
| Provider | Alibaba | |
| API model ID | gemini-3.8-flash | qwen3.8-max |
| Description | Google's most intelligent Flash model, engineered for long-horizon software engineering with 1M context. | Alibaba's flagship Qwen3.8 model with 1M context, native thinking, and vision, at lower cost than Qwen3.7-Max. |
| Context window | 1M tokens | 1M tokens |
| Max output tokens | 65,536 | 131,072 |
| Input cost ($/1M tokens) | $0.90 | $2.40 |
| Output cost ($/1M tokens) | $4.50 | $7.20 |
| Cache read cost ($/1M tokens) | $0.09 | $0.30 |
| Latency | low | medium |
| Reasoning model | Yes | Yes |
| Vision (image input) | Yes | Yes |
| Tool use | Yes | Yes |
See all model comparisons, full model specifications, benchmark results, or model comparison articles.