$ diff qwen-3-8 inkling
Qwen 3.8 Max vs Inkling
Values come from Qwen 3.8 Max and Inkling, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.
Alibaba QwenQwen 3.8 Max
- context
- 1M
- weights
- open
- $/M in
- $2
- $/M out
- $6
- context
- 1M
- weights
- open
- $/M in
- —
- $/M out
- —
What actually differs
- Context window
- Identical — both accept 1M tokens, so context length is not a reason to pick either.
- Weights
- Both publish weights — Qwen 3.8 Max under Open-weight announced for the week of 2026-08-10; not yet shipped, Inkling under Apache 2.0 (open-weight). Check the licences rather than assuming they permit the same commercial use.
- Recency
- Qwen 3.8 Max shipped 19 days after Inkling (2026-08-03 vs 2026-07-15).
Full spec
| Attribute | Qwen 3.8 Max | Inkling |
|---|---|---|
| Developer | Alibaba Qwen | Thinking Machines |
| Released | 2026-08-03 | 2026-07-15 |
| Context window | 1 million tokens | 1M |
| Pricing | $2 / M input · $6 / M output | not recorded |
| License | Open-weight announced for the week of 2026-08-10; not yet shipped | Apache 2.0 (open-weight) |
| Availability | Alibaba Cloud Model Studio (global), QwenWork | Hugging Face (BF16 and NVFP4), Tinker (fine-tuning), SGLang, vLLM, llama.cpp, Unsloth |