Models / Qwen3.6 Max Preview
AlibabaQwen3.6 Max Preview
16 published results from 4 sources. Each card shows where the number comes from and what it does not measure. The overall leaderboard combines them; here each stands alone.
- Provider
- Alibaba
- Sources
- 4
- Our benchmarks
- 0
- Price per million tokens
- $1.03 in · $6.16 out
OpenRouter list price, 1 Oct 2026 · 262,144-token context
Reported by others
1,452
Overall · rank 38 of 177
- Unit
- Arena rating, higher is better
- Range
- 1,445 to 1,460
- Sample
- 5414 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not an error rate: claims that cannot be checked on the web are skipped, and preference still carries most of the weight.
1,450
Business, management and finance · rank 27 of 402
- Unit
- Arena rating, higher is better
- Range
- 1,432 to 1,469
- Sample
- 1073 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,436
Creative writing · rank 18 of 407
- Unit
- Arena rating, higher is better
- Range
- 1,414 to 1,459
- Sample
- 759 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,499
Expert prompts · rank 8 of 359
- Unit
- Arena rating, higher is better
- Range
- 1,474 to 1,523
- Sample
- 553 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,449
Instruction following · rank 33 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,435 to 1,463
- Sample
- 1884 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,460
Overall · rank 40 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,452 to 1,468
- Sample
- 5414 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,446
Writing, literature and language · rank 19 of 408
- Unit
- Arena rating, higher is better
- Range
- 1,429 to 1,464
- Sample
- 1186 votes
- Configuration
- Qwen3.6 Max Preview
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
28.4
Artificial Analysis Intelligence Index · rank 111 of 314
- Unit
- index score, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not business work, and a blend: read the parts for any one task.
88.8%
GPQA Diamond · rank 81 of 276
- Unit
- % of questions, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not applied work; multiple-choice science questions.
30.8%
Humanity's Last Exam · rank 117 of 314
- Unit
- % of questions, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not everyday work; academic questions at the edge of expertise.
76.6%
IFBench · rank 14 of 215
- Unit
- % of instructions, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not judgement about what an instruction meant.
80.7%
Long-context reasoning (AA-LCR) · rank 53 of 309
- Unit
- % of questions, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not retrieval over your own document store.
43.9%
Terminal-Bench Hard · rank 36 of 213
- Unit
- % of tasks, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not other harnesses; Artificial Analysis no longer runs it on new models.
95.9%
Τ²-bench telecom · rank 14 of 214
- Unit
- % of tasks, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not your policies or systems; no longer run on new models.
52.0%
Correct answers · rank 21 of 85
- Unit
- % of questions, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 27 Aug 2026
- Not shown
- Not answers grounded in your documents; tests what the model remembers.
$4,254.19
Money after a year · rank 35 of 63
- Unit
- US dollars, higher is better
- Configuration
- Qwen3.6 Max Preview
- Measured
- 1 Oct 2026
- Not shown
- Not a real business; one simulated market with set rules.
Compare Qwen3.6 Max Preview with