qwen3.8-max vs qwen3.8-max-0902
Parameters, independent leaderboard coverage, real-task reports and open-weight availability are shown side by side; missing coverage is not a zero.
Parameters
| Comparison item | qwen3.8-max | qwen3.8-max-0902 |
|---|---|---|
| Model Type | LLM | LLM |
| Context window | 1,000,000 tokens | 1,000,000 tokens |
| Maximum output | 131,072 tokens | 131,072 tokens |
| Input/output modalities | text, image, video → text | text, image, video → text |
| Endpoint | openai · anthropic | openai · anthropic |
| Open weights | Unavailable | Unavailable |
| License | Proprietary hosted model | Proprietary hosted model |
Leaderboard coverage
Every independent board covering either model is retained; an em dash means that board has not published a figure for that model.
Real-task reports
qwen3.8-max
- AI Coding Daily LLM Coding Leaderboard — Implement and harden production-shaped projects in several stacks, including a Laravel app, a Go service and a React/TypeScript app, from the same prompts. Finding: Qwen 3.8 Max at high reasoning scored 47.08 ($0.45 and 10 min 21 s per run in OpenCode), 29th of 47 rows, below its 0902 snapshot (48.70). Limitation: One author's benchmark whose code-quality half is graded by a judge model; rows run in different harnesses, the page is updated as models are added (read on 2026-09-24, 47 rows), and the author sells premium tutorials and advertising.
qwen3.8-max-0902
- AI Coding Daily LLM Coding Leaderboard — Implement and harden production-shaped projects in several stacks, including a Laravel app, a Go service and a React/TypeScript app, from the same prompts. Finding: Qwen 3.8 Max (0902) at high reasoning scored 48.70 ($0.40 and 11 min 49 s per run in OpenCode), 22nd of 47 rows and above the original qwen3.8-max (47.08). Limitation: One author's benchmark whose code-quality half is graded by a judge model; rows run in different harnesses, the page is updated as models are added (read on 2026-09-24, 47 rows), and the author sells premium tutorials and advertising.
- kzhu's six-question LLM test leaderboard — Six questions scored out of 10: constrained sentence writing, two reasoning puzzles (replaced by pelican and compound-bow SVG animations in version 2.2), a browser OS with a space shooter, a 3D racing game and a Trello-style board. Finding: Qwen3.8 Max 0902 scored 47/60 on version 2.1 in the DeepSeek harness and 46 in Pi after re-running only the two new version-2.2 SVG questions, where the pelican flipped upside down and the bow pointed the wrong way (6 and 5). Limitation: One author's scores on six questions; rows use different harnesses and test versions (2.1 or 2.2), and a few rows' table total differs by one point from the author's note.