# Qwen3.8-Max-0902

Qwen3.8-Max-0902 is an AI model released by Qwen on Sep 2 2026. It has 2.4T parameters and a 1M context window. At release it scored 64.9% on NL2Repo-Bench, 70% on QwenSWEBench V2 and 29% on Terminal-Bench 3.0.

## Facts

| Field | Value |
| --- | --- |
| Model | Qwen3.8-Max-0902 |
| Developer | Qwen |
| Release date | Wednesday, Sep 2 2026 |
| Licensing | Proprietary |
| Parameters | 2.4T |
| Context window | 1M |

## Benchmark scores published at release

| Benchmark | Score | Source | What it measures |
| --- | --- | --- | --- |
| DeepSWE 1.1 | 69.3% | Lab | Artificial Analysis' independent test of deep, agentic software-engineering work — the AI has to plan and carry out substantial coding tasks end to end. (Version 1.1 of the test.) Higher is better. |
| NL2Repo-Bench | 64.9% | Lab | Tests whether the AI can turn a natural-language requirement into working code across an entire repository, not just produce a single function or patch. Higher is better. |
| QwenSWEBench V2 | 70% | Lab | Qwen's in-house coding benchmark, second version, built around complex real-world software-engineering tasks. Scores are not comparable with the first version. Higher is better. |
| Terminal-Bench 3.0 | 29% | Lab | command-line task completion (v3.0, much harder task set) |
| JobBench | 64% | Lab | Tests the AI on professional workplace tasks that require using real work tools — the kind of multi-step jobs an office worker handles. Higher is better. |
| CoWorkBench | 76.1% | Lab | Tests long-running office tasks across fields including computer science, finance, law, medicine, and other productivity work. Higher is better. |
| Toolathlon-Verified | 73.3% | Lab | Tests how well the AI uses everyday personal tools and apps to get things done — a human-checked version of Toolathlon. Higher is better. |
| AutomationBench | 50.8% | Lab | Tests whether the AI can run real multi-step business workflows — the kind of end-to-end office processes companies want to automate — from start to finish. Higher is better. |
| MMMU-Pro | 82.7% | Lab | A tougher version of MMMU — college-level questions that mix images, diagrams, and text together. Higher is better. |

## About Qwen3.8-Max-0902

Qwen3.8-Max-0902, released September 2, 2026, was a post-trained refresh of the Qwen3.8-Max Alibaba had shipped a month earlier rather than a new model underneath it. The 2.4-trillion-parameter base and the 1M-token context window were unchanged; what the lab changed was the post-training, which it said had been extended on coding and on Cowork, the agentic office work its own CoWorkBench measures. The dated suffix marked the checkpoint the way DeepSeek had been naming its own refreshes since V3-0324, and it kept the older checkpoint addressable instead of replacing it in place.

The launch table was set against the August checkpoint rather than against rival labs, and the coding rows carried most of the movement. Qwen3.8-Max-0902 scored 29.0% on Terminal-Bench 3.0 where the older checkpoint had managed 11.3%, 69.3% on DeepSWE 1.1 against 56.6%, and 64.9% on NL2Repo-Bench against 55.9%; on QwenSWEBench V2, Alibaba's own harder in-house set, it posted 70.0% to the earlier model's 55.1%, ahead of the 68.0% and 67.1% it reported for Claude Opus 5 and Fable 5. Agent work moved less. CoWorkBench went from 74.8% to 76.1% and Toolathlon-Verified from 72.5% to 73.3%, both still behind Claude Opus 5, while JobBench jumped from 53.4% to 64.0%. Multimodal scores were flat, at 82.7% on MMMU-Pro against 82.3%.

Pricing was $2 per million input tokens and $6 per million output, with cache hits at $0.17 explicit and $0.25 implicit — the first time the Max tier had been quoted on QwenCloud, the API Alibaba had introduced with Qwen3.8-Flash the previous week, rather than through Alibaba Cloud Model Studio. It launched API-only, with no weights published for the refreshed checkpoint, leaving the August base model the most recent Qwen3.8 weights available.

## Questions and answers

### When was Qwen3.8-Max-0902 released?

Qwen3.8-Max-0902 was released by Qwen on Wednesday, Sep 2 2026.

### Who made Qwen3.8-Max-0902?

Qwen3.8-Max-0902 was built by Qwen. Alibaba's AI lab, building the Qwen family. The most prolific publisher of open-weight models of any major lab, alongside a proprietary Max and Plus tier sold through Alibaba Cloud.

### What benchmark scores did Qwen3.8-Max-0902 get?

Qwen3.8-Max-0902 reports 9 tracked benchmark scores — DeepSWE 1.1: 69.3%; NL2Repo-Bench: 64.9%; QwenSWEBench V2: 70%; Terminal-Bench 3.0: 29%; JobBench: 64%; CoWorkBench: 76.1%; Toolathlon-Verified: 73.3%; AutomationBench: 50.8%; MMMU-Pro: 82.7%. Scores are the figures published at release by Qwen. It holds the best score among all models tracked here on NL2Repo-Bench, QwenSWEBench V2, Terminal-Bench 3.0, JobBench, CoWorkBench and AutomationBench.

### What is the context window of Qwen3.8-Max-0902?

Qwen3.8-Max-0902 has a context window of 1M. That is the maximum amount of input plus output the model can hold in a single request.

### How many parameters does Qwen3.8-Max-0902 have?

Qwen3.8-Max-0902 is reported at 2.4T parameters.

### Is Qwen3.8-Max-0902 open source?

No. Qwen3.8-Max-0902 is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after Qwen3.8-Max-0902?

Qwen's previous tracked release was Qwen3.8-Flash-Next on Aug 26 2026, 7 days earlier. It is the most recent Qwen model tracked on AI Release Tracker.


---

Canonical page: https://aireleasetracker.com/model/qwen/qwen3.8-max-0902
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
