# Qwen3.8-Max

Qwen3.8-Max is an AI model released by Qwen on Aug 3 2026. It has 2.4T parameters and a 1M context window. At release it scored 93% on PaperBench, 86.1% on OSWorld-Verified and 82% on BabyVision.

## Facts

| Field | Value |
| --- | --- |
| Model | Qwen3.8-Max |
| Developer | Qwen |
| Release date | Monday, Aug 3 2026 |
| Licensing | Proprietary |
| Parameters | 2.4T |
| Context window | 1M |

## Benchmark scores published at release

| Benchmark | Score | What it measures |
| --- | --- | --- |
| SWE-Bench Pro | 67.7% | Can the AI fix real bugs in real software? It's handed actual problems from open-source projects and has to write code that genuinely solves them. Higher is better. |
| PaperBench | 93% | reproducing the results of an ML research paper end to end |
| Terminal-Bench 2.1 | 86.6% | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? Higher is better. |
| JobBench | 53.4% | Tests the AI on professional workplace tasks that require using real work tools — the kind of multi-step jobs an office worker handles. Higher is better. |
| OSWorld-Verified | 86.1% | Can the AI actually operate a computer — clicking, typing, and using real apps — to finish tasks on its own? Higher is better. |
| CharXiv Reasoning | 88.4% | Can the AI read and reason about complex charts and figures, not just text? Higher is better. |
| BabyVision | 82% | Tests core visual reasoning — seeing and understanding images the way even young children can, which AIs often find surprisingly hard. Higher is better. |

## About Qwen3.8-Max

Qwen3.8-Max, released August 3, 2026, was the largest model Alibaba had built: 2.4 trillion total parameters with 95B active, a 1M-token context window, and hybrid thinking. Alibaba previewed it on July 19 and shipped it to the Alibaba Cloud API two weeks later, placing it fifth on Text Arena and second on Vision Arena at launch.

The launch numbers pointed at agents rather than raw coding. It scored 93.0% on PaperBench and 86.1% on OSWorld-Verified — both ahead of every model Alibaba benchmarked against, including GPT-5.6 Sol and Claude Fable 5 — along with 86.6% on Terminal-Bench 2.1 and 53.4% on JobBench. On conventional software engineering it was further back, at 67.7% on SWE-Bench Pro against the 80.3% Claude Fable 5 had posted at its own release two months earlier. Multimodal results were strong without a code tool: 88.4% on CharXiv chart reasoning and 82.0% on BabyVision.

Alibaba announced at release that it intended to publish the weights — which would have made it by a wide margin the largest model any lab had opened. That commitment, on a proprietary Max-tier flagship, cut against the pattern the Max line had followed since Qwen3-Max. It landed two and a half weeks after Moonshot AI's Kimi K3, in a stretch where the Chinese labs were shipping trillion-parameter models within weeks of each other.

## Questions and answers

### When was Qwen3.8-Max released?

Qwen3.8-Max was released by Qwen on Monday, Aug 3 2026.

### Who made Qwen3.8-Max?

Qwen3.8-Max was built by Qwen. Alibaba's AI lab, building the Qwen family. The most prolific publisher of open-weight models of any major lab, alongside a proprietary Max and Plus tier sold through Alibaba Cloud.

### What benchmark scores did Qwen3.8-Max get?

Qwen3.8-Max reports 7 tracked benchmark scores — SWE-Bench Pro: 67.7%; PaperBench: 93%; Terminal-Bench 2.1: 86.6%; JobBench: 53.4%; OSWorld-Verified: 86.1%; CharXiv Reasoning: 88.4%; BabyVision: 82%. Scores are the figures published at release by Qwen. It holds the best score among all models tracked here on PaperBench, OSWorld-Verified and BabyVision.

### What is the context window of Qwen3.8-Max?

Qwen3.8-Max has a context window of 1M. That is the maximum amount of input plus output the model can hold in a single request.

### How many parameters does Qwen3.8-Max have?

Qwen3.8-Max is reported at 2.4T parameters.

### Is Qwen3.8-Max open source?

No. Qwen3.8-Max is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after Qwen3.8-Max?

Qwen's previous tracked release was Qwen3.7-Plus on Jun 1 2026, 63 days earlier. It is the most recent Qwen model tracked on AI Release Tracker.


---

Canonical page: https://aireleasetracker.com/model/qwen/qwen3.8-max
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
