# Qwen3.5

Qwen3.5 is an AI model released by Qwen on Feb 16 2026. It has 397B parameters, a 1M context window and open weight. At release it scored 85% on MMMU, 88.4% on GPQA Diamond and 76.4% on SWE-Bench Verified.

## Facts

| Field | Value |
| --- | --- |
| Model | Qwen3.5 |
| Developer | Qwen |
| Release date | Monday, Feb 16 2026 |
| Licensing | Open Weight |
| Parameters | 397B |
| Context window | 1M |

## Benchmark scores published at release

| Benchmark | Score | What it measures |
| --- | --- | --- |
| SWE-Bench Verified | 76.4% | Real coding tasks pulled from open-source projects — the AI has to find and fix actual bugs. A human-checked version of the original SWE-Bench. Higher is better. |
| SWE-Bench Multilingual | 69.3% | Like SWE-Bench, but the coding problems span many programming languages, not just one. Tests how broadly the AI can code. Higher is better. |
| Terminal-Bench 2.0 | 52.5% | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? (Version 2.0 of the test.) Higher is better. |
| BrowseComp | 69% | Can the AI browse the web and track down hard-to-find answers? Higher is better. |
| Humanity's Last Exam (no tools) | 28.7% | Humanity's Last Exam — extremely hard expert questions across many subjects, written so you can't just look up the answer. “No tools” means the AI answers on its own. Higher is better. |
| GPQA Diamond | 88.4% | Graduate-level science questions in biology, physics, and chemistry — hard enough that subject-matter PhDs score around 65%. Higher is better. |
| OSWorld-Verified | 62.2% | Can the AI actually operate a computer — clicking, typing, and using real apps — to finish tasks on its own? Higher is better. |
| CharXiv Reasoning | 80.8% | Can the AI read and reason about complex charts and figures, not just text? Higher is better. |
| MMMU-Pro | 79% | A tougher version of MMMU — college-level questions that mix images, diagrams, and text together. Higher is better. |
| MMMU | 85% | Tests the AI on understanding images and text together across many college subjects. Higher is better. |

## About Qwen3.5

Qwen3.5, released February 16, 2026, was a 397B-parameter mixture-of-experts model with 17B active, published under Apache 2.0 with a 1M-token context window. Alibaba framed it around agent work rather than chat: the launch claim was that it could operate desktop and mobile applications directly, which put it in the same competitive frame as the computer-use models the US labs had been shipping.

On Alibaba's own harness it posted 76.4% on SWE-Bench Verified, 88.4% on GPQA Diamond and 62.2% on OSWorld-Verified. The last of those fit the agent pitch — GPT-5.2 managed 38.2% on the same table — even as Qwen3.5 trailed GPT-5.2 and Claude Opus 4.5 by roughly four points on SWE-Bench Verified and sat at 52.5% on Terminal-Bench 2.

It launched the same day as the proprietary Qwen3.5-Plus, making explicit the two-tier structure Qwen3-Max had introduced five months earlier. Smaller open members of the family followed quickly — 122B-A10B, 35B-A3B and 27B on February 24, then 9B, 4B, 2B and 0.8B variants on March 2, 2026 — repeating the size-ladder pattern that had made Qwen2.5 and Qwen3 so widely adopted.

## Questions and answers

### When was Qwen3.5 released?

Qwen3.5 was released by Qwen on Monday, Feb 16 2026.

### Who made Qwen3.5?

Qwen3.5 was built by Qwen. Alibaba's AI lab, building the Qwen family. The most prolific publisher of open-weight models of any major lab, alongside a proprietary Max and Plus tier sold through Alibaba Cloud.

### What benchmark scores did Qwen3.5 get?

Qwen3.5 reports 10 tracked benchmark scores — SWE-Bench Verified: 76.4%; SWE-Bench Multilingual: 69.3%; Terminal-Bench 2.0: 52.5%; BrowseComp: 69%; Humanity's Last Exam (no tools): 28.7%; GPQA Diamond: 88.4%; OSWorld-Verified: 62.2%; CharXiv Reasoning: 80.8%; MMMU-Pro: 79%; MMMU: 85%. Scores are the figures published at release by Qwen. It holds the best score among all models tracked here on MMMU.

### What is the context window of Qwen3.5?

Qwen3.5 has a context window of 1M. That is the maximum amount of input plus output the model can hold in a single request.

### How many parameters does Qwen3.5 have?

Qwen3.5 is reported at 397B parameters.

### Is Qwen3.5 open source?

Partly. Qwen3.5 is an open-weight model: the trained weights are free to download and run locally or on your own infrastructure, but the training data and code are not fully released and the license may restrict some commercial uses. It is not open source in the strict sense.

### What came before and after Qwen3.5?

Qwen's previous tracked release was Qwen3-Coder-Next on Feb 3 2026, 13 days earlier. It was followed by Qwen3.5-Plus on Feb 16 2026.


---

Canonical page: https://aireleasetracker.com/model/qwen/qwen3.5
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
