# Composer 2.5

Composer 2.5 is an AI model released by Cursor (now part of SpaceXAI) on May 18 2026. At release it scored 92% on Next.js Evals, 73% on Terminal-Bench 2.1 and 54% on SWE-Bench Pro.

## Facts

| Field | Value |
| --- | --- |
| Model | Composer 2.5 |
| Developer | Cursor (now part of SpaceXAI) |
| Release date | Monday, May 18 2026 |
| Licensing | Proprietary |

## Benchmark scores published at release

| Benchmark | Score | What it measures |
| --- | --- | --- |
| SWE-Bench Pro | 54% | Can the AI fix real bugs in real software? It's handed actual problems from open-source projects and has to write code that genuinely solves them. Higher is better. |
| SWE-Bench Multilingual | 79.8% | Like SWE-Bench, but the coding problems span many programming languages, not just one. Tests how broadly the AI can code. Higher is better. |
| CursorBench v3.2 | 56.1% | Cursor's own test of harder, real-world coding tasks inside a code editor, on the refreshed v3.2 task set. Scores aren't comparable with v3.1. Higher is better. |
| CursorBench v3.1 | 63.2% | Cursor's own test of harder, real-world coding tasks inside a code editor. Higher is better. |
| DeepSWE 1.0 | 18% | Artificial Analysis' independent test of deep, agentic software-engineering work — the AI has to plan and carry out substantial coding tasks end to end. Higher is better. |
| Next.js Evals | 92% | Vercel's open eval of how well AI coding agents build and migrate real Next.js apps — measured as the share of tasks the agent completes successfully. Higher is better. |
| Terminal-Bench 2.1 | 73% | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? Higher is better. |
| Terminal-Bench 2.0 | 69.3% | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? (Version 2.0 of the test.) Higher is better. |

## Questions and answers

### When was Composer 2.5 released?

Composer 2.5 was released by Cursor (now part of SpaceXAI) on Monday, May 18 2026.

### Who made Composer 2.5?

Composer 2.5 was built by Cursor (now part of SpaceXAI). Elon Musk's AI company building the Grok series of models. Founded in 2023 as xAI, now part of SpaceX. Cursor (Anysphere) and its Composer coding models were acquired in 2026 and are tracked here.

### What benchmark scores did Composer 2.5 get?

Composer 2.5 reports 8 tracked benchmark scores — SWE-Bench Pro: 54%; SWE-Bench Multilingual: 79.8%; CursorBench v3.2: 56.1%; CursorBench v3.1: 63.2%; DeepSWE 1.0: 18%; Next.js Evals: 92%; Terminal-Bench 2.1: 73%; Terminal-Bench 2.0: 69.3%. Scores are the figures published at release by Cursor (now part of SpaceXAI). It holds the best score among all models tracked here on Next.js Evals.

### Is Composer 2.5 open source?

No. Composer 2.5 is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after Composer 2.5?

SpaceXAI's previous tracked release was Grok 4.3 Beta on Apr 17 2026, 31 days earlier. It was followed by Grok 4.5 on Jul 8 2026.


---

Canonical page: https://aireleasetracker.com/model/xai/composer-2.5
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
