Roar AI
DocsRequest access
All models
Alibaba · Open weights

Qwen3.8-Max

Alibaba’s flagship, and the first Max-tier Qwen with downloadable weights.

Input$2.21per 1M tokens
Output$6.63per 1M tokens
BillingOne monthly invoiceevery model on one key
Request early accessSee every model’s rate

About Qwen3.8-Max

Qwen3.8-Max is Alibaba’s top model and the first Max-class Qwen whose weights you can download — a 2.4-trillion-parameter mixture of experts. Alibaba trained it to check its own visual output, and reports 67.7% on SWE-bench Pro, 92.6% on GPQA Diamond and 86.1% on the OSWorld computer-use test.

Through the API it reads images and video and holds a million tokens; the downloadable version is text-only with a shorter context. It thinks before answering, at three effort levels.

Where it’s strong

Frontier-level coding agents.

Operating desktop software.

Research and analysis across long documents.

Where to be careful

Every headline score is from Alibaba’s internal runs.

The licence on the open weights restricts very large AI-service companies.

Price

Input$2.21 per 1M tokens
Output$6.63 per 1M tokens
Cached input$0.2925 per 1M tokens
Cache write$3.25 per 1M tokens

For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $13.26 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.

Specs

Made byAlibaba
ReleasedAugust 2026
TypeChat
Context window1M tokens
Longest answer131K tokens
ReadsText, Images, Video
Thinks before answeringYes, adjustable
WeightsOpen (Qwen3.8-Max License)
Size2.4T total, 95B active
Model idqwen3.8-max

Call it

Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.

from openai import OpenAI

client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")

reply = client.chat.completions.create(
    model="qwen3.8-max",
    messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)

Questions

How much does Qwen3.8-Max cost?

Input is $2.21 per 1M tokens; output is $6.63 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $13.26. The rate is published and does not change from one request to the next.

Can I use Qwen3.8-Max alongside other models?

Yes. Every model on Roar AI is on the same API key and the same monthly invoice, so switching from Qwen3.8-Max to another model is a change to one word in your code.

Is Qwen3.8-Max open source?

Its weights are open under the Qwen3.8-Max License licence, so you could run it yourself. Through Roar AI you call it like any other model, with no servers to manage.