Qwen3.8-27B
Alibaba’s small dense Qwen: open, multimodal and unusually strong at coding for 27 billion parameters.
About Qwen3.8-27B
Qwen3.8-27B is the compact member of Alibaba’s Qwen3.8 generation — 27 billion parameters, dense, open under Apache 2.0 — and it punches well above its size. Alibaba reports 61.7% on SWE-bench Pro, above its own figure for Claude Opus 4.6, plus 89.2% on GPQA Diamond and 84.3% on the OSWorld computer-use test.
It reads images and video as well as text, holds 262K tokens natively, and thinks by default with adjustable effort. Its design mixes fast linear-attention layers with full attention, which keeps long inputs cheap to process.
Where it’s strong
Coding agents at a low price.
Operating desktop software.
Reading documents, images and video.
Where to be careful
Every headline score is Alibaba’s own.
Stretching it to a million tokens can hurt quality on shorter texts.
Price in rupees
For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $4.09 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.
Specs
qwen3.8-27bUsing it from Sri Lanka
Qwen3.8-27B is billed in rupees, on the same monthly invoice as every other model your team uses — no foreign card, no separate Alibaba account, and one key for all of it.
Prompts sent to Qwen3.8-27B are processed outside Sri Lanka. If a workload has to stay in the country, Qwen3.8-27B (Colombo) runs on our own servers in Colombo.
Alibaba says the Qwen family covers 201 languages and dialects, without naming Sinhala or Tamil. The same model also runs on our own server in Colombo, for work whose data has to stay in the country.
Call it
Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.
from openai import OpenAI
client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")
reply = client.chat.completions.create(
model="qwen3.8-27b",
messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)Compare Qwen3.8-27B
Questions
How much does Qwen3.8-27B cost in Sri Lanka?
Input is $0.455 per 1M tokens; output is $2.73 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $4.09. The rate is published and does not change from one request to the next.
Can I pay for Qwen3.8-27B in Sri Lankan rupees?
Yes. Usage is billed to one monthly invoice in rupees, together with every other model your team uses, and you can also pay by card. There is no separate Alibaba account to open and no foreign card needed.
Does Qwen3.8-27B keep data in Sri Lanka?
No — prompts sent to Qwen3.8-27B are processed outside Sri Lanka. If your data has to stay in the country, use Qwen3.8-27B (Colombo), which runs on our servers in Colombo.
Is Qwen3.8-27B open source?
Its weights are open under the Apache 2.0 licence, so you could run it yourself. Through Roar AI you call it like any other model, with no servers to manage.
