Roar AI
DocsRequest access
All models
OpenAI · Open weights

gpt-oss-120b

OpenAI’s larger open-weight model: o4-mini-class reasoning you could also run yourself.

Input$0.0481per 1M tokens
Output$0.221per 1M tokens
BillingOne invoice, in LKRwith every other model
Request early accessSee every model’s rate

About gpt-oss-120b

gpt-oss-120b was OpenAI’s first open-weight language model since GPT-2, released in August 2025 under Apache 2.0. OpenAI says it comes close to o4-mini on core reasoning tests and beats it on competition maths and health questions, and it scores 80.1% on GPQA Diamond. It was built to fit on a single 80GB GPU.

It reasons at low, medium or high effort and shows its full chain of thought. Its small active size makes it inexpensive to run, which makes it a sensible choice for maths, logic and structured tasks at volume.

Where it’s strong

Maths, science and logic problems.

Structured, step-by-step tasks at volume.

Agent steps that call functions.

Where to be careful

Its knowledge stops at June 2024, the oldest of the reasoning models here.

Text only.

Price in rupees

Input$0.0481 per 1M tokens
Output$0.221 per 1M tokens
Cached input$0.0975 per 1M tokens

For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $0.37 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.

Specs

Made byOpenAI
ReleasedAugust 2025
TypeChat
Context window131K tokens
Longest answer131K tokens
ReadsText
Thinks before answeringAlways on
WeightsOpen (Apache 2.0)
Size117B total, 5.1B active
Model idgpt-oss-120b

Using it from Sri Lanka

gpt-oss-120b is billed in rupees, on the same monthly invoice as every other model your team uses — no foreign card, no separate OpenAI account, and one key for all of it.

Prompts sent to gpt-oss-120b are processed outside Sri Lanka. If a workload has to stay in the country, Qwen3.8-27B (Colombo) runs on our own servers in Colombo.

OpenAI tested it in 14 languages, including Hindi and Bengali, averaging 81.3% on a multilingual knowledge test. Sinhala and Tamil were not among them.

Call it

Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.

from openai import OpenAI

client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")

reply = client.chat.completions.create(
    model="gpt-oss-120b",
    messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)

Questions

How much does gpt-oss-120b cost in Sri Lanka?

Input is $0.0481 per 1M tokens; output is $0.221 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $0.37. The rate is published and does not change from one request to the next.

Can I pay for gpt-oss-120b in Sri Lankan rupees?

Yes. Usage is billed to one monthly invoice in rupees, together with every other model your team uses, and you can also pay by card. There is no separate OpenAI account to open and no foreign card needed.

Does gpt-oss-120b keep data in Sri Lanka?

No — prompts sent to gpt-oss-120b are processed outside Sri Lanka. If your data has to stay in the country, use Qwen3.8-27B (Colombo), which runs on our servers in Colombo.

Is gpt-oss-120b open source?

Its weights are open under the Apache 2.0 licence, so you could run it yourself. Through Roar AI you call it like any other model, with no servers to manage.