Roar AI
DocsRequest access
All models
OpenAI · Chat

GPT-5.4 nano

OpenAI’s small model for classifying, extracting and ranking at scale.

Input$0.23per 1M tokens
Output$1.4375per 1M tokens
BillingOne invoice, in LKRwith every other model
Request early accessSee every model’s rate

About GPT-5.4 nano

GPT-5.4 nano is built for the plain, repetitive jobs: sorting text into categories, pulling fields out of documents, ranking results and acting as a small helper inside a bigger agent. It answers without reasoning by default, which keeps it fast, and can be asked to think when a task needs it.

It is not a general assistant. On OSWorld, the computer-use test, it scored 39.0% against 72.1% for its bigger sibling GPT-5.4 mini, and the newer Luna models now cover the same ground with a much longer memory.

Where it’s strong

Classifying and tagging large volumes of text.

Pulling structured fields out of documents.

Ranking and reranking search results.

Simple helper agents.

Where to be careful

Weak at operating software and at multi-step agent work.

Its knowledge stops at August 2025, and its 400K context is well under the newer models’ million.

Price in rupees

Input$0.23 per 1M tokens
Output$1.4375 per 1M tokens
Cached input$0.023 per 1M tokens
Batch input$0.115 per 1M tokens
Batch output$0.7188 per 1M tokens

For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $2.13 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.

Specs

Made byOpenAI
ReleasedMarch 2026
TypeChat
Context window400K tokens
Longest answer128K tokens
ReadsText, Images
Thinks before answeringYes, adjustable
WeightsClosed
Model idgpt-5.4-nano

Using it from Sri Lanka

GPT-5.4 nano is billed in rupees, on the same monthly invoice as every other model your team uses — no foreign card, no separate OpenAI account, and one key for all of it.

Prompts sent to GPT-5.4 nano are processed outside Sri Lanka. If a workload has to stay in the country, Qwen3.8-27B (Colombo) runs on our own servers in Colombo.

Call it

Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.

from openai import OpenAI

client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")

reply = client.chat.completions.create(
    model="gpt-5.4-nano",
    messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)

Questions

How much does GPT-5.4 nano cost in Sri Lanka?

Input is $0.23 per 1M tokens; output is $1.4375 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $2.13. The rate is published and does not change from one request to the next.

Can I pay for GPT-5.4 nano in Sri Lankan rupees?

Yes. Usage is billed to one monthly invoice in rupees, together with every other model your team uses, and you can also pay by card. There is no separate OpenAI account to open and no foreign card needed.

Does GPT-5.4 nano keep data in Sri Lanka?

No — prompts sent to GPT-5.4 nano are processed outside Sri Lanka. If your data has to stay in the country, use Qwen3.8-27B (Colombo), which runs on our servers in Colombo.

Is GPT-5.4 nano open source?

No. OpenAI does not publish the weights, so it is only available through an API like this one.