Roar AI
DocsRequest access
All models
OpenAI · Chat

GPT-6 Luna

OpenAI’s cheapest GPT-6, with the same million-token memory as Astra.

Input$0.115per 1M tokens
Output$0.575per 1M tokens
BillingOne invoice, in LKRwith every other model
Request early accessSee every model’s rate

About GPT-6 Luna

Luna is the volume tier of GPT-6 — the model for jobs you run thousands of times a day. It keeps the flagship’s full 1.05M-token context and 128K-token answers, which is unusual at this price. At maximum effort OpenAI reports 66.6% on DeepSWE, which it puts level with Claude Opus 5 at medium effort.

Its reasoning can be switched off completely, so it also works as a fast plain model. Independent testing puts its intelligence roughly level with GPT-5.6 Luna: the gain is in price, not ability.

Where it’s strong

Sorting, tagging, extracting and routing at high volume.

Cheap helper agents inside a larger system.

Summarising very long documents in one pass.

Where to be careful

Not a step up in ability over GPT-5.6 Luna — the same class of model at a lower price.

Small-model mistakes still happen: one reviewer got Kubernetes files that passed every check and then crashed on start. Test generated code before shipping it.

Price in rupees

Input$0.115 per 1M tokens
Output$0.575 per 1M tokens
Cached input$0.0115 per 1M tokens
Cache write$0.1437 per 1M tokens
Batch input$0.0575 per 1M tokens
Batch output$0.2875 per 1M tokens

For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $0.92 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.

Specs

Made byOpenAI
ReleasedSeptember 2026
TypeChat
Context window1.05M tokens
Longest answer128K tokens
ReadsText, Images
Thinks before answeringYes, adjustable
WeightsClosed
Model idgpt-6-luna

Using it from Sri Lanka

GPT-6 Luna is billed in rupees, on the same monthly invoice as every other model your team uses — no foreign card, no separate OpenAI account, and one key for all of it.

Prompts sent to GPT-6 Luna are processed outside Sri Lanka. If a workload has to stay in the country, Qwen3.8-27B (Colombo) runs on our own servers in Colombo.

Call it

Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.

from openai import OpenAI

client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")

reply = client.chat.completions.create(
    model="gpt-6-luna",
    messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)

Questions

How much does GPT-6 Luna cost in Sri Lanka?

Input is $0.115 per 1M tokens; output is $0.575 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $0.92. The rate is published and does not change from one request to the next.

Can I pay for GPT-6 Luna in Sri Lankan rupees?

Yes. Usage is billed to one monthly invoice in rupees, together with every other model your team uses, and you can also pay by card. There is no separate OpenAI account to open and no foreign card needed.

Does GPT-6 Luna keep data in Sri Lanka?

No — prompts sent to GPT-6 Luna are processed outside Sri Lanka. If your data has to stay in the country, use Qwen3.8-27B (Colombo), which runs on our servers in Colombo.

Is GPT-6 Luna open source?

No. OpenAI does not publish the weights, so it is only available through an API like this one.