Roar AI
DocsRequest access
All models
Anthropic · Chat

Claude Haiku 4.5

Anthropic’s smallest, fastest model, for chat, support and work split across many small agents.

Input$1.15per 1M tokens
Output$5.75per 1M tokens
BillingOne invoice, in LKRwith every other model
Request early accessSee every model’s rate

About Claude Haiku 4.5

Haiku 4.5 is the quick one. When it came out in October 2025 it matched the coding of the then-current Sonnet 4 at a third of the cost and more than twice the speed, scoring 73.3% on SWE-bench Verified. It was also the first Haiku that could think before answering and operate a computer.

Anthropic pitches it as a worker a bigger model directs — many Haikus running in parallel under one Opus — and for anything a customer waits on, like live chat. It is the oldest model in Anthropic’s current line, and that shows in what it knows.

Where it’s strong

Live chat and customer-support agents, where speed matters most.

Sub-agents running in parallel under a larger model.

High-volume sorting, tagging and data extraction.

Quick coding help and prototypes.

Where to be careful

Its knowledge is old: reliable to February 2025, more than a year behind the 5.x models.

Anthropic only guarantees it until mid-October 2026. For new work, plan for a successor.

Smaller limits than newer models: 200K tokens of context and 64K of output.

Price in rupees

Input$1.15 per 1M tokens
Output$5.75 per 1M tokens
Cached input$0.115 per 1M tokens
Cache write$1.4375 per 1M tokens
Batch input$0.575 per 1M tokens
Batch output$2.875 per 1M tokens

For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $9.20 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.

Specs

Made byAnthropic
ReleasedOctober 2025
TypeChat
Context window200K tokens
Longest answer64K tokens
ReadsText, Images, PDFs
Thinks before answeringYes, adjustable
WeightsClosed
Model idclaude-haiku-4-5

Using it from Sri Lanka

Claude Haiku 4.5 is billed in rupees, on the same monthly invoice as every other model your team uses — no foreign card, no separate Anthropic account, and one key for all of it.

Prompts sent to Claude Haiku 4.5 are processed outside Sri Lanka. If a workload has to stay in the country, Qwen3.8-27B (Colombo) runs on our own servers in Colombo.

Anthropic’s published language tests for Haiku cover Hindi (92% of its English score) and Bengali (90%), but not Sinhala or Tamil. Test it on your own messages first.

Call it

Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.

from openai import OpenAI

client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")

reply = client.chat.completions.create(
    model="claude-haiku-4-5",
    messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)

Questions

How much does Claude Haiku 4.5 cost in Sri Lanka?

Input is $1.15 per 1M tokens; output is $5.75 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $9.20. The rate is published and does not change from one request to the next.

Can I pay for Claude Haiku 4.5 in Sri Lankan rupees?

Yes. Usage is billed to one monthly invoice in rupees, together with every other model your team uses, and you can also pay by card. There is no separate Anthropic account to open and no foreign card needed.

Does Claude Haiku 4.5 keep data in Sri Lanka?

No — prompts sent to Claude Haiku 4.5 are processed outside Sri Lanka. If your data has to stay in the country, use Qwen3.8-27B (Colombo), which runs on our servers in Colombo.

Is Claude Haiku 4.5 open source?

No. Anthropic does not publish the weights, so it is only available through an API like this one.