Gemini 3.8 Flash
Google’s newest Flash: near-flagship coding at Flash speed.
About Gemini 3.8 Flash
Gemini 3.8 Flash is the third Flash release in six weeks and Google’s most capable so far. On the DeepSWE coding test it scored 73.7%, just under Claude Opus 5 and ahead of GPT-5.6 Sol, and Google reports wins over Opus 5 on finance and legal agent tests. It reads the same mix as Gemini Pro — text, images, audio, video and PDFs — in a million-token window.
Once it starts, it is fast: about 300 tokens a second in independent tests. Starting is the slow part — about 13 seconds to the first token, against a median of 3 — and it is wordy, so tasks can cost more than its rates suggest.
Where it’s strong
Coding agents at a mid-range price.
Working through large volumes of video, audio and PDFs.
Finance and legal analysis agents.
Batch jobs where total throughput matters more than the first word.
Where to be careful
Slow to begin answering, which makes it a poor fit for live chat.
Verbose: one review measured tasks costing about 40% more than on Gemini 3.7 Flash at the same rates.
Google’s model card notes its safety behaviour in languages other than English slipped compared with 3.7 Flash.
Price
For scale: 1,000 requests of about 3,000 tokens in and 1,000 out cost $6.90 — a long prompt and a page of answer each. Every rate is published and stays the same from one request to the next. See every model.
Specs
gemini-3.8-flashCall it
Use the OpenAI SDK you already have — point it at Roar AI and name the model. Read the docs.
from openai import OpenAI
client = OpenAI(base_url="https://api.roar-ai.com/v1", api_key="roar_live_…")
reply = client.chat.completions.create(
model="gemini-3.8-flash",
messages=[{"role": "user", "content": "Summarise this contract in plain English."}],
)Compare Gemini 3.8 Flash
Questions
How much does Gemini 3.8 Flash cost?
Input is $0.8625 per 1M tokens; output is $4.3125 per 1M tokens. For scale, 1,000 requests of about 3,000 tokens in and 1,000 out cost $6.90. The rate is published and does not change from one request to the next.
Can I use Gemini 3.8 Flash alongside other models?
Yes. Every model on Roar AI is on the same API key and the same monthly invoice, so switching from Gemini 3.8 Flash to another model is a change to one word in your code.
Is Gemini 3.8 Flash open source?
No. Google does not publish the weights, so it is only available through an API like this one.
