Best for long context
Large context window for documents and transcripts.
Model · Google
Gemini 3.8 Flash is served by Google through OpenRouter. It provides 1.05M tokens of context and supports text, vision, audio, video, files inputs.
| Provider | |
| OpenRouter ID | google/gemini-3.8-flash |
| Modalities | Text, Vision, Audio, Video, Files |
| Best for | Multimodal, Long Context |
openrouter.ai/api/v1/models Verified 2026-10-10 Drift-checked daily 06:00 CST Best for long context
Large context window for documents and transcripts.
Best multimodal
Handles images, files, audio or video inputs.
These examples use the 2026-10-10 input and output prices. Cache discounts, provider surcharges and tool fees are excluded.
| Workload | Input | Output | Estimated cost |
|---|---|---|---|
| Quick test | 20,000 | 5,000 | $0.034 |
| Product chat month | 500,000 | 125,000 | $0.844 |
| Document batch | 5,000,000 | 1,000,000 | $7.50 |
Run these examples from your server or edge function. Keep the OpenRouter API key private and do not expose it in browser code.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "google/gemini-3.8-flash",
"messages": [{ "role": "user", "content": "Summarize this document in five bullet points." }],
"max_tokens": 500
}' const response = await fetch("https://openrouter.ai/api/v1/chat/completions", {
method: "POST",
headers: {
Authorization: `Bearer ${process.env.OPENROUTER_API_KEY}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
model: "google/gemini-3.8-flash",
messages: [{ role: "user", content: "Summarize this document in five bullet points." }],
max_tokens: 500,
}),
});
const data = await response.json();
console.log(data.choices[0].message.content); Compare models View on OpenRouter ↗
Gemini 3.8 Flash is a Google model available through OpenRouter. It supports text, vision, audio, video, files inputs and is suited for multimodal, long context workloads.
The 2026-10-10 snapshot lists $0.750 per 1M input tokens and $3.75 per 1M output tokens. Verify current pricing on OpenRouter before production.
Gemini 3.8 Flash offers 1.05M tokens of context in the current snapshot. Provider-specific limits and surcharges may apply.
If your input is words and pictures, Haiku 5.5 is the cheaper answer by a wide margin, and it is positioned for exactly the high-volume jobs (summarisation, subagents, browser use) that cheap tiers exist for. Reach for Gemini Flash when the input is a video, a recording, or a folder of mixed files — that is the capability worth 7.5x, not the model's badge.
Pricing and context data are based on a 2026-10-10 snapshot. Provider limits can change; always verify on OpenRouter before production.