HunterAlphaHub
OpenRouter model reference Facts from the public catalogue, dated and labelled
Jev Laya Verified
2026-10-10

Model · Google

Gemini 3.8 Flash

Gemini 3.8 Flash is served by Google through OpenRouter. It provides 1.05M tokens of context and supports text, vision, audio, video, files inputs.

Snapshot 2026-10-10 google/gemini-3.8-flash
Context window
1.05M tokens
Input / 1M tokens
$0.750
Output / 1M tokens
$3.75

Key facts

Provider Google
OpenRouter ID google/gemini-3.8-flash
Modalities Text, Vision, Audio, Video, Files
Best for Multimodal, Long Context
Source openrouter.ai/api/v1/models Verified 2026-10-10 Drift-checked daily 06:00 CST

Best-fit workloads

Best for long context

Large context window for documents and transcripts.

Best multimodal

Handles images, files, audio or video inputs.

Cost estimates

These examples use the 2026-10-10 input and output prices. Cache discounts, provider surcharges and tool fees are excluded.

Workload Input Output Estimated cost
Quick test 20,000 5,000 $0.034
Product chat month 500,000 125,000 $0.844
Document batch 5,000,000 1,000,000 $7.50

Strengths

  • ✓ The widest input set here: text, image, video, file and audio all go in at $0.75 / $3.75 per million
  • ✓ 1,048,576-token window at a Flash price
  • ✓ Google's Flash successor to 3.7 Flash, which this hub also carries

Limitations

  • ! Text output only, despite reading video and audio
  • ! 65,536-token output ceiling
  • ! Image generation is a different Google row (Nano Banana 2.1) — this one reads pictures, it does not draw them

OpenRouter API quickstart

Run these examples from your server or edge function. Keep the OpenRouter API key private and do not expose it in browser code.

cURL

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemini-3.8-flash",
    "messages": [{ "role": "user", "content": "Summarize this document in five bullet points." }],
    "max_tokens": 500
  }'

JavaScript / TypeScript

const response = await fetch("https://openrouter.ai/api/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.OPENROUTER_API_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    model: "google/gemini-3.8-flash",
    messages: [{ role: "user", content: "Summarize this document in five bullet points." }],
    max_tokens: 500,
  }),
});

const data = await response.json();
console.log(data.choices[0].message.content);

How to evaluate Gemini 3.8 Flash

  1. Run five real tasks from your product, not generic demo prompts.
  2. Record token usage, latency and output quality.
  3. Estimate cost using your real input/output mix.
  4. Compare it against one model with a similar price and one stronger model.

Compare models View on OpenRouter ↗

Gemini 3.8 Flash FAQ

What is Gemini 3.8 Flash?

Gemini 3.8 Flash is a Google model available through OpenRouter. It supports text, vision, audio, video, files inputs and is suited for multimodal, long context workloads.

How much does Gemini 3.8 Flash cost on OpenRouter?

The 2026-10-10 snapshot lists $0.750 per 1M input tokens and $3.75 per 1M output tokens. Verify current pricing on OpenRouter before production.

How large is Gemini 3.8 Flash's context window?

Gemini 3.8 Flash offers 1.05M tokens of context in the current snapshot. Provider-specific limits and surcharges may apply.

Head-to-head comparisons

Gemini 3.8 Flash vs Claude Haiku 5.5

If your input is words and pictures, Haiku 5.5 is the cheaper answer by a wide margin, and it is positioned for exactly the high-volume jobs (summarisation, subagents, browser use) that cheap tiers exist for. Reach for Gemini Flash when the input is a video, a recording, or a folder of mixed files — that is the capability worth 7.5x, not the model's badge.

Pricing and context data are based on a 2026-10-10 snapshot. Provider limits can change; always verify on OpenRouter before production.

Send us a link

A project built with Jev, a post, a video, a correction, a tip. We open the link, check it says what you said it says, and write the entry ourselves.

Required a link, and an email to reply to. Optional everything else.

Add context — all optional

One or two sentences about what it does, in your words. We write the entry ourselves.

Cost, latency, a benchmark — anything you measured. We attribute these to you.

We store what you type and email it to ourselves. No IP address, no user agent, no referrer — the same rule as the mailing list.

What happens to what you send →