HunterAlphaHub
OpenRouter model reference Facts from the public catalogue, dated and labelled
Jev Laya Verified
2026-10-10

Model · inclusionAI

Ling 3.1 Flash

Ling 3.1 Flash is served by inclusionAI through OpenRouter. It provides 262K tokens of context and supports text inputs.

Snapshot 2026-10-10 inclusionai/ling-3.1-flash
Context window
262K tokens
Input / 1M tokens
$0
Output / 1M tokens
$0

Key facts

Provider inclusionAI
OpenRouter ID inclusionai/ling-3.1-flash
Modalities Text
Best for Budget
Source openrouter.ai/api/v1/models Verified 2026-10-10 Drift-checked daily 06:00 CST

Best-fit workloads

Best for budget

Low input/output cost for high-volume workloads.

Cost estimates

These examples use the 2026-10-10 input and output prices. Cache discounts, provider surcharges and tool fees are excluded.

Workload Input Output Estimated cost
Quick test 20,000 5,000 $0
Product chat month 500,000 125,000 $0
Document batch 5,000,000 1,000,000 $0

Strengths

  • ✓ Listed at $0 in and $0 out per million on the day we read it
  • ✓ 262,144-token window; 25B active parameters out of 560B in a hybrid reasoning mixture-of-experts
  • ✓ A free text row with a window big enough for real documents rather than a demo

Limitations

  • ! Text only — no images, files or audio
  • ! Free listings on this catalogue have been withdrawn before, so $0 here is a reading with a date on it, not a promise
  • ! 32,768-token output ceiling, the smallest here after Nano Banana

OpenRouter API quickstart

Run these examples from your server or edge function. Keep the OpenRouter API key private and do not expose it in browser code.

cURL

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "inclusionai/ling-3.1-flash",
    "messages": [{ "role": "user", "content": "Summarize this document in five bullet points." }],
    "max_tokens": 500
  }'

JavaScript / TypeScript

const response = await fetch("https://openrouter.ai/api/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.OPENROUTER_API_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    model: "inclusionai/ling-3.1-flash",
    messages: [{ role: "user", content: "Summarize this document in five bullet points." }],
    max_tokens: 500,
  }),
});

const data = await response.json();
console.log(data.choices[0].message.content);

How to evaluate Ling 3.1 Flash

  1. Run five real tasks from your product, not generic demo prompts.
  2. Record token usage, latency and output quality.
  3. Estimate cost using your real input/output mix.
  4. Compare it against one model with a similar price and one stronger model.

Compare models View on OpenRouter ↗

Ling 3.1 Flash FAQ

What is Ling 3.1 Flash?

Ling 3.1 Flash is a inclusionAI model available through OpenRouter. It supports text inputs and is suited for budget workloads.

How much does Ling 3.1 Flash cost on OpenRouter?

The 2026-10-10 snapshot lists $0 per 1M input tokens and $0 per 1M output tokens. Verify current pricing on OpenRouter before production.

How large is Ling 3.1 Flash's context window?

Ling 3.1 Flash offers 262K tokens of context in the current snapshot. Provider-specific limits and surcharges may apply.

Pricing and context data are based on a 2026-10-10 snapshot. Provider limits can change; always verify on OpenRouter before production.

Send us a link

A project built with Jev, a post, a video, a correction, a tip. We open the link, check it says what you said it says, and write the entry ourselves.

Required a link, and an email to reply to. Optional everything else.

Add context — all optional

One or two sentences about what it does, in your words. We write the entry ourselves.

Cost, latency, a benchmark — anything you measured. We attribute these to you.

We store what you type and email it to ourselves. No IP address, no user agent, no referrer — the same rule as the mailing list.

What happens to what you send →