google/gemini-3.1-flash-lite

google/gemini-3.1-flash-lite
by google
p50 TTFT805 ms
INPUT$0.25/ 1M tokens
OUTPUT$1.50/ 1M tokens
p50 TTFT805 ms7d
p95 TTFT3.07 s7d
TRAFFIC2.8Mtokens / 7d

Code samples

Call from any SDK

OpenAI-compatible — keep the SDK you already use

  • OpenAI SDKhttps://api.orcarouter.ai/v1
  • Gemini SDKhttps://api.orcarouter.ai
from openai import OpenAI

client = OpenAI(
    base_url="https://api.orcarouter.ai/v1",
    api_key="$ORCAROUTER_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-3.1-flash-lite",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Pricing

Input / 1M tokens$0.250
Output / 1M tokens$1.50
Cache read / 1M$0.025
CurrencyUSD

Cost calculator

Tokens / month10MM
Input share70%%
Estimated / month $6.25 · With prompt caching $5.46

Estimate based on list price

Token & cost estimator

Input tokens: 20Cost per request: $0.000755

Estimate only — actual token counts depend on the provider's tokenizer.

Performance

p50 TTFT
805 ms
Output speed
77.6 tok/s
p95 TTFT
3.07 s
Error rate
0%

Public benchmarks

Source: Design Arena

How it compares

google/gemini-3.1-flash-liteGemini 3.1 Pro PreviewGemini 3.1 Pro Preview Custom ToolsGemini 3 Flash Preview
Input $/M$0.25$2.00$4.00$0.50
Output $/M$1.50$12.00$18.00$3.00
Context1.0M1.0M1.0M
Quality5/1010/1010/109/10
Compare side-by-sideCompare side-by-sideCompare side-by-sideCompare side-by-side

FAQ

How much does google/gemini-3.1-flash-lite cost on OrcaRouter?
google/gemini-3.1-flash-lite is priced at $0.25 per 1M input tokens and $1.50 per 1M output tokens via OrcaRouter. Pricing is pulled live from the routing layer.
What is google/gemini-3.1-flash-lite's context window?
google/gemini-3.1-flash-lite supports a context window of — tokens. Use long-context features (RAG, summarisation) up to that limit.
How do I call google/gemini-3.1-flash-lite via the OpenAI SDK?
Set OpenAI base_url to https://api.orcarouter.ai/v1, supply your OrcaRouter API key, and pass model="google/gemini-3.1-flash-lite" in the chat.completions.create call.
Does OrcaRouter rate-limit google/gemini-3.1-flash-lite?
Per-model rate limits follow your OrcaRouter plan. Free tiers ship with conservative caps; paid tiers lift them. Check /pricing for current quotas.

Embed this badge

google/gemini-3.1-flash-lite$0.25/M in805ms p50via OrcaRouter
HTML <a href="https://www.orcarouter.ai/models/google/gemini-3.1-flash-lite" target="_blank"> <img src="https://www.orcarouter.ai/embed/google/gemini-3.1-flash-lite.svg" alt="google/gemini-3.1-flash-lite on OrcaRouter" /> </a>
Markdown [![google/gemini-3.1-flash-lite](https://www.orcarouter.ai/embed/google/gemini-3.1-flash-lite.svg)](https://www.orcarouter.ai/models/google/gemini-3.1-flash-lite)

Model card as data

GET /api/public/models/google/gemini-3.1-flash-liteOpen
Machine-readable:/llms.txt/llms-full.txt