google/gemini-3.1-flash-tts-preview

google/gemini-3.1-flash-tts-preview
Audio
by google

Google's fast, cost-efficient text-to-speech model for real-time audio generation, accessed via OrcaRouter.

Inputtext
p50 TTFT10.00 s
INPUT$1.00/ 1M tokens
OUTPUT$20.00/ 1M tokens
p50 TTFT10.00 s7d
p95 TTFT10.00 s7d
TRAFFIC125.1Ktokens / 7d

google/gemini-3.1-flash-tts-preview is a text-to-speech model from Google, part of the Gemini Flash family. It accepts text input and generates spoken audio output. The model is optimized for speed…

What is google/gemini-3.1-flash-tts-preview?

Who is this model designed for?

What are the key input and output specifications?

How does OrcaRouter provide access to this model?

Code samples

Call from any SDK

OpenAI-compatible — keep the SDK you already use

  • OpenAI SDKhttps://api.orcarouter.ai/v1

Supported parameters

  • voice

Pricing

Input / 1M tokens$1.00
Output / 1M tokens$20.00
CurrencyUSD

Cost calculator

Tokens / month10MM
Input share70%%
Estimated / month $67.00

Estimate based on list price

Token & cost estimator

Input tokens: 20Cost per request: $0.0100

Estimate only — actual token counts depend on the provider's tokenizer.

Performance

p50 TTFT
10.00 s
Output speed
3071 tok/s
p95 TTFT
10.00 s
Error rate
3.0%

Public benchmarks

Source: Design Arena

How it compares

google/gemini-3.1-flash-tts-previewGemini 3.1 Pro PreviewGemini 3.1 Pro Preview Custom ToolsGemini 3 Flash Preview
Input $/M$1.00$2.00$4.00$0.50
Output $/M$20.00$12.00$18.00$3.00
Context1.0M1.0M1.0M
Quality2/1010/1010/109/10
Compare side-by-sideCompare side-by-sideCompare side-by-sideCompare side-by-side

FAQ

How much does it cost to use google/gemini-3.1-flash-tts-preview on OrcaRouter?
Pricing is $1.00 per 1 million input tokens and $20.00 per 1 million output tokens, billed at Google's provider rate with zero markup from OrcaRouter. There are no additional platform fees. You are charged for both input and output tokens consumed.
What is the context window (maximum input length) for this model?
The maximum input length (context window) for google/gemini-3.1-flash-tts-preview is not specified in the available facts. As a preview model, limits may be imposed by Google. You should test with your expected text lengths or consult Google's documentation for the preview.
What are the main strengths of this TTS model?
Strengths include very low input token cost ($1/M tokens), fast generation due to Flash architecture, and simple text-to-speech conversion. It is designed for real-time applications and batch processing of short audio. OrcaRouter provides zero-markup pricing and OpenAI-compatible API access.
How does this model compare to other TTS models like OpenAI's tts-1?
OpenAI's TTS models charge per character (~$0.015/1k characters) and have known voices. Google's Flash TTS uses token-based pricing; input tokens are cheap but output tokens are expensive. Latency is expected to be similar. Quality and voice options vary; you should test both with your content to compare naturalness.
What data handling policies apply when using this model through OrcaRouter?
Data handling follows Google's policies for Gemini models and OrcaRouter's privacy policy. OrcaRouter does not log or store your prompts or outputs beyond what is necessary for API routing. Google may process data according to its own terms. Review both OrcaRouter's terms and Google's model-specific terms for compliance.
How can I call this model using the OpenAI Python SDK?
Set your API base to https://api.orcarouter.ai/v1, use your OrcaRouter API key, and set model="google/gemini-3.1-flash-tts-preview". Example: client = OpenAI(base_url="https://api.orcarouter.ai/v1", api_key="sk-..."). Then send a chat completion request. The audio output will be in the response; you may need to decode base64 or use streaming.
Does this model support streaming audio output?
Streaming support is not confirmed in the provided facts. Many TTS models support streaming by sending audio chunks as they are generated. You can test by setting stream=true in your API request. If supported, you will receive incremental response data. Otherwise, the complete audio blob is returned at the end.
What input modalities does this model accept?
This model accepts text input only. It converts the text into audio output. There is no support for other modalities like image or video input. Ensure your prompts are plain text.
Is there any context caching available for this model on OrcaRouter?
The available facts do not mention caching features for this model. OrcaRouter may offer caching for some models, but it is not specified. To reduce costs for repeated prompts, consider caching generated audio at your application layer.
How do I migrate from using Google's native API to OrcaRouter's API for this model?
Change your base URL to https://api.orcarouter.ai/v1, use your OrcaRouter API key, and set model to "google/gemini-3.1-flash-tts-preview". Adjust authentication and endpoint structure. OrcaRouter's API is OpenAI-compatible so you can use standard SDKs. No code changes beyond these parameter updates are needed.

Embed this badge

google/gemini-3.1-flash-tts-preview$1.00/M in10000ms p50via OrcaRouter
HTML <a href="https://www.orcarouter.ai/models/google/gemini-3.1-flash-tts-preview" target="_blank"> <img src="https://www.orcarouter.ai/embed/google/gemini-3.1-flash-tts-preview.svg" alt="google/gemini-3.1-flash-tts-preview on OrcaRouter" /> </a>
Markdown [![google/gemini-3.1-flash-tts-preview](https://www.orcarouter.ai/embed/google/gemini-3.1-flash-tts-preview.svg)](https://www.orcarouter.ai/models/google/gemini-3.1-flash-tts-preview)

Model card as data

GET /api/public/models/google/gemini-3.1-flash-tts-previewOpen
Machine-readable:/llms.txt/llms-full.txt