Google: Gemini 3.8 Flash

google/gemini-3.8-flash
Compare Model Lab

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Modalities
Text + Image + Video + File + Audio → Text
Input / Output Price$0.375 / $1.88From · per 1M tokens · see provider prices
Context1.05Mtokens · maximum across providers
ReleasedSep 2, 2026OpenRouter public metadata
Availability & data sources Listed in CLSSAI · 6 reference endpoints
ACCESS

Where this model works

OpenRouter testing and CLSSAI API routing are independent decisions with separate sources.

OpenRouter testTestable with your OpenRouter key
Testable

Exact capability metadata is available. CLSSAI can prepare the Audio understanding request without guessing the model contract.

Evidence
Exact ID google/gemini-3.8-flash · /chat/completions
Mapped mode
Audio understanding
CLSSAI APICallable through CLSSAI API
Callable

The exact model ID is listed in the current CLSSAI public directory. This proves current route eligibility, not a completed inference for every mode.

Evidence
Exact ID google/gemini-3.8-flash is directory-listed
Audio analysis test status

The request contract and frontend path are ready. A real OpenRouter output has not yet been recorded for this exact model and mode.

Request contract
Build verified
/chat/completions · Schema tests · 2026-08-05
Frontend path
Available
Audio analysis form · 2026-08-05
Real model output
Not yet verified
No keyed OpenRouter smoke record

Providers

Compare published prices and endpoint capabilities. Select a provider for context limits, caching prices and supported parameters.

$0.375$1.88$0.038----99.78%
$0.375$1.88$0.038----86.6%
$0.75$3.75$0.075----99.90%
$0.75$3.75$0.075----99.16%
$1.35$6.75$0.135----99.87%
$1.35$6.75$0.135----99.52%

USD per 1M tokens. Prices and metrics describe OpenRouter endpoints, not CLSSAI routing. “--” means no measurement was supplied. Uptime: last 30 minutes.

Pricing

Published base rates across providers, in USD per 1M tokens. Actual cost depends on the provider, caching and prompt length.

Input / 1M tokens$0.375 – $1.35
Output / 1M tokens$1.88 – $6.75
Cache read / 1M tokens$0.038 – $0.135

Effective Pricing: request-weighted costs and historical prices are not supplied by the public source.

Performance

Latency and throughput history are not supplied by the public endpoint API. Check OpenRouter for its latest performance charts.

View on OpenRouter

Uptime

Successful requests in the last 30 minutes, as reported by OpenRouter. A snapshot does not show historical reliability.

Google AI Studio Flex99.78%
Google Global Flex86.6%
Google AI Studio99.90%
Google Global99.16%
Google AI Studio Priority99.87%
Google Global Priority99.52%

Benchmarks

No scored evaluations are included in the public model metadata. View the source for benchmarks and their methodology.

View on OpenRouter

Apps

Application rankings require verified traffic data, which is not included in this source.

View on OpenRouter

Activity

Historical token and request volumes are not supplied by the public API. OpenRouter shows activity on its own network.

View on OpenRouter

Frequently asked questions

Provider availability, context limits and API capabilities.

Which providers serve Google: Gemini 3.8 Flash?

Google AI Studio Flex, Google Global Flex, Google AI Studio, Google Global, Google AI Studio Priority, Google Global Priority are listed in the OpenRouter reference data. Check CLSSAI availability above before making a request.

What context length is available for Google: Gemini 3.8 Flash?

The largest available context window is 1,048,576 tokens.

Which API parameters does Google: Gemini 3.8 Flash support?

The providers list 12 distinct supported parameters. Open an endpoint row to inspect them.

Explore

Continue comparing models or start integrating.