Skip to content
Dashboard

Gemini 3.8 Flash

Gemini 3.8 Flash is the next iteration in the Gemini 3 model family, featuring algorithmic improvements to its core reasoning foundation. It supports customizable thinking configurations to control the mix of quality, cost and latency.

Input and output price
50% off
Prices from: Input $0.75, Output $3.75, Per 1M tokens
24h uptime
Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({
model: 'google/gemini-3.8-flash',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingThroughput

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the docs for more info.

Your use is subject to Google's Terms & Privacy Policies.