Loading...
stable-gemini-3.8-flashgoogleAvailableStableReleased 2026-09-02Served on a dedicated channel we operate directly rather than the shared pool: no silent model substitution, no waiting behind pool contention, and prompt caching and thinking pass through intact. Same Gemini 3.8 Flash weights — the premium buys a predictable path to them, not a different model.
Measured availability 99.2%
This model runs on a channel we operate directly, separate from the shared Token Codex pool.
This is a single direct connection, not a multi-provider pool: it trades the pool's failover for predictable behaviour. If this channel is down, requests to this model fail rather than falling back.
Measured on our side. A dash means the window has no measurement yet — never a failure.
Official vs our price (rate 0.4x)
| Tier | Our price | Discount |
|---|---|---|
| Input | $0.300/M | -60% |
| Output | $1.50/M | -60% |
| Cache read | $0.030/M | -60% |
| Cache write | $0.300/M | -60% |
Stable dedicated channel
Available/v1/messages