Vertex model API pricing
Vertex Gemini 3.1 Flash-Lite API pricing
Vertex Gemini 3.1 Flash-Lite API pricing is $0.125 per 1M input tokens and - per 1M output tokens in the ByteCosts provider pricing index. A combined 1M input plus 1M output token estimate is not calculated because one of those prices is missing. Context window: -. Prompt cache prices: read $0.013, write -, 1h write - per 1M tokens. The confidence grade is A+: Read from the provider's official pricing page. The source was last checked 2026-07-19. A reference monthly cost is not calculated because either input or output price is missing. This is a planning estimate; real invoices can differ because taxes, discounts, commitments, gateway fees, and non-token charges are excluded.
Open the pricing index - Vertex / Gemini 3.1 Flash-Lite →
Pricing snapshot
| Field | Value |
|---|---|
| Provider | Vertex |
| Model | Gemini 3.1 Flash-Lite |
| Input / 1M tokens | $0.125 |
| Output / 1M tokens | - |
| Cache read / 1M tokens | $0.013 |
| Cache write / 1M tokens | - |
| Cache write 1h / 1M tokens | - |
| Context window | - |
| Reference monthly workload | not calculated |
The reference workload is 1M input tokens plus 250K output tokens per month. It excludes taxes, discounts, gateway fees, and non-token charges.
Reference workload examples
Each example is pure arithmetic from this committed record's input and output token prices. The assumptions are shown inline.
| Example | Assumed monthly tokens | Estimated token cost |
|---|---|---|
| Light reference | 1M input + 250K output tokens/month | not calculated |
| Production reference | 10M input + 2.5M output tokens/month | not calculated |
| Heavy reference | 100M input + 25M output tokens/month | not calculated |
Examples exclude taxes, discounts, commitments, gateway fees, cache effects, and non-token charges unless those charges are part of the committed token price record.
Gemini 3.1 Flash-Lite pricing across providers
Gemini 3.1 Flash-Lite is listed by 2 providers in the ByteCosts index. Per-million-token list prices, lowest input first.
| Provider | Input / 1M | Output / 1M |
|---|---|---|
| Vertex (this page) | $0.125 | - |
| Deep Infra | $0.250 | $1.50 |
When this model is cost-effective
ByteCosts cannot rank Gemini 3.1 Flash-Lite within Vertex by combined input+output price because one of the token prices is missing.
Cached input reads are listed at $0.013 per 1M tokens, 10% of the fresh input rate, so repeated cached prompts have a lower committed input-token cost than fresh prompts.
No context window is committed for this record, so there is no long-context cost signal to apply.
Cheaper alternatives
These records have a lower committed input+output list price than this page's record. Same-provider records and same-model cross-provider records are preferred when available.
| Record | Input / 1M | Output / 1M | Combined / 1M + 1M | Detail page |
|---|---|---|---|---|
| No lower-priced committed record | - | - | - | - |
Evidence and freshness
ByteCosts grades this row A+. Read from the provider's official pricing page. The record was first tracked on 2026-05-30 and last checked on 2026-07-19.
The source URL is kept with the record so readers can verify the provider page before making a purchasing or architecture decision. Prices change often, so quote the last-checked date with any copied number.
Watch-outs
- audio surcharge
Price history
| When | Field | Was | Now | Change |
|---|---|---|---|---|
| 2026-06-01 | cacheRead | new | $0.025 | new |
| 2026-06-01 | audioInput | new | $1.00 | new |
Frequently asked questions
Is Gemini 3.1 Flash-Lite free to use?
This ByteCosts row is not marked free. Use the listed token rates and verify the provider source for discounts, commitments, and account-specific terms.
Does this page include all possible charges?
No. The page focuses on source-backed token pricing and known cache/context fields. It does not include taxes, enterprise commitments, gateway markups, storage, retrieval, fine-tuning, or platform fees unless those charges are present in the source-backed row.
Vertex Gemini 3.1 Flash-Lite API pricing. ByteCosts. Updated 2026-07-19. https://bytecosts.com/tools/ai-provider-pricing/google-vertex--gemini-3.1-flash-lite/