# Gemini 3.5 Flash

Gemini 3.5 Flash is a language model from Google with a 1M-token context window. Provider list price is $1.50 per million input tokens and $9 per million output; on Allocate you pay $1.60 and $9.63. It is a closed model served over API; the weights are not published.

## Pricing

| | Provider list | On Allocate |
| --- | --- | --- |
| Input, per M tokens | $1.50 | $1.60 |
| Output, per M tokens | $9 | $9.63 |

Prices checked 2026-07-21.

## Facts

| Field | Value |
| --- | --- |
| Lab | Google |
| Modality | Language |
| Context window | 1M tokens |
| License | Proprietary API |
| Open weights | No |
| Fine-tunable | No |
| Catalog id | google/gemini-3.5-flash |

## What a real workload costs

Take 1,000,000 requests a month at 1,200 input and 350 output tokens each: 1,200M input and 350M output tokens. At list prices that is 1,200 × $1.50 + 350 × $9 = $4,950 a month. Billed on Allocate it is $5,297.

## Where it fits

Google’s workhorse: 1M tokens of context, vision built in, and a $1.50 per million input list price. The default choice for high-volume routes where cost per request decides the experience.

- High-volume support and triage
- Document extraction at scale
- Vision and OCR pipelines

## Common questions

### How much does Gemini 3.5 Flash cost per million tokens?

Provider list price is $1.50 per million input tokens and $9 per million output tokens. On Allocate you pay $1.60 in and $9.63 out.

### What context window does Gemini 3.5 Flash have?

1,000,000 tokens (1M). At roughly 0.75 words per token, that is about 750k words of English text per request.

### Can I fine-tune Gemini 3.5 Flash?

No. Gemini 3.5 Flash is a closed model served over API; the weights are not published. If you want a model you can train and own, start from an open-weights base in the catalog and fine-tune that.

### How do I call Gemini 3.5 Flash on Allocate?

Send google/gemini-3.5-flash in the model field of the OpenAI-compatible endpoint at api.allocate.network/v1, or point a route name (like prod/support-agent) at it so you can swap the model later without a deploy.

---

[HTML page](https://allocate.network/models/gemini-3-5-flash) · [Machine-readable catalog](https://allocate.network/catalog.json)
