Gemini 3.5 Flash Lite
Current version: 3.5 updated
The most cost-effective model in the Gemini 3.5 line
What's new in Gemini 3.5 Flash Lite
- The most affordable option in the Gemini 3.5 line
- Upgraded agentic skills and accurate tool calling
- 1M-token context at a minimal price
- Built for subagents and background tasks in multi-agent workflows
What it's the best tool for
- Fastest model in the Gemini 3.5 line at up to 350 tokens per second
- 1M-token context window
- Terminal-Bench 2.1: 54% vs 31% for the previous Lite
- OSWorld-Verified 74% and SWE-Bench Pro 54.2% — ahead of Gemini 3 Flash
- Built for agentic search and document processing
- The most affordable tier in the Gemini 3.5 family
When to reach for something else
- Trails bigger Gemini models and flagships on complex multi-step reasoning
- No deep-thinking mode — the model is tuned for speed
- Google content filters on sensitive topics
- Heavy coding on large repos is better served by a higher tier
How Gemini 3.5 Flash Lite responds
Four scenarios where it pays for itself
More about Gemini 3.5 Flash Lite
Gemini 3.5 Flash Lite Online — Google's Fastest Lite Model
Gemini 3.5 Flash Lite is the fastest and most affordable model in Google's Gemini 3.5 line, released July 21, 2026. It streams up to 350 tokens per second and handles a 1M-token context. Try it online on NetRoom — right in the browser, no VPN required.
What it does well
Google built Flash Lite for high-throughput, low-latency workloads: agentic search, document processing, bulk classification and data extraction. The jump over the previous Lite generation is real: Terminal-Bench 2.1 at 54% vs 31%, long-context GDM-MRCR v2 at 72.2% vs 60.1%, and GDPval-AA v2 at 1140 points vs 642. It even beats Gemini 3 Flash — a model a tier above — scoring 54.2% on SWE-Bench Pro and 74% on OSWorld-Verified.
Speed is the feature
At 350 tokens per second, answers land faster than you can read the first line. That is exactly what chatbots, support desks and live interfaces need: no spinners, no dead air. Add the million-token window and a full contract stack, a hundred-page report or a large slice of a codebase fits into a single request.
When to pick Gemini 3.5 Flash Lite
Choose it when speed and cost matter most: PDF summarization, contract parsing, fast subagents inside multi-agent pipelines, bots serving thousands of requests a day. It is the cheapest tier in the Gemini 3.5 family, so bulk workloads get noticeably lighter on the budget. For deep multi-step reasoning or heavy coding on big repos, look at the flagship models in the NetRoom catalog.
How to start
Sign up on NetRoom, top up your balance, pick Gemini 3.5 Flash Lite from the model list and go. You pay for actual tokens only — no subscriptions, no hidden terms.
What changed Gemini 3.5 Flash Lite
- + Added text model Gemini 3.5 Flash Lite (Google).
Version history of Gemini 3.5 Flash Lite
| Version | Date | What changed |
|---|---|---|
| 3.5 current |
|
Try Gemini 3.5 Flash Lite
right now
Free access to basic models. No card, no obligations.