Skip to content

Gemini 3.5 Flash-Lite

Last updated: 37m ago

by Google · Pricing

High confidence

Gemini 3.5 Flash-Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

77
Overall Score244
Rank #77 of 320 in Coding(244 24h)
Top 24% · Methodology v3
Score Trend
77/100
14-day history
API Pricing
$0.3/M in
$2.5/M out
Context Window
1.0M
65.5K max output
1.0M token context
Released 2026-07-21
#77range #61-#93Top 24%
#1#320

Signal Overview

Benchmarks75Capabilities100Pricing98Recency100Context96Output80

Score Breakdown

SignalStrengthWeightImpact
Benchmarksjust now
75
30%+22.4
Capabilitiesjust now
100
20%+20.0
Recencyjust now
100
15%+15.0
Pricingjust now
98
15%+14.6
Context Windowjust now
96
10%+9.6
Output Capacityjust now
80
10%+8.0

Benchmark Performance

Benchmark Scores(0 benchmarks + Arena Elo)

LMSYS Arena Elo

1460

Percentile

93.3

Weight

30%

No task benchmark data available yet for this model.

View all benchmarks

Capabilities

Reasoning
Vision
Function Calling
JSON Mode
Streaming
Web Search
Image Output

Modalities

Input
text
image
video
file
audio
Output
text

Recent Google releases

View this model against the provider’s recent shipping cadence.

Reviews

Community and practitioner feedback adds real-world signal on top of benchmarks and pricing.

Reviews

Be the first to review this model

Share your experience with Gemini 3.5 Flash-Lite and help the community make better decisions.

Frequently Asked Questions

Gemini 3.5 Flash-Lite by Google excels in the Coding category, where it ranks #77 with a composite score of 77/100. Gemini 3.5 Flash-Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows. It is particularly strong in areas highlighted by its top benchmark performance and adoption metrics, making it suitable for both individual developers and enterprise teams looking for a reliable coding solution.
Gemini 3.5 Flash-Lite is priced at $0.30 per million input tokens and $2.50 per million output tokens (USD). Contact the provider for volume discounts and enterprise pricing. Pricing is competitive within the coding category and reflects the model's quality-to-cost ratio.
In the Coding category, Gemini 3.5 Flash-Lite holds rank #77 out of 320 models tracked. Its quality rank is #77 and adoption rank is #77. You can use our comparison tool at /compare to see detailed side-by-side metrics with specific alternatives. Key differentiators include its composite scoring across benchmarks, community sentiment, and real-world adoption rates.
Gemini 3.5 Flash-Lite has been evaluated across 6 different signals. Its strongest areas include Capabilities (100/100), Benchmarks (75/100), Pricing (98/100). These scores are derived from industry-standard benchmarks, community ratings, and real-world performance metrics. The composite score of 77/100 reflects a weighted combination of all tracked signals.
Gemini 3.5 Flash-Lite is a paid model, though some providers may offer trial credits or limited free tiers for evaluation. Check Google's website for current free tier availability and promotional offers.
Gemini 3.5 Flash-Lite supports a 1,049K token context window (1,048,576 tokens total). That translates to roughly 786,432 words in a single prompt. This is large enough to process entire codebases, research papers, or long conversation histories in one shot.
Gemini 3.5 Flash-Lite can generate up to 66K output tokens (65,536 tokens) per response. That is roughly 49,152 words. This is enough for generating complete code files, detailed reports, or long-form content in a single response.
Gemini 3.5 Flash-Lite supports image understanding (vision), function/tool calling, structured JSON output, extended reasoning/chain-of-thought, web search, streaming responses. Function calling lets you integrate it with external APIs and tools programmatically. Vision support means it can analyze images, screenshots, and diagrams alongside text. These capabilities determine which workflows and integrations the model can handle natively.
Gemini 3.5 Flash-Lite was developed by Google. It was released on July 21, 2026. You can access it through Google's API. Check our provider page for all models from Google and how they compare against each other.
Pick Gemini 3.5 Flash-Lite when you need a solid balance of cost and capability for everyday development tasks, content generation, and standard API integrations. If your task is straightforward text completion or classification, a cheaper model might give you 90% of the quality at a fraction of the price. Run a quick benchmark on your actual use case before committing.
You can access Gemini 3.5 Flash-Lite through Google's API using standard HTTP requests or their official SDK. Most providers support OpenAI-compatible endpoints, so switching between models often requires changing just the model name in your API call. Streaming is supported for real-time token-by-token output. For production use, implement proper error handling, rate limiting, and cost monitoring.

Key Info

ProviderGoogle
CategoryCoding
Max Output65.5K tokens
LicenseProprietary
Statusstable
Data updated: Jul 21, 2026Benchmarks: Jul 21, 2026Status: Jul 21, 2026

Pricing Tools

Pricingper 1M tokens
Best value
82% cheaper than category average
Input
$0.30
-87% vs avg
Output
$2.50
-77% vs avg

Cost Estimator

Input: 70%Output: 30%
Est. monthly cost$9.60
Category average$49.78

You save $40.18/month vs category average

Access & Availability

Hosted APIAvailable
PlaygroundAvailable
Open weightsNo
Hugging FaceNot listed

Why This Rank

+Benchmarks
+Capabilities
+Recency
+Pricing

Similar Models

View all