Gemini 3.1 Flash-Lite
FeaturedBy Google · Visit provider →
Gemini 3.1 Flash-Lite is Google's most cost-efficient model, designed for high-volume production workloads. It handles classification, extraction, and simple generation tasks at an extremely competitive price point.
Key Strengths
- Google's lowest-cost Gemini model
- Excellent throughput for batch processing
- 1M context at a budget price
- Good for simple multimodal tasks
Best Use Cases
Pricing Details
| Metric | Value |
|---|---|
| Input (per 1M tokens) | $0.25 |
| Output (per 1M tokens) | $1.50 |
| Context window | 1M tokens |
| Provider | |
| Architecture | Lightweight Gemini architecture, 1M context, optimized for maximum cost efficiency |
High-volume, cost-efficient
Benchmark Performance
Source: Artificial Analysis (independent benchmarking). Updated 2026-08-08.
Recent AI News
Databricks wanted $1B, got $15B in demand — so it raised $5B at a $190B valuation
TechCrunch AI · Aug 13, 2026
IBM bets big on OpenAI, will retrain tens of thousands of consultants
TechCrunch AI · Aug 13, 2026
Google slashes Gemini 3.7 Flash pricing by 50% to $0.75 per 1M tokens
MarkTechPost · Aug 13, 2026
OpenAI's CRO exits after 9 months as Wiz COO Dali Rajic takes over sales
TechCrunch AI · Aug 13, 2026
Google Ships Gemini 3.7 Flash Three Weeks After 3.6 — Coding Scores Jump 9 Points
Ars Technica AI · Aug 13, 2026
Get smarter about AI
The sharpest AI news, curated daily. Delivered free to your inbox.
Pricing data is sourced from official API documentation and updated weekly. Actual costs may vary based on prompt caching, batch processing, service tiers, and regional pricing. Inblix is not affiliated with Google.