AI Pulse by Inblix

Gemini 3.1 Flash-Lite

Featured

By Google · Visit provider →

Gemini 3.1 Flash-Lite is Google's most cost-efficient model, designed for high-volume production workloads. It handles classification, extraction, and simple generation tasks at an extremely competitive price point.

Input Price
$0.25
per 1M tokens
Output Price
$1.50
per 1M tokens
Context Window
1M
tokens
Calculate cost for Gemini 3.1 Flash-Lite

Key Strengths

  • Google's lowest-cost Gemini model
  • Excellent throughput for batch processing
  • 1M context at a budget price
  • Good for simple multimodal tasks

Best Use Cases

Large-scale content classificationDocument parsing and extractionHigh-volume chat moderationCost-sensitive production pipelines

Pricing Details

Metric Value
Input (per 1M tokens) $0.25
Output (per 1M tokens) $1.50
Context window 1M tokens
Provider Google
Architecture Lightweight Gemini architecture, 1M context, optimized for maximum cost efficiency

High-volume, cost-efficient

Benchmark Performance

GPQA Diamond 82.2 / 100
Coding Index 34.7 / 100
Intelligence Index 25.6 / 100
Humanity's Last Exam 17.2 / 100

Source: Artificial Analysis (independent benchmarking). Updated 2026-08-08.

Recent AI News

Get smarter about AI

The sharpest AI news, curated daily. Delivered free to your inbox.

Pricing data is sourced from official API documentation and updated weekly. Actual costs may vary based on prompt caching, batch processing, service tiers, and regional pricing. Inblix is not affiliated with Google.