Google logo

Google: Gemini 3.1 Flash Lite

google/gemini-3.1-flash-lite

Used in WordGPTTry in WordGPT

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Modalities

TextImageVideoFiles / PDFAudioText

In / out price

$0.25 / $1.5 per 1M

Context

1.05M

Released

May 7, 2026

Why use Gemini 3.1 Flash Lite

  • Understands images
  • Understands audio
  • Reads PDFs and other files
  • Step-by-step reasoning mode
  • Tool and function calling (agent-ready)
  • Structured (JSON) outputs
  • Very long context: whole books or codebases in one prompt

Pricing

Input$0.25 / 1M tokens
Output$1.5 / 1M tokens
Cached input (read)$0.025 / 1M tokens
Cache write$0.083 / 1M tokens
Image input$0.25 / 1M tokens

List prices via OpenRouter, checked daily. Real cost depends on the provider and on reasoning tokens.

Cost calculator

Pick a task, set how often you run it, and compare what it costs on each model.

ModelPer runPer monthPer 1,000 runs
Google: Gemini 3.1 Flash Lite$0.0032$3.20$3.20

Estimates from list prices via OpenRouter; real bills vary by provider and reasoning tokens.

Will it fit? Context window simulator

Each square is about one page (500 words). Colored squares are your content; grey squares are free space in the model's context window.

Your content

  • Novel (90,000 words)

Total: about 120,000 tokens (≈ 90,226 words). Token counts are estimates.

Models

Google: Gemini 3.1 Flash Lite

1,048,576 tokens ≈ 1,577 pages

Fits: uses 11% of the window (room for 8× this much).

Use Google: Gemini 3.1 Flash Lite in WordGPT

WordGPT already runs this model behind its writing and agent tools. Free plan, no credit card.

Try it free

More from Google