Google

Google: Gemini 3.1 Flash Lite

google/gemini-3.1-flash-lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic workflows, simple data extraction, and applications where responsiveness and API cost are the primary constraints. Supports full thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs. Priced at half the cost of Gemini 3 Flash.

  • Context window: 1,048,576 tokens
  • Input: text, image, video, file, audio
  • Output: text
  • Pricing: $0.25/M input tokens, $1.5/M output tokens

View on OpenRouter. Model data sourced from OpenRouter.