Back to Model
G

Gemini 3.5 Flash-Lite

Model Sep 22, 2026

A lightweight Gemini tier designed for cost-sensitive, high-volume inference.

Visit primary link 1 source attached

Gemini 3.5 Flash-Lite at a glance

A lightweight Gemini tier designed for cost-sensitive, high-volume inference.

  • Provider: Google
  • Focus: fast, efficient
  • Access: Provider catalog

Access and deployment

Gemini 3.5 Flash-Lite is listed here under Google’s model catalog. Use that catalog to confirm the exact API identifier and current availability; the display name on this page is not an API alias.

Evaluation checklist

Measure quality at your target latency and concurrency. For local deployment, check memory requirements at the context length and quantization you intend to use.

Preview aliases change quickly, so verify the stable model ID, supported inputs, output modes and deprecation schedule.

Pricing and usage limits

Pricing, quotas and license conditions are not independently verified in this entry. Check the linked official source before choosing a production deployment. For a hosted deployment, check input and output charges, regional availability and rate limits for the exact model ID.

Sources