gemini-3.5-flash-lite
Common Name: Gemini 3.5 Flash-Lite
Google's fastest, most cost-effective Gemini 3.5 model, delivering 350 output tokens/sec for high-throughput tasks like agentic search, document processing, and translation.
Specifications
Performance (7-day Average)
Pricing
Availability Trend (24h)
Performance Metrics (24h)
Similar Models
Google's most efficient workhorse model designed for speed and low-cost. Improved across key benchmarks for reasoning, multimodality, code and long context while being 20-30% more efficient.
Preview of Google's next-generation Gemini 3 Flash model, optimized for speed with frontier intelligence combined with superior search and grounding capabilities.
Google's most cost-efficient Gemini 3 series model, optimized for high-volume agentic tasks, translation, and simple data processing with 2.5X faster time to first token than 2.5 Flash.
Gemini 3.1 Flash Image generation model designed for speed and efficiency, effective for quick interactive image responses and high throughput.