Gemini 3.1 Flash-Lite
Google's Gemini 3-series model for high-volume lightweight tasks, first released in preview on March 3, 2026. Gemini API general availability and Google Cloud's GA announcement are dated May 7. It supports text, image, video, audio and PDF input with text output, a 1,048,576-token input limit and a 65,536-token output limit. Standard Gemini API pricing is $0.25/$1.50 per MTok (input/output) for text/image/video input and text output; audio input costs $0.50 per million tokens. Google recommends Gemini 3.5 Flash-Lite as its replacement and lists May 7, 2027 as the earliest shutdown date.
- Provider
- Context
- 1MCatalog · As of Aug 11, 2026Stale · 31d old
- yno subscription tier
- Pro
- Released
- Mar 3, 2026
- Speed
- Fast
- Reasoning
- Advanced
- Modality
- Multimodal
Benchmarks
Quality
- AA Intelligence Index
- AA Coding Index
- GPQA Diamond
- 82.2%Artificial Analysis · Date unknownAge unknown
Speed and latency
- Output speed
API pricing
- API input cost
- API output cost
Benchmark variant: Gemini 3.1 Flash-Lite
AA data retrieved Sep 11, 2026 · Artificial Analysis
Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology
Capabilities
- Text Generation
- Vision
- Video
- Code
- Speed
Use cases
- High-volume translation
- Audio-file transcription
- Structured data extraction
- Document summarization
Strengths
- High-volume lightweight tasks
- 1,048,576-token input limit
- Text, image, audio and video input
- Structured outputs
Best for
Real-time multimodal applications needing Gemini 3.x quality at production scale
No credit card required
Related models
Frequently asked
- What is Gemini 3.1 Flash-Lite?
- Google's Gemini 3-series model for high-volume lightweight tasks, first released in preview on March 3, 2026. Gemini API general availability and Google Cloud's GA announcement are dated May 7. It supports text, image, video, audio and PDF input with text output, a 1,048,576-token input limit and a 65,536-token output limit. Standard Gemini API pricing is $0.25/$1.50 per MTok (input/output) for text/image/video input and text output; audio input costs $0.50 per million tokens. Google recommends Gemini 3.5 Flash-Lite as its replacement and lists May 7, 2027 as the earliest shutdown date.
- How much does Gemini 3.1 Flash-Lite cost in yno.ai?
- Gemini 3.1 Flash-Lite is available on the Pro tier of yno.ai.
- What can Gemini 3.1 Flash-Lite do?
- Gemini 3.1 Flash-Lite is best for Real-time multimodal applications needing Gemini 3.x quality at production scale. Its main capabilities include Text Generation, Vision, Video, Code, Speed.
- How does Gemini 3.1 Flash-Lite compare to other models?
- Gemini 3.1 Flash-Lite excels at High-volume lightweight tasks and is recommended for High-volume translation, Audio-file transcription, Structured data extraction, Document summarization. See related models below.
- How do I use Gemini 3.1 Flash-Lite in yno.ai?
- Sign up for yno.ai, select Gemini 3.1 Flash-Lite from the model picker, and start chatting.