- Beats the older 2.5 flash
- Runs at 363 tokens/sec (insane at that price)
- $0.25 per 1M input tokens
-Built for high-volume agent tasks, translation and simple data processing
- Faster and cheaper than the last model.
Insane model for high workload tasks at a low cost.
Google DeepMind@GoogleDeepMindGemini 3.1 Flash-Lite has landed. It’s our most cost-efficient Gemini 3 series model yet, built for intelligence at scale. Here’s what’s new 🧵





