A Flash model is a relatively small Geminimodel optimized for speed, low latency, and efficient responses. It is intended for applications where fast interaction is more important than using the largest available model.
A family of relatively small Gemini models optimized for speed
and low latency. Flash models are designed for a wide
range of applications where quick responses and high throughput are crucial.