Gemini 1.5 Flash is a foundation model that performs well at a variety of multimodal tasks such as visual understanding, classification, summarization, and creating content from image, audio and video. It's adept at processing visual and text inputs such as photographs, documents, infographics, and screenshots.
Gemini 1.5 Flash is designed for high-volume, high-frequency tasks where cost and latency matter. On most common tasks, Flash achieves comparable quality to other Gemini Pro models at a significantly reduced cost. Flash is well-suited for applications like chat assistants and on-demand content generation where speed and scale matter.
Usage of Gemini is subject to Google's Gemini Terms of Use(opens in new tab).
#multimodal
Modalities
Context
1M
Released
May 14, 2024
Knowledge Cutoff
May 2024
Token volume and request traffic to this model over time.
Gemini 1.5 Flash is a foundation model that performs well at a variety of multimodal tasks such as visual understanding, classification, summarization, and creating content from image, audio and video. It's adept at processing visual and text inputs such as photographs, documents, infographics, and screenshots.
Gemini 1.5 Flash has a 1,000,000 token context window.
Gemini 1.5 Flash accepts text and images as input and returns text.
Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash Lite and 26 more are other text models from Google.
Gemini 1.5 Flash was released on May 14, 2024. Its knowledge cutoff is May 31, 2024.