Gemini Omni Flash Preview launched on Google Cloud
Google Cloud has launched a preview of Gemini Omni Flash, a multimodal model optimised for video, images, and text. The model is available via the Gemini Enterprise Agent Platform.

What happened?
Google Cloud has introduced Gemini Omni Flash (Preview), a multimodal AI model capable of handling video, image, and text-based tasks. The model is specifically optimised for video generation and can produce video output alongside text responses from a single model. It is available as part of the Gemini Enterprise Agent Platform.
Key facts
| Modellnamn | Gemini Omni Flash (Preview) |
|---|---|
| Plattform | Gemini Enterprise Agent Platform, Google Cloud |
| Huvudfokus | Videogenerering, bild- och textuppgifter |
| Maximala input tokens | 131 072 |
| Maximala output tokens | 57 920 |
”Gemini Omni Flash (Preview) is a multimodal model designed for video, image, and text tasks. It is optimized for video generation, offering video output alongside text responses in a single model.”
”Deploy example app" requires a Google Cloud project with billing and Agent Platform API enabled.”
Why it matters
The launch of Gemini Omni Flash represents a further development in multimodal AI models, with a clear focus on video. By integrating video input and output directly into the model, developers can create more dynamic and interactive AI applications. This capability can reduce the complexity of building AI solutions that require advanced media handling.
Who is affected?
This launch primarily affects developers and companies using or planning to use the Google Cloud platform for AI development, particularly those working in media, content creation, or applications requiring multimodal interactions. Users of the Agent Platform API require a Google Cloud project with billing enabled.
Impact on the EU
Gemini Omni Flash Preview is available via the Google Cloud platform. Availability within the EU follows Google's general cloud services, but users must ensure their use complies with the General Data Protection Regulation (GDPR) and the upcoming EU AI Act.
What else you should know
The model ID for the preview is `gemini-omni-flash-preview`. It features a maximum input of 131,072 tokens and an output of 57,920 tokens. Features such as system instructions, Gemini Live API, and structured output are not supported in this preview version; however, video generation, video editing, and video referencing are confirmed features.
Quick answers about this story
Vad har hänt?
När hände det?
Varför spelar det roll?
Vilka funktioner stöds för videogenerering?
Vem kan använda Gemini Omni Flash Preview?
The link opens in a new window and leads to the publisher's own site.
Källan är en aggregator eller syndikering — vi rekommenderar att verifiera hos primärutgivaren.
Get similar news straight to your inbox
The reader's room
Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.
Sign in to submit a comment or question.
Read the article through your role
- Decide whether this affects strategy over 6–12 months or is just noise.
- Discuss with leadership: do we own the right question or does ownership need to move?
- Ask: what risk are we taking by NOT acting on this this quarter?
Generated angle — not editorial analysis of "Gemini Omni Flash Preview launched on Google Cloud"