Skip to content
Explore models

Video

Gemini Omni Flash 1.1

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video with synchronized native audio from a text prompt, grounded in Gemini's real-world knowledge and physics understanding, with cinematic camera control expressed in natural language.

  • Text to video
  • Image to video
Create with Gemini Omni Flash 1.1
Preview for Gemini Omni Flash 1.1

Catalog preview for Gemini Omni Flash 1.1. Review your own result before publishing.

Use Gemini Omni Flash 1.1 through the API.

Start with this model ID for the endpoint featured on this page:

google/gemini-omni-flash/v1.1/text-to-video

Available modes

Check the current model schema for supported inputs and options. Other tasks may use different model IDs.

From quote to result.

Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.

Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.