Skip to main content
Generate a video with audio from a text prompt using Google’s Gemini Omni Flash model. Optionally provide reference images and/or videos to guide or edit the result. Describe the desired length (3-10s) and aspect ratio (16:9 or 9:16) directly in the prompt.

Inputs

Notes:
  • If an image input contains multiple frames, each frame counts toward the maximum of 14 images.
  • Reference videos are always sent inline to the model. Up to 10 images are uploaded as URLs; any images beyond the first 10 are sent inline. The total inline media size (all videos plus any inline images) must stay under about 90 MB, otherwise the node raises an error.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 1b7ca51d07cfb6a166cfed2a7e7174fd62f3290abcc1bdfdce94369dda242d3f