Skip to main content
Generate a video with audio from a text prompt using Google’s Gemini Omni Flash model. Optionally provide reference images and/or videos to guide or edit the result. Describe the desired length (3-10s) and aspect ratio (16:9 or 9:16) directly in the prompt.

Inputs

Common Inputs

Omni Flash Inputs

Reference Inputs

Notes:
  • If an image input contains multiple frames, each frame counts toward the maximum of 14 images.
  • When reference images or videos are provided, the total encoded media size must stay under about 90 MB; otherwise the node raises an error.
  • When no reference images or videos are provided, the node generates the video from the text prompt alone.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 648844868affb68298d2eac8ac20095bfe378d32e721396781de330ef6a6d69f