> ## Documentation Index
> Fetch the complete documentation index at: https://docs.comfy.org/llms.txt
> Use this file to discover all available pages before exploring further.

# GeminiVideoOmni - ComfyUI Built-in Node Documentation

> Generate a video with audio from a text prompt using Google's Gemini Omni Flash model.

Generate a video with audio from a text prompt using Google's Gemini Omni Flash model. Optionally provide reference images and/or videos to guide or edit the result. Describe the desired length (3-10s) and aspect ratio (16:9 or 9:16) directly in the prompt.

## Inputs

### Common Inputs

| Parameter | Description | Data Type | Required | Range |
| - | - | - | - | - |
| `model` | The Gemini video model used to generate the video. | DYNAMIC\_COMBO | Yes | "Omni Flash" |
| `seed` | Seed controls whether the node should re-run; results are non-deterministic regardless of seed (default: 42). | INT | Yes | 0 to 2147483647 |

### Omni Flash Inputs

| Parameter | Description | Data Type | Required | Range |
| - | - | - | - | - |
| `prompt` | Describe the video to generate. Specify the length and aspect ratio directly in the prompt, e.g. "a 6-second clip in 16:9". Length may be 3-10 seconds; the aspect ratio must be 16:9 (landscape) or 9:16 (portrait). The output is 720p, 24 FPS, with audio. | STRING | Yes | Minimum 1 character after stripping whitespace |
| `temperature` | Controls randomness. Lower is more focused/deterministic, higher is more varied (default: 1.0). | FLOAT | No | 0.0 to 2.0 |
| `top_p` | Nucleus sampling: sample from the smallest token set whose cumulative probability reaches top\_p (default: 0.95). | FLOAT | No | 0.0 to 1.0 |

### Reference Inputs

| Parameter | Description | Data Type | Required | Range |
| - | - | - | - | - |
| `images` | Growable slot: connect one or more reference images (`image_1`...`image_14`) to guide or animate the video. Up to 14 images in total. | IMAGE | No | 0 to 14 images |
| `videos` | Growable slot: connect one or more reference videos (`video_1`...`video_3`) to guide or edit. Up to 3 videos, each up to 10 seconds long. | VIDEO | No | 0 to 3 videos, each max 10 seconds |

Notes:

* If an image input contains multiple frames, each frame counts toward the maximum of 14 images.
* When reference images or videos are provided, the total encoded media size must stay under about 90 MB; otherwise the node raises an error.
* When no reference images or videos are provided, the node generates the video from the text prompt alone.

## Outputs

| Output Name | Description | Data Type |
| - | - | - |
| `VIDEO` | The generated video with audio from the Gemini model. | VIDEO |
| `STRING` | Any text response from the model, such as reasoning or explanations. | STRING |

> This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! [Edit on GitHub](https://github.com/Comfy-Org/embedded-docs/blob/main/comfyui_embedded_docs/docs/GeminiVideoOmni/en.md)

***

**Source fingerprint (SHA-256):** `648844868affb68298d2eac8ac20095bfe378d32e721396781de330ef6a6d69f`


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.