> For the complete documentation index, see [llms.txt](https://docs.pletor.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.pletor.ai/models/video-models/gemini-omni-flash-1.md).

# Gemini Omni Flash

### Overview

Generates from text, image, and video inputs, and lets you refine results through natural language.

Best suited for short clips and iterative editing.

Caps at 10 seconds. Prefer [Seedance 2.0](/models/video-models/seedance-2.0.md) or [Kling 3.0](/models/video-models/kling-3.0.md) for longer durations or production-grade cinematography.

#### Key updates

* **Multimodal referencing:** Combine text, image, and video inputs in one generation to control composition and maintain consistency.
* **Conversational video editing:** Refine and edit videos using natural language, no need to re-generate from a full prompt for small changes.
* **Real-world knowledge:** Draws on Gemini's general knowledge (history, biology, narrative logic) to construct more coherent scenes.

#### Weaknesses

* **10-second cap:** Generations are limited to 10 seconds; longer durations are planned but not yet available.
* **Limited aspect ratios**: 16:9 or 9:16.
* **Video reference limitation:** Short video references (up to 3s) are accepted by the request schema but aren't correctly processed by the model yet.
* **Consistency across scene changes:** Character consistency can degrade during scene changes or panning movements.
