gemini-omni-flash-api

Installation
SKILL.md

Gemini Omni Flash Skill

This skill uses the Gemini Omni Flash model (gemini-omni-flash-preview) to perform text to video generation, image to video generation and video editing.

[!WARNING] Important Regional Restrictions: Uploading videos to use for video edits is NOT available in the EEA, Switzerland, the United Kingdom, and some US states. If a video-to-video edit completes quickly with empty outputs (total_output_tokens: 0 or no video content), it is likely due to this restriction.

Core capabilities

  1. Video editing and refinement: Editing existing videos (maximum duration 10 seconds), applying stylistic changes, or performing inpainting/outpainting.
  2. Text to video: Generating videos from a text prompt.
  3. First-frame to video: Generating videos from a single input image.
  4. Image-referenced generation: Using style, character, or object references from images to guide video generation.

Workflow

  1. Analyze request: Determine the target task (e.g., first-frame-to-video, reference-guided editing) and identify any input media assets.
  2. Run SDK scripts:
Installs
619
GitHub Stars
3.9K
First Seen
Jun 30, 2026
gemini-omni-flash-api — google-gemini/gemini-skills