gemini-omni-flash-api
Installation
SKILL.md
Gemini Omni Flash Skill
This skill uses the Gemini Omni Flash model (gemini-omni-flash-preview) to perform text to video generation, image to video generation and video editing.
[!WARNING] Important Regional Restrictions: Uploading videos to use for video edits is NOT available in the EEA, Switzerland, the United Kingdom, and some US states. If a video-to-video edit completes quickly with empty outputs (
total_output_tokens: 0or no video content), it is likely due to this restriction.
Core capabilities
- Video editing and refinement: Editing existing videos (maximum duration 10 seconds), applying stylistic changes, or performing inpainting/outpainting.
- Text to video: Generating videos from a text prompt.
- First-frame to video: Generating videos from a single input image.
- Image-referenced generation: Using style, character, or object references from images to guide video generation.
Workflow
- Analyze request: Determine the target task (e.g., first-frame-to-video, reference-guided editing) and identify any input media assets.
- Run SDK scripts: