Floyo
Floyo
Workflows
API
Pricing
Floyo
Floyo
Workflows
API
Pricing

Gemini Omni Flash 1.1 for Video Edit

Edit a video with a text instruction using Gemini Omni Flash 1.1, Google's multimodal video model. Upload a clip, describe the change, and hit run.

Gemini Omni Flash 1.1
V2V
Video
Video Edit

70

Gen time: -- secs

Nodes & Models

GeminiOmniFlash11VideoEdit_floyo
VideoToFrames
LoadVideo
CreateVideo
SaveVideo

ABOUT THE WORKFLOW

Edit a Video
Upload a video clip and describe what to change. The model reads the scene, applies the edit, and returns a new version with the rest of the clip intact. Add or replace objects, swap backgrounds, restyle a shot. That's it.

Partner node. This workflow calls an external API, so each run uses credits from your API wallet. No API key needed. Floyo handles the connection.

Model

  • Gemini Omni Flash 1.1 by Google DeepMind. A multimodal video model strong at natural-language editing, object replacement, and scene-aware changes with native audio.


HOW IT WORKS

Step 1. Upload your video
The clip you want to edit. Keep it under 10 seconds. The model reads the full scene before applying any change.
Works great with: short clips · product shots · landscape footage · social content

Step 2. Describe the change
Say what to replace, add, or remove, and what should stay the same. Be specific. "Replace the two biplanes with two large dragons, keeping the same flight path" works better than "add dragons."

Step 3. Hit run and download
The model applies the edit and returns a new video with audio. Preview it in the workflow, then download.
Ready for: Premiere Pro · DaVinci Resolve · After Effects · any editor

First time? Leave every setting as-is. The defaults (720p · auto format) are the right starting point for almost everyone.


RECOMMENDED SETTINGS

Quick-start guide. Find the goal that matches yours and copy the settings.

  • Standard edit (most people) — 720p · auto format. The right starting point for almost everyone.

  • Quick preview or test — 360p. Fastest and cheapest way to check whether the edit landed before committing to a higher resolution.

  • Final output for delivery — 1080p or 4K. Use when the clip goes into a client project or broadcast. Higher resolution costs more per second.

  • Social-first vertical clip — Record or crop your source to 9:16 before uploading. The model preserves the aspect ratio of the input.

  • The edit changes too much of the scene — Be explicit about what to keep. "Replace the car with a truck, keep the road and sky unchanged" narrows the scope.

  • The edit is not landing — Describe the change in concrete terms. Name the object, say where it is, and say what replaces it. Vague instructions like "make it cooler" rarely work.

Prompt: Describe the specific change and what to preserve. "Replace the two biplanes with dragons soaring through the sky, keeping the same flight path and formation above the forest" is clearer than "add dragons to the video." One clear instruction per run gives the best results.


LEARN

📹 Videos

✨ Quick links


USE CASES

🎬 VFX & Post-Production
Replace objects, swap backgrounds, or restyle a shot without reopening your 3D pipeline. Upload a clip, describe the swap, and get a revised version in seconds.

📱 Social Media & Short-Form Content
Edit a clip for a different platform or audience. Change a product, adjust the setting, or swap elements between takes without reshooting.

🛍️ Product & E-commerce Video
Update a product in an existing video without a new shoot. Swap a colourway, place the product in a different environment, or change on-screen details.

🎨 Concept & Pre-Vis
Test how a scene reads with different elements before committing to a full VFX build. Swap objects, change a setting, or prototype a look.

📚 Education & Training
Replace objects or labels in instructional footage to localise content, update visuals, or create scenario variations from one source clip.


WHAT WORKS BEST / WHAT TO AVOID

✅ Works great

  • Short clips under 10 seconds

  • Clear object replacements and swaps

  • Specific, concrete edit instructions

  • Well-lit footage with distinct subjects

⚠️ May produce softer results

  • Clips longer than 10 seconds (must be trimmed first)

  • Vague instructions like "make it better" or "improve it"

  • Edits that require complex physics (fluid, fire, cloth)

  • Multiple conflicting changes in a single prompt


FAQ

What is Gemini Omni Flash 1.1?
Gemini Omni Flash 1.1 is Google DeepMind's multimodal video model, released on August 27, 2026. It processes text, image, audio, and video together. For video editing, you upload a clip and describe the change in plain language. The model applies the edit while preserving the parts of the scene you did not mention. Output is 3 to 10 seconds at 24 fps with audio generated alongside the picture.

How does AI video editing with Gemini Omni Flash work?
You upload a video (10 seconds or shorter) and write a prompt that says what to change and what to keep. The model reads the full clip, identifies the elements you named, and builds a new version with the edit applied. Everything you did not mention stays intact. The output includes generated audio that matches the new scene.

What resolution does Gemini Omni Flash 1.1 support?
The model outputs at 360p, 720p, 1080p, or 4K. 720p is the default and covers most use cases. 1080p and 4K are upscaled outputs, so they cost more per second of generation. Use 360p for fast drafts and 1080p or 4K for final delivery.

Is Gemini Omni Flash 1.1 different from other AI video editors like Runway or Kling?
Most AI video tools generate from scratch or apply broad style transfers. Gemini Omni Flash edits an existing clip with natural-language instructions, preserving the parts of the scene you want to keep. It also generates synchronized audio alongside the video, so the output includes ambience and effects. Kling and Seedance produce stronger cinematic motion quality, while Gemini Omni Flash is faster and better at iterative text-based edits on existing footage.

Does Gemini Omni Flash 1.1 output have a watermark?
Yes. Google applies an invisible SynthID watermark to all generated video. It does not change how the video looks, but it is programmatically detectable for AI provenance. The watermark survives re-encoding and resizing.

Can I use Gemini Omni Flash 1.1 video edits commercially?
Gemini Omni Flash 1.1 is a proprietary Google model. Commercial use is generally allowed, but it is governed by Google's terms rather than an open license. Review Google's current usage policies, note that all outputs carry the SynthID watermark, and confirm you have the rights to any footage you upload.

How to run Gemini Omni Flash 1.1 video editing online?
You can run Gemini Omni Flash 1.1 video editing online through Floyo. No installation, no setup, no API key to wire up. Open the workflow in your browser, upload your video, describe the edit, and hit run. Free to try.


WHY FLOYO?

Floyo is the only platform with team collaboration for ComfyUI in the browser. You run workflows with no install. You share run history, assets, and models across your team. You pay only when you generate. Floyo supports open-source and closed-source models.

A designer runs an edit and likes the result. A teammate opens that exact run from shared history and keeps going. No file handoffs. No version confusion.

For studios and enterprise teams, Floyo adds private workspaces, pooled resources, and a team usage dashboard. Other ComfyUI cloud tools run for one person at a time. Floyo runs for the whole team, with transparent per-generation costs.


Ready to try it?
Upload a video, describe the change, and run it. The settings are already set.

→ Launch Workflow, Free

Questions? Watch the free course or check the FAQ above.

Read more

N