GoogleGoogle

Gemini Omni Flash
One Model, 360p to 4K, Any Direction

Gemini Omni Flash v1.1 applies Gemini's multimodal reasoning to video - generate from text, animate an image with an optional end frame, blend up to 10 images and 3 reference videos, or just describe the edit you want. Output from a fast 360p draft up to a clean 4K pass.

Gemini Omni Flash v1.1 is available on faktry: text-to-video, image-to-video, reference-to-video, and instruction-based editing, all priced per second by resolution from 360p up to 4K.

What Is Gemini Omni Flash?

Gemini Omni Flash is Google DeepMind's video generation and editing model, updated to v1.1 on August 27, 2026 with expanded resolution options and reference controls. It generates video from a text prompt, a single image, or up to 10 reference images and 3 reference video clips, and edits an existing clip from plain-language instructions - all in one model, with audio controlled directly through the prompt. Output ranges from a fast, low-cost 360p draft up to a clean 4K final pass, and image-to-video now supports an optional end frame for precise first-to-last transitions.

Gemini Omni Flash at a glance

Developer
Google DeepMind
Released
v1.1, August 27, 2026
Type
Text-to-video, image-to-video, reference-to-video, and video editing model
Max video length
3-10 seconds per generation (editing bills the real uploaded clip length)
Resolution
360p, 720p, 1080p, 4K
References
Up to 10 images and 3 reference video clips (each up to 3s)
Audio
Controlled directly through the prompt
On faktry
Available now

Last updated: · Model details published by Google

Gemini Omni Flash, Live on faktry

Four ways to generate and edit video with Gemini Omni Flash, all priced per second by resolution.

Text to Video

Generate a video directly from a text prompt, with audio controlled through the prompt itself.

2-20 credits / second, by resolution
3-10s · 360p, 720p, 1080p, or 4K · 16:9 or 9:16

Image to Video

Animate a single reference image into video, with an optional end frame for precise first-to-last transitions.

2-20 credits / second, by resolution
3-10s · 360p, 720p, 1080p, or 4K · optional end frame

Reference to Video

Combine up to 10 reference images and 3 reference video clips, bound inline via '<IMAGE_REF_n>' and '<VIDEO_REF_n>' tags.

2-20 credits / second, by resolution
3-10s · 360p, 720p, 1080p, or 4K · up to 10 images + 3 videos

Video Editing

Upload an existing clip and describe the change in plain language - no timeline, no manual masking.

2-20 credits / second, by resolution
Billed on real video length, up to 60s · 360p to 4K

Resolution, duration, and reference inputs are all configurable per generation. Lower resolutions cost less per second.

Why Choose Gemini Omni Flash?

One model for four video operations, with resolution and reference controls built in.

3-10 Second Clips

Generate text-to-video, image-to-video, and reference-to-video clips from 3 to 10 seconds per pass.

360p to 4K Output

Render a fast, low-cost 360p draft or go straight to a clean 4K final pass, priced per second at each tier.

Multi-Image & Video References

Combine up to 10 reference images and 3 reference video clips, bound inline in your prompt with '<IMAGE_REF_n>' and '<VIDEO_REF_n>' tags.

Conversational Video Editing

Upload a clip and describe the change in plain language - no timeline, no manual masking.

In-Prompt Audio Control

Describe the music, dialogue, or silence you want directly in the prompt - Omni Flash generates it alongside the video.

Gemini's Real-World Reasoning

Carries Gemini's knowledge of history, biology, and narrative logic into every frame for coherent scenes.

Capabilities

A closer look at what Gemini Omni Flash can generate, control, and edit.

Video Generation & Editing

Text-to-video, image-to-video with an optional end frame, reference-to-video from images and video clips, and instruction-based editing of existing clips
3 to 10 second clips per generation; editing bills the real length of the uploaded clip, up to 60 seconds
360p, 720p, 1080p, or 4K output, priced per second at each tier
Up to 10 reference images and 3 reference video clips (each up to 3 seconds), bound inline via '<IMAGE_REF_n>' and '<VIDEO_REF_n>' tags
Describe the change in plain language and the model applies it directly - no timeline, no manual masking
Audio is controlled directly through the prompt - describe the music, dialogue, or silence you want

Built for Fast-Moving Teams

Omni Flash trades render complexity for speed and a natural-language workflow.

E-Commerce & Product

Turn a single product photo into a moving showcase, or pin a start and end frame to control exactly how a shot transitions.

Social & Short-Form Content

Go from prompt to a vertical, sound-aware clip in one pass - no separate audio generation step.

Rapid Video Editing

Skip the timeline. Describe a style change or fix and let conversational editing apply it directly.

Storytelling & Concepting

Use Gemini's real-world knowledge to keep narrative logic, settings, and details coherent across a scene.

How It Works

Creating and editing video with Gemini Omni Flash is simple and fast.

1

Describe or Upload

Write a text prompt, add up to 10 reference images and 3 reference videos, or upload an existing clip to edit.

2

Configure Your Settings

Choose resolution from 360p to 4K, aspect ratio, and duration for your generation.

3

Generate & Download

Get your video in moments, with SynthID watermarking and commercial usage rights included.

Can I try Gemini Omni Flash for free?

Start generating and editing video today with free credits.

Free credits to try us

100 credits included - no card required
Access to Gemini Omni Flash v1.1 generation and editing
360p to 4K output, 16:9 and 9:16 aspect ratios
Commercial usage rights included
Get Started Now

Frequently Asked Questions

What is Gemini Omni Flash?

Gemini Omni Flash is Google DeepMind's video generation and editing model, updated to v1.1 on August 27, 2026. It combines Gemini's multimodal reasoning with video generation and editing, and is available on faktry for text-to-video, image-to-video, reference-to-video, and instruction-based video editing.

Can I try Gemini Omni Flash for free?

Yes. Sign up at faktry.ai and you get 100 free credits instantly, no credit card required. That is enough to generate your first Gemini Omni Flash videos straight away. After that there is no subscription: you pay only for what you actually generate.

What resolutions does Gemini Omni Flash v1.1 support?

360p, 720p, 1080p, and 4K, each priced per second - render a fast, low-cost 360p draft or go straight to a clean 4K final pass.

Can I use a reference video, not just images?

Yes. Reference-to-video accepts up to 3 reference video clips (each up to 3 seconds long) alongside up to 10 reference images, bound inline in your prompt with tags like '<IMAGE_REF_0>' and '<VIDEO_REF_0>'.

Does image-to-video support a start and end frame?

Yes. Image-to-video accepts an optional end frame image - provide a first and last frame and the model interpolates the motion between them.

What's the maximum video duration?

Text-to-video, image-to-video, and reference-to-video generations support 3-10 second clips. Video editing has no duration input of its own - it's billed on the real length of the uploaded clip, up to 60 seconds.

Is generated video watermarked, and can I use it commercially?

Video generated with Gemini Omni Flash carries Google's SynthID watermark and passes through safety filters. All video generated on faktry, including with Omni Flash, can be used commercially.