Access Gemini Omni Flash - also searched as Gemini Omni or Omni Flash - through EvoLink's unified video API. Test text-to-video, image-to-video, reference-to-video, and conversational video editing before you integrate.
Google·Video Generation·Available
From $0.015 / 1K video output tokens$0.018 official price-15%
Text-to-VideoImage-to-VideoReference-to-VideoConversational Video EditNative Audio
Production routeThis rate reflects platform-side availability — only confirmed server errors (HTTP 500 / empty response) count as failures. User-side issues (content moderation, invalid params, cancellation) plus rate limits, timeouts and auth errors are excluded. Before real traffic arrives, empty buckets may display as available.Live
Live
Model Highlights
3-10s or Auto duration · 720p output · native audio included
Pick a mode, write a prompt, and see the estimated cost before you spend anything.
Create task
Fixed to gemini-omni-flash-text-to-video
Duration
Auto lets the provider pick the length; the estimate above assumes 10 seconds.
Aspect ratio
Estimated: ~$0.87059.1102 cr
Ready
Ready
History
Results are kept for 24 hours — download anything you want to keep.
Your generation history will appear here
Know your Gemini Omni Flash testing cost before you add credits.
Gemini Omni Flash is billed by tokens, not by the second. The final charge always uses the tokens the model actually returned.
First Test Cost
Text-to-Video · 3s · 720p
-15%Minimum charge
Credits18.0961
One 3s 720p sample
Approx. Cost$0.267
About 35 tests with $10 in credits.
In Playground, pick Text to Video, set Duration to 3s, and send a short prompt with no input media - that is the cheapest valid request on this page. Output is always 720p, so cost scales with duration, not resolution.
Testing Budget Guide
Pick an amount based on how many tests you expect.
Estimates only. Gemini Omni Flash bills by token, so the final charge always uses the usage the model actually returned. Auto duration is reserved as 10s and settled against the seconds actually produced.
Model Pricing
Model PricingToken billing
Gemini Omni Flash bills by token across three categories: video output, input (prompt + media), and other output (thinking tokens). Final charges always reflect the usage actually returned by the model.
Item
Rule
Rate
Billing
Video Outputper 1K tokens
Generated video frames and audio
$0.015 / 1K tokens-15%
1.0115 cr / 1K tokens$0.018 Official
Output
Inputper 1K tokens
Prompt text, reference images, and input video
$0.0013 / 1K tokens-15%
0.0867 cr / 1K tokens$0.0015 Official
Input
Other Outputper 1K tokens
Thinking tokens and internal reasoning
$0.0077 / 1K tokens-15%
0.5202 cr / 1K tokens$0.0090 Official
Output
Video Output
Generated video frames and audio · Output
$0.015 / 1K tokens-15%
1.0115 cr / 1K tokens$0.018 Official
Input
Prompt text, reference images, and input video · Input
$0.0013 / 1K tokens-15%
0.0867 cr / 1K tokens$0.0015 Official
Other Output
Thinking tokens and internal reasoning · Output
$0.0077 / 1K tokens-15%
0.5202 cr / 1K tokens$0.0090 Official
Rates shown are per 1,000 tokens. The actual charge for each request is calculated from the usage values returned in the API response.
Price your own Gemini Omni Flash request
Gemini Omni Flash API on EvoLink
Use Gemini Omni Flash on EvoLink for text-to-video, image-to-video, reference-to-video, and video editing through one unified video API. Public discussion often frames Gemini Omni as a video counterpart to Nano Banana because it brings multimodal video creation and conversational editing into short-form workflows.
On EvoLink, the practical value is API access: EvoLink model IDs, async task workflow, callback support, token-based usage visibility, and the same API key used for Veo, Seedance, Kling, and other video models.
Input
Text, images, or one input video
Output
720p MP4 with native audio
Duration
3-10s, or Auto
Resolution
720p (fixed)
Reference assets
Up to 6 reference images
What can you build with Gemini Omni API?
Text to Video
gemini-omni-flash-text-to-video
Generate a 720p video from a prompt with 3-10 second or Auto duration.
Image to Video
gemini-omni-flash-image-to-video
Animate one input image with a prompt and 3-10 second or Auto duration.
Reference to Video
gemini-omni-flash-reference-to-video
Use 1-6 reference images to guide a 3-10 second or Auto-duration video.
Video Edit
gemini-omni-flash-video-edit
Edit one MP4 input video with a natural-language instruction.
Chat-Based Video Editing
Generate a clip with Gemini Omni, then refine it in conversation — "make the lighting warmer", "replace the red car". The workflow is designed for iterative edits while preserving the surrounding scene, subject identity, and motion as much as the selected route supports.
· Multi-turn refinement in one chat thread
· Scene continuity during edits
· Less regenerate-from-scratch work
Object Replacement and Scene Rewrite
Swap an object in frame, remove an unwanted element, or rewrite a scene while preserving identity and motion. Useful for ad creative iteration and product variant rendering without external editing tools.
· Swap or remove objects in frame
· Preserve identity and motion
· Ad creative and product variant iteration
Reference Image Workflow
Pass a reference image and Gemini Omni anchors character identity, lighting, and color across the generated video. Combine with chat-based editing to refine specific shots without losing visual consistency.
· Anchor character identity from reference image
· Consistent lighting and color across clips
· Combine with chat editing for refinement
Audio-Capable Video Generation
Gemini Omni Flash routes can return short video outputs with audio where supported by the selected mode, reducing the need to stitch a separate TTS or sound-design pipeline into first-pass generation.
· Video output with audio where supported
· Useful for first-pass short-form concepts
· Less separate audio pipeline work
Two ways to use Gemini Omni Flash: EvoLink API or Agent
Use the EvoLink API for product integration and batch jobs, or call it from Codex, Claude, or Gemini for fast creative and development workflows. Both paths share the same EvoLink API key, balance, model routes, and task history.
Option 1
Integrate with the EvoLink API
Best for: product backends, batch jobs, automated workflows
Call EvoLink’s unified video API from your server and control the model ID, parameters, task queue, callbacks, and result storage.
1Validate output and cost with a real brief in Playground
2Create an EvoLink API key in the console
3Pick the model ID that matches your task from the route cards above
4Submit the task and retrieve the result by polling or HTTPS callback
Best for: creative and development tasks in Codex, Claude, and Gemini
Give the Agent your output goal, assets, and acceptance criteria. It can choose the route, assemble the request, track the task, and return the result without requiring you to hand-code every step.
1Set EVOLINK_API_KEY in your local environment; never put it in code or a prompt
2Describe the output goal, references, and the key parameters
3Ask the Agent to call a Gemini Omni Flash route and save the task ID
4Let the Agent poll the final state, download the result, and report failure reasons
What teams build with the Gemini Omni Flash API
Four workflows that map onto the four routes. Each one is a different input shape, not a different quality tier - pick by what you have on hand.
Short-form ad variants from a brief
Generate 3-10 second cuts from the prompt alone, then iterate on wording rather than re-shooting. Native audio means the first pass is already reviewable without a separate TTS or sound pass.
Animating a product still
Image-to-video takes exactly one input image and adds controlled motion and lighting while holding the subject. Useful when you have approved product photography but no video budget.
Identity-consistent scenes from references
Reference-to-video accepts 1-6 images to anchor a character, palette, or style across multiple generated shots. Order carries no meaning here - the set is read as a whole.
Conversational edits on an existing clip
Video edit rewrites objects, style, or lighting in one MP4 up to 10 seconds. Length and aspect ratio follow the input, so the output stays drop-in compatible with the original cut.
Gemini Omni Flash API code example and error handling
This example shows the shortest runnable flow: create a task, save the returned task ID, then poll or wait for an HTTPS callback. Open the API tab for the complete parameter and response reference.
A Gemini Omni Flash task that reaches a final failed state is not charged. Keep the task ID, inspect the final state, and identify whether the cause is validation, review, or execution before retrying.
Callback not received
callback_url must be a public HTTPS URL, not localhost or a private network address. Always keep the task ID as a polling fallback.
Gemini Omni vs Veo 3.1 vs Seedance 2.0 — Side-by-side comparison
Three models commonly shortlisted for production video workflows in 2026. All three accessible through one EvoLink API key.
Feature
Gemini Omni
Veo 3.1
Seedance 2.0
EvoLink price
Token-based
From $0.04/s
From $0.045/s
Quality
720p
720p / 1080p, 4K upscaling where available
480p / 720p / 1080p
Native audio
Yes
Yes
Yes
Reference control
Text + image + chat edit
Text + image
Text + image + video + audio
Video length
3-10s / Auto
Short clips with Extend for longer scenes where supported
Three steps to your first Gemini Omni video task. Same integration pattern as Veo 3.1, Seedance 2.0, and Kling 3.0.
1
Step 1 — Get Your API Key
Sign up on EvoLink.ai and generate your API key from the dashboard. No Google Cloud project required.
2
Step 2 — Submit Generation Task
POST to /v1/videos/generations with one of the Gemini Omni Flash model names and your prompt. Add duration for 3-10 second or Auto generation modes, image_urls for image-to-video or reference-to-video, video_urls for video edit, and callback_url for completion notification. The API processes asynchronously and returns a task id.
3
Step 3 — Retrieve Video Result
Use the task ID to poll the status endpoint, or wait for the callback_url webhook. When status reaches completed, you receive a download URL for the generated MP4. Links are valid for 24 hours.
Gemini Omni API Capabilities
Technical specifications for production video workflows.
Editing
Chat-Based Video Editing
Multi-turn refinement in a conversational workflow, with scene continuity depending on the selected route and input quality.
Output
720p, 3-10s / Auto Clips
720p output with configurable 3-10 second or Auto clips for generation modes. Auto is estimated as 10 seconds. Video edit accepts one MP4 input up to 10 seconds.
Modes
Text-to-Video and Image-to-Video
T2V from prompts and I2V with reference image input. Chat editing applies to outputs of either mode.
Audio
Audio-Capable Video Output
Short video outputs can include audio where supported by the selected Gemini Omni Flash route.
Consistency
Long-Context Character Consistency
Designed for stronger continuity across multi-input and edit-heavy workflows; validate consistency on your own production prompts.
Workflow
Async API with Task ID and Callback
Submit a task, receive an ID, poll status or configure a callback_url. Same lifecycle as other EvoLink video models.
Cost Example — Gemini Omni pricing estimates
100 × 3-10s/Auto clips for social media batch
Use current Pricing tab rates
1,000 × 3-10s/Auto clips/month at production scale
Use current Pricing tab rates
1 generation + 3 edits multi-turn workflow
Use current Pricing tab rates
Use the Pricing tab above for current token-based rates. Select the workflow by changing the model parameter.
Why developers choose Gemini Omni on EvoLink
Gemini Omni is most interesting for workflow rather than raw fidelity alone: multimodal inputs, conversational editing, and a practical EvoLink route for testing it beside Veo, Seedance, and Kling with one API key.
Chat-Native Editing Workflow
Gemini Omni is positioned around conversational video editing, while Veo 3.1 and Seedance 2.0 are usually evaluated first as generation routes. For multi-turn refinement, this is the workflow difference to test.
Long-Context Character Consistency
Gemini Omni is reported to benefit from Gemini context and world knowledge for continuity across multi-input and edit-heavy workflows. Treat this as a behavior to evaluate in your own storyboard or short-video pipeline.
No Google Cloud Project — Same Async Pattern as Veo and Seedance
No GCP setup, no Vertex billing, no separate region approval. If you already run video generation through EvoLink, adding Gemini Omni is a one-parameter change — same request shape, same task lifecycle as Veo 3.1, Seedance 2.0, and Kling.
What is Gemini Omni and how is it different from Veo 3.1?
Gemini Omni is Google's multimodal video model family announced at Google I/O 2026, with Omni Flash discussed as a short-form video route for text, image, video, and audio inputs. Compared with Veo 3.1, Gemini Omni is more interesting for conversational editing and multi-input workflows, while Veo remains a strong cinematic generation baseline.
How much does Gemini Omni API cost on EvoLink?
Billing follows the usage tokens returned by the API, with separate token meters for input, video output, and other output. Check the Pricing table above for current rates.
Do I need a Google Cloud project to use Gemini Omni through EvoLink?
No. EvoLink provides access via one API key. No Google Cloud project, no Vertex billing, no separate region approval. Same authentication as Veo 3.1 and Seedance 2.0 on EvoLink.
Which Gemini Omni Flash modes are supported?
Four modes are available: gemini-omni-flash-text-to-video, gemini-omni-flash-image-to-video, gemini-omni-flash-reference-to-video, and gemini-omni-flash-video-edit. All share the same async video API endpoint.
Does Gemini Omni support callback_url and webhooks?
Yes. Pass a callback_url (HTTPS) when submitting the task and EvoLink can POST task updates to your endpoint when the task reaches a terminal state. Polling the task status endpoint also works if you do not provide a callback URL.
How do I handle failed Gemini Omni tasks?
Failed tasks return a failed status with an error reason. For application-level retry, inspect the error, keep the original parameters for debugging, and resubmit only when the input or transient failure mode is clear.
Can I edit a previously generated video with Gemini Omni?
Yes — this is one of Gemini Omni's main workflow differences. Use a natural-language edit instruction and validate how well the selected route preserves the surrounding scene, subject identity, and motion across iterations.
What's the maximum clip length for Gemini Omni?
Generation modes support configurable 3-10 second or Auto clips. Auto is estimated as 10 seconds for reservation. Video edit accepts one MP4 input up to 10 seconds. For longer narratives, chain multiple clips using long-context character consistency.
Does Gemini Omni support image-to-video?
Yes. Pass a reference image URL and Gemini Omni uses it as an identity anchor for the generated video.
How does Gemini Omni compare to Seedance 2.0 and Veo 3.1?
Seedance 2.0 has strong benchmark and multimodal reference signals, while Veo 3.1 remains a strong cinematic generation baseline with advanced Flow and extension workflows. Gemini Omni is different because developers are evaluating it for conversational editing, multi-input generation, and short-form iteration.
Can I access all Google video models through one EvoLink API key?
Yes. EvoLink exposes Gemini Omni, Veo 3.1, Nano Banana 2, and the rest of the Gemini family through a single API key. Switch by changing the model parameter.
EVOLINK · PRICE EST.gemini-omni-flash
Auto estimated as 10s · real-time
Figures are pre-bill estimates. Actual charges follow the upstream usage tokens returned by the model.