Omni Flash
Gemini Omni is Google DeepMind's video model that turns any input — video, image, text, or audio — into one cohesive video. Access it through Velokey's unified endpoint with below-list pricing.
Use cases
Marketing & social video · Educational & explainer video
Input
Text, image, or video
Output
Video
Billing
Per second
Sign up to get $0.5 in free credits — about 20 free images
Generate a video
Enter a prompt and click Generate
Omni Flash video examples
👆 Click on any example above to view full details, prompt, and parameters
Pricing details
Transparent usage-based pricing with no hidden fees.
Detailed Pricing
Usage-based pricing breakdown
| Spec | Price | Official | Savings |
|---|---|---|---|
| Input/1M tokens | 1.2 Credits≈$1.2 | 1.5 Credits | -20% |
| Output/1M tokens | 14 Credits≈$14 | 17.5 Credits | -20% |
Billing Rules
- Output price varies by model, resolution, quality, duration, or token usage.
- Uploaded reference files may be charged separately when supported.
- Add-on capabilities such as web search or image search may be billed per request.
- Token-based models are billed by input and output tokens.
- If the primary route is unavailable, Velokey may automatically switch to a stable fallback route when possible.
- Usage is pay-as-you-go with transparent credit deduction.
* Final cost depends on the generated result.
Related models
Explore other models in the same family.
GoogleSave 20%Veo 3.1
Veo 3.1 is Google's state-of-the-art video model with richer audio, true-to-life textures, stronger prompt adherence, plus Ingredients/Frames-to-Video and Extend.
OpenAISave 20%Sora 2
OpenAI's Sora 2 flagship video model with synchronized audio, improved physical realism, and stronger controllability and consistency.
ByteDanceSave 20%Doubao Seedance 2.0
Seedance 2.0 is designed for smoother camera movement, expressive character motion, and rich stylized sequences.
AlibabaSave 20%Wan 2.7
Wan 2.7 offers reliable multi-scene video generation with balanced quality, speed, and motion stability.
Affordable Gemini Omni API
Access Google DeepMind's Gemini Omni on Velokey — turn any input into one cohesive video, edit it conversationally, and get physics-grounded realism, with a unified endpoint and below-list pricing for production apps.
Meet Gemini Omni and its API access
Gemini Omni is Google DeepMind's video model built to create anything from any input — starting with video. It accepts video, image, text, and audio inputs and generates one cohesive video output, turning any reference into a single unified result. You can edit any video through natural, step-by-step conversation with consistency across turns, and it pairs an intuitive understanding of physics with Gemini's world knowledge for grounded storytelling. Velokey provides an affordable, easy-to-integrate Gemini Omni API: a unified endpoint, a clear request flow, and below-list pricing for production creative pipelines, marketing tools, and automated video workflows.
Any input to a cohesive video
Accepts video, image, text, and audio inputs and generates a single cohesive video output. Turn any reference — image, text, video, or audio — into one unified result, starting from any modality.
View Documentation →
Conversational step-by-step editing
Edit any video through natural, step-by-step conversation. Each edit builds on the previous one, maintaining consistency across turns so you refine shots without re-rendering from scratch.
View Documentation →
Physics-grounded realism + world knowledge
An intuitive understanding of physics — gravity, kinetic energy, and fluid dynamics — combined with Gemini's knowledge of history, science, and cultural context for grounded, believable storytelling.
View Documentation →
Motion & style transfer and character swapping
Text-to-video synchronization, character and object swapping with motion consistency, sketch-to-video translation, and motion and style transfer from any reference.
View Documentation →
How to deploy the Gemini Omni API on Velokey
Get started with just a few steps.
Sign up and get an API Key
Sign in to the Velokey console to generate an API Key that authenticates every video generation and editing request sent to Gemini Omni.
Configure request parameters
Set model to gemini-omni and provide your inputs — video, image, text, or audio references — plus a prompt describing the video to generate or the edit to apply.
Integrate and start calling
Send requests to Velokey's unified endpoint to generate cohesive video and continue the conversation to edit it step by step, with usage managed in one console.
Real-world use cases for Gemini Omni
From marketing to education and production, covering many creative video scenarios.
Marketing & social video
Turn a brief, an image, or a rough clip into a polished cohesive video, then refine it conversationally for every channel and format.
Conversational editing workflows
Refine shots step by step through natural conversation, with each edit building on the last and consistency held across turns.
Educational & explainer video
Ground storytelling in Gemini's world knowledge of history, science, and culture for accurate explainer and learning content.
Reference-driven production
Turn any reference — image, text, video, or audio — into one cohesive output, unifying scattered assets into a single video.
Motion & style transfer
Transfer motion and style from a reference onto new footage, with character and object swapping that keeps motion consistent.
Physics-accurate simulation & product viz
Use grounded physics — gravity, kinetic energy, fluid dynamics — for believable product visualization and simulation-style shots.
Why choose Velokey for Gemini Omni
Discount on list price
Call Gemini Omni at a lower price on Velokey — pay-as-you-go per second of video with no forced subscription, reducing upfront cost.
Unified API interface
One API Key accesses Gemini Omni and other models — no juggling multiple accounts and keys, simplifying integration.
Complete docs and integration guidance
Endpoint references, parameter details, and input examples help you wire up any-input generation and conversational editing quickly.
Stable and highly available
Automatic failover and stable fallback routes maintain availability when the primary route is down, keeping production reliable.
API Reference
Complete API call examples and parameter descriptions
Endpoint
https://api.velokey.ai/v1/chat/completionsAuthentication
Bearer YOUR_API_KEYRequest Example
curl https://api.velokey.ai/v1/chat/completions \
-X POST \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "omni-flash",
"prompt": "A cinematic shot...",
"duration": 5,
"resolution": "720p"
}'Response Example
{
"id": "req_abc123",
"model": "omni-flash",
"created": 1234567890,
"data": { ... }
}Gemini Omni API FAQ
What inputs does Gemini Omni accept?
It accepts video, image, text, and audio inputs and generates one cohesive video output. You can turn any reference — in any of these modalities — into a single unified result.
How does conversational video editing work?
You edit any video through natural, step-by-step conversation. Each edit builds on the previous one, so the model maintains consistency across turns without restarting from scratch.
How realistic is the motion?
Gemini Omni has an intuitive understanding of physics — gravity, kinetic energy, and fluid dynamics — combined with Gemini's world knowledge of history, science, and culture for grounded, believable storytelling.
What editing and transfer features are supported?
Text-to-video synchronization, character and object swapping with motion consistency, sketch-to-video translation, and motion and style transfer from references.
How does it perform on benchmarks?
In human evaluation it leads on video editing (overall preference and instruction following), ranks top on text-to-video (MovieGenBench), and leads on image-to-video (VBench I2V).
How is it billed on Velokey?
Gemini Omni is billed per second of generated video through Velokey's unified endpoint, with a discount on list price and usage managed in one console.
Start building with Omni Flash today
Use one API key to access low-cost, stable AI models through Velokey.