
gemini-omni-videoOmni Flash (Stable) is the full-suite Gemini Omni video generator. One model covers text-to-video, image-to-video (up to 7 reference images), and video-to-video rewriting (video_list), plus optional reusable voices (audio_ids, via the Omni Audio endpoint) and consistent characters (character_ids, via the Omni Character endpoint). Choose 720p / 1080p / 4k, 6 / 8 / 10 seconds, 16:9 or 9:16, with an optional seed. Per-call pricing — 720p and 1080p share $0.35-$0.6, 4k runs $0.7-$1.0; video-input generations are flat ($0.8 for 720p/1080p, $1.2 for 4k). Note the quota: images×1 + videos×2 + character_ids×1 must total ≤ 7.
Text / image / video inputs in one model, plus reusable voices and consistent characters
Per-call pricing from $0.35 — text / image / video input, voice & character in one model
Fuse styles, characters and scenes; or pass 1 video (video_list) to rewrite footage
Build a voice (Omni Audio) or character (Omni Character) once, then reuse via audio_ids / character_ids
Leave empty for text-to-video, or add up to 7 reference images for image-to-video / multi-image fusion.
Upload one clip (≤100MB, ≤10s — the upstream model rejects longer reference videos) to rewrite existing footage. With a video, output duration is set by the model (the Duration buttons are ignored) and pricing switches to the flat video tier ($0.8 for 720p/1080p, $1.2 for 4k).
Generated video will appear here
Provide URLs and click Generate
Omni Flash (Stable) is a Video Generation API provided by Google. Omni Flash (Stable) is the full-suite Gemini Omni video generator. One model covers text-to-video, image-to-video (up to 7 reference images), and video-to-video rewriting (video_list), plus optional reusable voices (audio_ids, via the Omni Audio endpoint) and consistent characters (character_ids, via the Omni Character endpoint). Choose 720p / 1080p / 4k, 6 / 8 / 10 seconds, 16:9 or 9:16, with an optional seed. Per-call pricing — 720p and 1080p share $0.35-$0.6, 4k runs $0.7-$1.0; video-input generations are flat ($0.8 for 720p/1080p, $1.2 for 4k). Note the quota: images×1 + videos×2 + character_ids×1 must total ≤ 7. Through APIMODELS platform, you can access this model via a unified API with transparent pay-as-you-go pricing. Current pricing: 720p / 1080p · 4s: $0.35, 720p / 1080p · 6s: $0.4, 720p / 1080p · 8s: $0.5, 720p / 1080p · 10s: $0.6, 4k · 4s: $0.7, 4k · 6s: $0.8, 4k · 8s: $0.9, 4k · 10s: $1, + video input · 720p / 1080p: $0.8, + video input · 4k: $1.2.
Related model: Omni Flash — the streamlined per-second option
Quickly generate brand promotion videos for ad campaigns and social media marketing.
Create compelling short-form video content for platforms like TikTok, Instagram, and YouTube.
Generate product feature demonstrations and tutorials to improve user conversion.
Produce course explanations, knowledge explainers, and training videos at low cost.
Omni Flash (Stable) is available through APIMODELS at: 720p / 1080p · 4s: $0.35, 720p / 1080p · 6s: $0.4, 720p / 1080p · 8s: $0.5, 720p / 1080p · 10s: $0.6, 4k · 4s: $0.7, 4k · 6s: $0.8, 4k · 8s: $0.9, 4k · 10s: $1, + video input · 720p / 1080p: $0.8, + video input · 4k: $1.2. Billing is pay-as-you-go — you only pay for what you generate.
Sign up at APIMODELS, get your API key, and call our unified API endpoint. We provide detailed API documentation with code examples in cURL, Python, and Node.js.
APIMODELS offers the same Omni Flash (Stable) model through our aggregation platform. We provide a unified API interface so you do not need separate accounts for each provider - one API key to access all models.
Omni Flash (Stable) (gemini-omni-video) is billed per call, costs less (from $0.35 at 720p), and supports the full suite — video-to-video (video_list), custom voices (audio_ids), and character consistency (character_ids). Omni Flash (gemini-omni-flash) is billed per second and renders fast but is more streamlined. Most use cases are better on Stable; use Omni Flash when you just need quick text-to-video / image-to-video.
On APIMODELS, Omni Flash (Stable) runs alongside 60+ models on one API key and one balance, so choosing is about fit, not lock-in. It supports T2V / I2V / V2V, Up to 7 Ref Images, Voice + Character, 720p / 1080p / 4k, 6 / 8 / 10s, and you can weigh it on price and capability against other Video Generation models, then switch by changing a single model-name string — no new account or integration. Browse every Video Generation option with live pricing at apimodels.app/models.
Omni Flash (Stable) supports: T2V / I2V / V2V, Up to 7 Ref Images, Voice + Character, 720p / 1080p / 4k, 6 / 8 / 10s. See the APIMODELS docs for full parameters and call examples.
Yes. APIMODELS exposes Omni Flash (Stable) through a single unified API and one key — no separate provider accounts, and no need to handle each provider's regional network access yourself.
We support Stripe (Visa, Mastercard, and other international cards) and Alipay. Credits are available instantly after payment.
Prompts shared by their authors — copy and adapt them. Each one credits its author and links back to the original post.
Condor Heroes characters teach English word dream
神雕侠侣主角趣味讲单词 dream 教程
by @nicekate8888
Pizza night UGC Domino’s vlog
VIDEO PROMPT — "Pizza Night Vlog" (UGC iPhone Style) Duration: 15 seconds | Aspect Ratio: 16:9 | Style: Authentic UGC / iPhone selfie-vlog, handheld, natural light, slight motion blur, TikTok/Reels energy — NOT cinematic, NOT overly polished. Feels like a real creator filmed this on their phone. Product Reference: Use the uploaded Domino's Pepperoni Pizza image as the only product reference. Keep crust thickness, cheese texture, bake color, pepperoni placement, and proportions identical in every cut — no redesigning the pizza. Camera: iPhone 15 Pro front + back camera switching, handheld, natural wobble, autofocus hunting slightly (realistic), vertical-style framing cropped to 16:9, occasional finger near lens edge, natural room lighting + phone flash reflections on the pizza box. Character Description Name (for reference): Mia Awoman in her mid-20s, naturally attractive and beautiful with an approachable, girl-next-door charm — not overly done up. Wavy sandy-blonde hair pulled back loosely, light natural makeup, wearing a cozy oversized cream sweater. Warm, genuine smile, expressive eyes, casual energetic personality like a real lifestyle vlogger. Sits in a softly lit modern kitchen/living room. Shot Breakdown SHOT 1 (0–2s) — The Grab Selfie-angle, she's mid-laugh holding up the Domino's box to camera. Quick jump cut. Dialogue: "Okay so it's officially pizza night—" SHOT 2 (2–4s) — The Open Cut to overhead handheld shot, box flips open, steam rising off the pizza, slight camera shake as she leans in. SHOT 3 (4–6s) — The Zoom Quick zoom-punch into the pizza, phone camera autofocus adjusts naturally, cheese and pepperoni in focus, ambient kitchen sounds. SHOT 4 (6–8s) — The Pull Cut to her hands lifting a slice, natural cheese pull, filmed from a slightly low candid angle like a friend filming across the table. SHOT 5 (8–10s) — The Reaction Cut back to selfie-cam, she takes a bite, eyes widen, quick genuine reaction. Dialogue: "Oh my god, that's so good." SHOT 6 (10–12s) — The Candid Cutaway Jump cut to a close, slightly shaky shot of the pizza box on the counter, her hand grabbing another slice off-frame, casual b-roll energy. SHOT 7 (12–14s) — The Wrap-Up Back to selfie angle, she grins at camera, holding slice up like a toast. Dialogue: "Dominos, y'all know what to do." SHOT 8 (14–15s) — End Tag Quick freeze/cut to the box logo close-up, natural handheld wobble, soft text overlay in casual font: "pizza night = solved 🍕" — cut to black. Look & Feel Warm indoor lighting, slightly grainy natural phone sensor look, imperfect framing, real reactions, minimal dialogue (3 short lines total), authentic pacing with hard jump cuts instead of smooth transitions. Negative Prompt cinematic grade, overly smooth camera moves, studio lighting, professional voiceover, staged acting, CGI look, plastic cheese, distorted pepperoni, extra fingers, warped hands, text glitches, logo distortion, overly polished commercial feel.
by @ShamiWeb3
The World's Unluckiest Superhero
A documentary about a superhero who has extremely bad luck and ends up saving people by accident through the destruction caused by his own misfortune. Dialogue in English. Scene direction with unique composition. Every cut, every camera angle, and every movement is of exceptionally high quality; the composition is guided by an experienced film director. The comedy is genuinely interesting, and throughout the 15 seconds everything unfolds in a perfectly crafted way, with a comedic payoff that can make anyone laugh.
by @NACHOS2D_
We curate copy-ready prompt libraries — every entry shows its full text and a sample result, ready to adapt.