Transform this raw talking-head video into a premium, viral-quality social media reel with high-retention editing, cinematic motion graphics, and modern creator-style visuals.
Do not change, replace, regenerate, or alter anything in the original footage itself. Keep the original video, frames, speaker, and audio exactly as they are — only add elements on top: kinetic text, motion graphics, B-roll overlays, color grading, and sound design. Everything you add must be a layer over the untouched original footage, never a modification of it.
All added visuals must match what the speaker is actually saying. Every background element, B-roll clip, graphic, icon, visual metaphor, and on-screen keyword must be directly based on the words being spoken at that exact moment. Do not insert generic, random, or unrelated visuals. Listen to the spoken content and only show supporting visuals that accurately represent that specific point — the background must always reflect the topic being discussed, never a different or off-topic subject.
Preserve the original speaker, original audio, and original clothing color completely unchanged — do not alter, recolor, or modify the subject's appearance, clothing, skin tone, hand movements, or gestures in any way. Mirror all natural gestures and body language exactly as they appear in the raw footage. Maintain precise lip-sync with the original spoken audio throughout.
Keep zooming and punch-in effects to a minimum — use them only occasionally and subtly to emphasize a strong point, not as a constant effect. When reframing or cropping, never cut off the speaker or remove any part of the spoken content; the full talking point and the subject must always remain visible and intact. Framing changes are allowed only as gentle, added emphasis, never in a way that loses footage or speech.
On-screen typography must be in English only. Do not display full sentences or full captions — show only the main keywords and key phrases that carry the core message, and these keywords must be taken directly from what the speaker is actually saying. Caption design should be a dominant visual element, not plain subtitles. Use bold cinematic typography where these key words become oversized design elements on screen. Build layered text compositions with foreground and background typography. Let key phrases appear behind the speaker, partially masked by the subject, creating a premium 3D depth effect.
Typography Style: Large editorial fonts, luxury cinematic typography, mixed font weights, layered text compositions, kinetic typography, motion-tracked text, depth and parallax effects, premium 3D text treatments.
Color Style: Build the entire typography and graphic palette around a rich, cinematic red-based scheme. Use a deep crimson-to-scarlet primary accent, blended into warm complementary tones — burnt orange, golden amber highlights, and deep maroon/black shadows for contrast and luxury feel. Important keywords should carry premium red gradient treatments with a soft warm glow and subtle bloom, creating a bold, high-energy, attention-grabbing look. Ensure the red palette feels cinematic and intentional — never flat or harsh — by layering gradients, depth shadows, and gentle lighting so the typography integrates naturally with the footage rather than overpowering the subject.
Whenever the speaker makes a strong point, create huge cinematic text moments — let key words dominate the screen, use layered typography behind the speaker, add scale animations, depth, shadows, lighting effects, and subtle motion. The words shown must match the exact point being made at that moment.
Add seamless transitions, speed ramps, motion blur transitions, and smooth camera movements where appropriate. Include relevant B-roll, graphics, UI animations, visual metaphors, icons, callouts, and motion graphic overlays to visually support the spoken content — all directly tied to what is being said. Do not add any subscribe button, follow button, or social-action UI elements.
Enhance color grading, contrast, exposure, and subject separation for a cinematic finish — without altering the subject's natural appearance or clothing color.
Add professional sound design including whooshes, impacts, swipes, clicks, risers, transition sounds, and subtle cinematic audio layering — all synced precisely to the original spoken audio.
Maintain fast, engaging pacing with meaningful visual changes every few seconds to maximize viewer retention.
Match the typography aesthetic to the uploaded reference style screenshots. The final output should feel like a premium reel produced by a top-tier content agency, where typography and captions are central visual storytelling elements.
RUNS ON
Gemni Omni Flash
How to use
Steps. Run it right.
Don't just paste and pray. Read how to get the best results from this prompt.
How to Edit Videos with Google Flow (Omni Flash)
Google Flow’s Gemini Omni Flash model lets you transform any video clip just by describing what you want — no timeline, no editing software needed. Here’s the full process:
Open Google Flow — head to labs.google/fx/tools/flow. Omni Flash requires a Google AI subscription (Plus, Pro, or Ultra), since it runs on AI credits.
Create a new project inside Flow.
Prepare your video. Omni Flash works in 10-second clips. If your video is longer, split it into 10-second pieces first using any editor.
Add your clip to the project window.
Select Omni Flash as your model.
Choose your uploaded video as the reference.
Paste your prompt and hit Enter. Flow will generate the edited result — then watch the magic happen.
Repeat for each clip you want to transform.
Combine all clips into one final video using any video editing app — or stitch them together right inside Flow.
Use cases
When to reach for this prompt.
◆Video Editing
◆Social Media
Pro tips
Things I'd add.
◇
Tip: Use Claude or ChatGPT to write and refine your prompt before pasting it into Flow. Tweak the wording until it matches the exact look you want.
◇
Note: Every Omni-generated video carries an invisible SynthID watermark identifying it as AI-made.