Write what you want to see like you are telling a friend. Example: "rain-soaked Tokyo alley, neon signs, steam". Omni gives you real motion AND the sound of the place — rain, street noise, everything.
02
02Step 2
Drop one photo in
Add one clear photo of your character as a reference image. Omni keeps the same face in every shot. New place, new light, same person.
03
03Step 3
Make it talk
Put the words in quotes in your prompt: she says: "This is not a real video." The voice is generated together with the video. No dubbing, no lip-sync tools.
Copy · paste into Omni
The same woman from the reference photo in a softly lit studio, medium close-up, speaking to camera, she says: "This is not a real video." hyperrealistic, natural voice. No text, no watermarks.
Want this running for you?
Build it step by step inside AI Creative OS, with me in the thread.