Gemini Omni is a multimodal AI video creation platform for turning text, images, video, and audio references into cinematic video. Create from a prompt, guide results with visual references, refine scenes through natural-language instructions, and produce high-resolution video ā all directly in your browser.
Hey Product Hunt š
We built Gemini Omni to make advanced AI video creation feel less like operating complicated editing software and more like having a conversation with a creative partner.
Most AI video workflows still involve jumping between separate tools for prompting, image references, video generation, editing, and upscaling. We wanted to bring those steps together into one browser-based workflow.
With Gemini Omni, you can:
š¬ Generate videos from natural-language prompts
š¼ļø Use images and video as creative references
š¬ Refine scenes through conversational instructions
š„ Control characters, scenes, camera movement, and visual direction
⨠Iterate from quick drafts to high-resolution outputs
The platform is built for creators, filmmakers, marketers, designers, and anyone experimenting with multimodal video generation.
New users can start with free credits, with no credit card required.
We'd especially love feedback from the Product Hunt community on the multimodal workflow and the overall creation experience.
What would you most want to control through conversation when creating an AI video?