Google's any-input video model. Create 4-10 second clips with native audio from a prompt, an image, three reference images, or an existing video — then reshape scenes with plain language. Grounded in Gemini's real-world knowledge for coherent physics and motion.
~2-4 min
curl https://www.flixly.ai/api/v1/generate \
-H "Authorization: Bearer flx_live_..." \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-omni-flash",
"prompt": "A beautiful sunset over the ocean",
"type": "TEXT_TO_VIDEO"
}'Get your API key from the Developer Portal.
Sora 2 with enhanced quality and longer durations
State-of-the-art video generation model with exceptional quality and understanding
Premium video generation model with native audio
ByteDance's newest video model. Generates up to 30 seconds in a single shot with synchronized audio — dialogue, music, and effects in the same pass — from a prompt, a start frame, or up to 50 image, video, and audio references.