← All models
LyviaVideo

Lyvia Lumi

Video generation with audio — text-to-video or image-to-video

Try Lyvia Lumi

About

Lyvia Lumi is a Pro-exclusive video model that generates video with synchronized audio. It works as text-to-video from a prompt alone, or image-to-video when you provide a start image. Supports durations from 4 to 10 seconds at 720p, with negative prompt support to exclude unwanted elements. One of the most affordable video models in the Lyvia lineup.

CinematicCreativeSocial MediaAmbient

How to prompt

  • Describe both the visual and the sound: "rain falling on a window, soft lo-fi music in the background"
  • Use negative prompts to exclude artifacts: "blurry, distorted faces, text overlays"
  • Provide a start image for image-to-video — the model animates forward from it
  • Shorter durations (4–6s) tend to be more coherent; use 8–10s for slower, ambient scenes
  • At 40 credits per 4s clip, Lumi is great for rapid iteration before committing to a higher-end model

Best for

Social media clips, ambient/mood video, quick video iteration, budget-friendly video with audio, image animation

Good to know

Pro subscribers only. 720p output only. Limited to 1 reference image. Shorter max duration (10s) compared to other video models.

Ready to create with Lyvia Lumi?

Open Lyvia