Skip to content

About Text to Video

Write the scene the way you would brief a crew — who is in it, what happens, how it is lit — and pick a camera move and a model — Lunar Motion, Veo, Kling and more.

Models you can use

  • Lunar Motion 2.5 by Lunar Labs — Our flagship. Thirty-second takes with sound, lip sync and nine-image character lock.
  • Veo 3.1 by Google — Photoreal footage with dialogue and sound design, up to native 4K.
  • Kling 3.0 Pro by Kuaishou — Fluid human motion and precise start and end frames.
  • Seedance 2.0 by ByteDance — Sharp 1080p with references, motion transfer and sound.
  • Hailuo H3 by MiniMax — Expressive physics and dramatic action, fast.
  • Wan 3.0 by Alibaba — Faithful to long, detailed prompts, with sound.

Direct the camera

Pick from 23 named camera moves — dolly in, crash zoom, orbit, crane, FPV drone — and set the film look: camera body, lens, focal length, aperture, shutter, film stock, lighting and colour palette. See every move.

How it works

  1. Step 1

    Write the scene

    Describe who is in it, what happens, where and when. Short and concrete beats long and vague.

  2. Step 2

    Choose the model and the camera

    Pick a model, a length, a resolution and, if you like, a named camera move such as a dolly in or an orbit.

  3. Step 3

    Generate

    The price is on the button before you press it. The clip, with native sound where the model has it, lands in your library.

Questions

Can AI text to video make sound?
Yes. Lunar Motion 2.5, Veo 3.1, Kling 3.0 Pro, Seedance 2.0 and Wan 3.0 generate sound with the picture; you can steer it with a line of sound direction.
How much does it cost?
A 5-second 480p SD clip on Lunar Motion 2.5 is 50 credits; longer, sharper or other models cost more, and the exact price is on the Generate button before you press it. New accounts get 60 free credits to try it.
Can I use the videos commercially?
Yes, on any paid plan: Starter, Creator and Studio include commercial use. Free-plan downloads carry a watermark.