Pick the version and the input type
Choose the named version of PixVerse, then text to video, first-frame to video, last-frame, or reference still, according to what you actually have.
AI Video
PixVerse is an AI video generator for text to video and image to video. Describe a shot in words, or upload a first frame, last frame, or reference still, to get a short clip. Duration, aspect ratio and resolution are chosen before you generate, because they cannot be changed afterwards.
Choose the named version of PixVerse, then text to video, first-frame to video, last-frame, or reference still, according to what you actually have.
Name the subject, the camera move and the light. If you uploaded a frame, describe only what moves: what enters, what drifts, where the camera goes.
Lock length, aspect ratio and resolution before you run. Video takes longer than an image; you can leave the page while it renders.
A written description is enough to start: subject, camera move, light, and the one action that happens in the clip. Use this when you do not yet have a still to animate.
Upload a still as the first frame and describe only the motion. Some versions of PixVerse also take a last frame or a reference image, so the clip can start and land on pictures you supply. That is the most reliable way to control how the shot looks.
A longer duration stretches the same action; it does not add plot. If more has to happen, generate more clips and cut them together.
PixVerse is an AI video model for turning text or pictures into a short clip. Typical jobs are text to video, image to video, first-frame to video, and first-and-last-frame.
Yes, when the selected version accepts a still. Upload a first frame (and a last frame if that mode is listed) and describe the motion only.
Text to video is enough to start. A first frame is the most reliable way to control what the clip looks like. Use last frame when the shot has to land on a specific pose.
Duration is chosen before you generate and cannot be changed afterwards. A longer setting stretches the same action. For more events, generate more clips.