video
7
modify
pixverse modify — edit video content using text prompts. swap subjects, add elements, or restyle regions with reference images.

avatar
pixverse avatar — generate talking avatar video from a portrait image. supports audio file or text-to-speech input at up to 1080p.

lip-sync
pixverse lip sync — align speech to mouth movement in video. supports audio file input or text-to-speech with selectable voices.

fusion
pixverse fusion — generate video from reference images. compose 1-3 subject/background images into a video scene with text-guided motion.

extend
pixverse extend — extend an existing video with ai-generated continuation. supports 5s or 8s extensions at up to 1080p.

c1
pixverse c1 — advanced video generation model. text-to-video and image-to-video with up to 1080p resolution and 15s duration. per-second billing.

v6
pixverse v6 — versatile video generation model. text-to-video and image-to-video with up to 1080p resolution and 15s duration. per-second billing.
explore more on inference.sh
pixverse is one of many providers on the grid. discover hundreds of apps across image, video, audio, and more.
we use cookies
we use cookies to ensure you get the best experience on our website. for more information on how we use cookies, please see our cookie policy.
by clicking "accept", you agree to our use of cookies.
learn more.