Skip to main content

Wan 2.5 Speak

Wan 2.5 Speak is the Lipsync Studio version of Alibaba’s Wan 2.5. Instead of uploading an audio file, you write what the subject should say in Audio Text, and the model generates the video with matching speech. In the model list it is described as “The best in class open source model with audio”.

Specifications

How to use

1

Open Lipsync Studio

Go to Video and select Lipsync Studio from the Video Tools tabs.
2

Select the model

Open Select Model and choose Wan 2.5 Speak.
3

Set aspect ratio, duration and resolution

Choose 16:9 or 9:16, 5s or 10s, and a resolution.
4

Add your image, script and prompt

Upload the image under Upload Image. Type the spoken lines in Audio Text, and describe the scene or motion in Prompt.
5

Generate

Click Generate.

Other Lipsync Studio models