> ## Documentation Index
> Fetch the complete documentation index at: https://docs.imagine.art/llms.txt
> Use this file to discover all available pages before exploring further.

# Kling Avatars 2.0 Pro

> Kling AI's audio-driven avatar model in ImagineArt Lipsync Studio — turn a single image into a talking video.

## Kling Avatars 2.0 Pro

Kling Avatars 2.0 Pro is Kling AI's avatar model for talking videos. You give it one image and an audio track, and it animates the subject's lips, face and upper body to match the speech. It works with photos as well as stylized or illustrated characters. In ImagineArt it appears in **Lipsync Studio**.

## Specifications

| Feature | Details |
| - | - |
| **Developer** | Kling AI (Kuaishou) |
| **Where to find it** | Video Studio → **Lipsync Studio** |
| **Inputs** | One image (JPG, JPEG, PNG or WEBP, up to 30MB), audio and an optional prompt |
| **Audio** | Upload an audio file with **Add Audio**, or create one with **Generate Audio** |
| **Duration** | Set by the length of your audio |
| **Settings** | No resolution, duration or aspect ratio controls |

## How to use

<Steps>
  <Step title="Open Lipsync Studio">
    Go to **Video** and select **Lipsync Studio** from the Video Tools tabs.
  </Step>

  <Step title="Select the model">
    Open **Select Model** and choose **Kling Avatars 2.0 Pro**.
  </Step>

  <Step title="Add your image, audio and prompt">
    Upload the image under **Upload Image**, then click **Add Audio** or **Generate Audio**. Use **Prompt** to guide expression and gestures.
  </Step>

  <Step title="Generate">
    Click **Generate**.
  </Step>
</Steps>

<Tip>
  To generate speech with Kling's own audio instead of uploading a track, use [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) in Lipsync Studio, which takes an **Audio Text** script.
</Tip>

## Other Lipsync Studio models

| Model | Resolution | Inputs |
| - | - | - |
| **Kling Avatars 2.0 Pro** | — | Image + audio + prompt |
| [Fabric 1.0](/ai-models/video/fabric-1-0) | 480p–720p | Image + audio |
| [Infini Talk](/ai-models/video/infini-talk) | 480p–720p | Image + audio + prompt |
| [OmniHuman](/ai-models/video/omnihuman) | — | Image + audio + prompt |
| [Wan 2.5 Speak](/ai-models/video/wan-2-5-speak) | 480p–1080p | Image + audio text + prompt |


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.