Wan Dancer AI Video Generator
Generate a choreographed dance clip from a single portrait and music track using the Wan Dancer AI Video Generator

video0

video1

video2

video3

video4

AI Video Prompt Generator

Feedback

AI Ad Video Example

Loading...

Wan Dancer AI Video Generator

Create seamless dance animations from one photo and any song using the Wan Dancer AI Video Generator. Enjoy 720p output at 30fps with over 60 seconds of steady motion—free and open-source on Apache-2.0.

All Tools

Discover our comprehensive AI-powered animation toolkit

Why Use the Wan Dancer AI Video Generator

Developed by Alibaba Tongyi Lab, the Wan Dancer AI Video Generator (Wan-Dancer-14B) transforms a reference portrait and audio track into a choreographed dance video at 720p/30fps. It delivers rhythm-synchronized movement that remains visually stable beyond one minute, all without motion capture data.

  • Beat-Aligned Movement From Your Audio
    Choreography is derived directly from your uploaded track, so each motion locks precisely onto the rhythm rather than following a preset loop.
  • One Portrait, Consistent Identity
    A single head-to-toe photo anchors facial features, hairstyle, and wardrobe so the performer stays recognizable from opening to closing frame.
  • Sustained Visual Stability Beyond 60 Seconds
    The two-pass planning pipeline avoids the typical breakdown point near 20 seconds that affects most diffusion approaches, holding structure for a full minute or more.

How to Use the Wan Dancer AI Video Generator

Follow three straightforward steps to produce a fully rhythm-synced dance clip from a portrait and a song.

Wan Dancer AI Video Generator Key Features

Five trained dance genres, publicly available model weights, and full ComfyUI compatibility make this music-to-dance engine ideal for extended single-subject routines.

Audio-Waveform Choreography Engine

Motion vectors are computed from the actual frequency spectrum, guaranteeing steps match real percussion instead of an approximate loop pattern.

Extended Duration Without Drift

A global-then-local two-pass rendering strategy preserves spatial integrity well beyond 20 seconds — easily covering an entire chorus section.

Reference Subject Tracking

Facial landmarks, hair shape, and garment details extracted from your input photo stay locked throughout every sequence, keeping the performer identifiable start to finish.

Crisp 720p / 30fps Rendering

Fluid half-inch motion renders at 60 source samples per second, purpose-built for vertical platforms like TikTok, Reels, and Shorts.

Five Distinct Dance Traditions

Coverage spans Chinese classical, K-pop, street freestyle, tap, and Latin ballroom, letting one portrait adapt to a broad spectrum of musical moods.

Public Weights Under Apache-2.0

Full model checkpoints are downloadable from Hugging Face and ModelScope, accompanied by ComfyUI nodes and LoRA adapters for personalized routines.

FAQ

Wan Dancer AI Video Generator — FAQ

Answers to the most frequent questions about setup, output quality, supported styles, and licensing.

1

What exactly does this tool do?

It reads one head-to-toe portrait plus an audio file and renders a 720p 30fps routine where every step lines up with the beat — no motion-capture session required. The model behind it, Wan-Dancer-14B, comes from Alibaba Tongyi Lab and ships openly under Apache-2.0.

2

What is the underlying pipeline?

First, a coarse planner scans the entire song and plots keyframes across its duration; second, an interpolator fills every intermediate frame with fine-grained movement. Separating these two phases prevents the gradual drift that plagues single-pass generators.

3

Which inputs should I prepare?

Gather a well-lit vertical photograph showing the full body, a clean audio recording, and one sentence describing the dance style. All three together give the strongest results.

4

Is there a maximum clip length?

Target durations run up to roughly 90 seconds while retaining consistent framing and timing. By comparison, most diffusion-based alternatives start degrading around the 20-second mark.

5

Which choreography genres are included?

Five categories ship out of the box: Chinese classical, K-pop, street, tap, and Latin. Simply name one in your prompt to steer the routine accordingly.

6

Can I self-host or fine-tune the model?

Absolutely. Both Hugging Face and ModelScope host the complete Apache-2.0 release, including inference scripts, ComfyUI nodes, and LoRA adapters you can adapt to bespoke choreography.

Start Your First Dance Video Now

Upload a portrait, pick a track, and let the Wan Dancer AI Video Generator choreograph a beat-accurate clip in seconds.