AI Video ModelText & image modesDynamic motionFlexible formats

Fast, flexible generation—built to iterate.

Wan 2.1 supports multiple creation modes and strong motion handling—ideal for rapid experimentation, style exploration, and campaign sprints.

Wan 2.1 flexible AI video modes
Wan 2.1 Text to video

Text to video

Text-to-video and image-to-video in the same creative loop

Wan 2.1 Image to video

Image to video

Strong motion handling and energetic scene dynamics

Wan 2.1 Style explore

Style explore

Flexible aspect ratios for modern content formats

1

Text & image modes

Text-to-video and image-to-video in the same creative loop

2

Dynamic motion

Strong motion handling and energetic scene dynamics

3

Flexible formats

Flexible aspect ratios for modern content formats

4

Style exploration

Fast style exploration for concepts, music visuals, and A/B hooks

Best for

Real scenarios you can ship today.

Fast concept sprints for campaignsStylized motion tests for music and editsSocial content variations across formats

Workflow

How it works

  1. 1

    Pick a mode

    Start from text or an image depending on whether you need a blank canvas or a locked look.

  2. 2

    Explore styles fast

    Generate variations to compare motion, mood, and format without leaving Cuta.

  3. 3

    Lock the winner

    Keep the strongest take, refine the prompt, and export for your campaign.

FAQ

Frequently asked

When should I use Wan 2.1?+

Use Wan when you need speed and flexibility—concepting, style tests, and social variations that benefit from quick iteration.

Does Wan support image-to-video?+

Yes. Bring a still into motion, or start from a text prompt when you want a fresh scene.

Is Wan good for social formats?+

Yes. Flexible aspect ratios make it easy to produce vertical, square, or landscape cuts for different channels.

Can Wan handle stylized looks?+

Wan is a strong pick for exploring multiple visual styles quickly before you commit to a final direction.