wan-video

v2026.09.24

Generate AI videos with Wan (Alibaba) via AceDataCloud API. Use when creating videos from text prompts or animating images into video. Supports text-to-video, image-to-video, reference video transfer, multi-resolution (480P-1080P), and optional audio.

GitHub
安装命令
npx skhub add acedatacloud/wan-video
Markdown
SKILL.md

Wan Video Generation

Generate AI videos through AceDataCloud's Wan (Alibaba) API.

Setup: See authentication for token setup.

Quick Start

curl -X POST https://api.acedata.cloud/wan/videos \
  -H "Authorization: Bearer $ACEDATACLOUD_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"action": "text2video", "prompt": "a dolphin jumping through ocean waves at golden hour", "model": "wan2.6-t2v"}'

Async: See async task polling. Poll via POST /wan/tasks with {"id": "..."}.

Models

ModelTypeBest For
wan2.6-t2vText-to-VideoCreating video from text description
wan2.6-i2vImage-to-VideoAnimating a still image into video
wan2.6-r2vReference Video-to-VideoCharacter extraction and transfer from reference video
wan2.6-i2v-flashImage-to-Video (Fast)Quick image-to-video generation
wan3.0-videoAll-in-OneText, frames, reference media, files, and public links

Workflows

1. Text-to-Video

POST /wan/videos
{
  "action": "text2video",
  "prompt": "a time-lapse of flowers blooming in a meadow",
  "model": "wan2.6-t2v",
  "resolution": "720P",
  "duration": 5
}

2. Image-to-Video

Animate a still image into a video clip.

POST /wan/videos
{
  "action": "image2video",
  "prompt": "gentle wind blows through the scene",
  "model": "wan2.6-i2v",
  "image_url": "https://example.com/landscape.jpg",
  "resolution": "720P",
  "duration": 5
}

3. Image-to-Video (Flash)

Faster image-to-video generation with reduced latency.

POST /wan/videos
{
  "action": "image2video",
  "prompt": "camera slowly pans across the landscape",
  "model": "wan2.6-i2v-flash",
  "image_url": "https://example.com/scene.jpg"
}

4. Reference Video Transfer

Extract characters or timbres from a reference video and transfer them into a new generation.

POST /wan/videos
{
  "action": "text2video",
  "prompt": "the character walks through a futuristic city at night",
  "model": "wan2.6-r2v",
  "reference_video_urls": ["https://example.com/reference.mp4"]
}

5. Multi-Cut Editing

Generate a video with multiple shots rather than a single continuous take.

POST /wan/videos
{
  "action": "text2video",
  "prompt": "a chef preparing a meal in a busy kitchen",
  "model": "wan2.6-t2v",
  "shot_type": "multi",
  "duration": 10
}

6. Video with Audio

Enable audio generation alongside the video.

POST /wan/videos
{
  "action": "text2video",
  "prompt": "ocean waves crashing on a rocky shore",
  "model": "wan2.6-t2v",
  "audio": true
}

Parameters

ParameterRequiredValuesDescription
actionYes"text2video", "image2video"Action type
promptYesstringScene description
modelYes"wan2.6-t2v", "wan2.6-i2v", "wan2.6-r2v", "wan2.6-i2v-flash"Model
image_urlFor image2videostringSource image URL (required for image-to-video)
negative_promptNostring (max 500 chars)Content to exclude from generation
reference_video_urlsFor r2varray of stringsReference videos for character/timbre extraction
shot_typeNo"single", "multi"Continuous shot or multi-cut editing
audioNobooleanEnable audio in the generated video
audio_urlNostringReference audio URL
resolutionNo"480P", "720P", "1080P"Output resolution (default: 720P)
sizeNostringThe size of the generated video
durationNo5, 10, 15Video duration in seconds
prompt_extendNobooleanEnable LLM-based prompt rewriting
callback_urlNostringAsync webhook notification URL

Gotchas

  • image_url is required for wan2.6-i2v and wan2.6-i2v-flash models
  • reference_video_urls is used only with wan2.6-r2v for character/timbre transfer
  • negative_prompt has a maximum length of 500 characters
  • Supported durations are 5, 10, or 15 seconds only
  • Default resolution is 720P; use 1080P for higher quality at increased cost
  • shot_type: "multi" produces multi-cut edits rather than a single continuous shot

MCP: pip install mcp-wan | Hosted: https://wan.mcp.acedata.cloud/mcp | See all MCP servers

Wan 3 All-in-One

wan3.0-video infers the workflow from media. Supported types are first_frame, last_frame, reference_image (up to 10), reference_video (up to 5, 15 seconds total), reference_audio (up to 5, 15 seconds total), file, and link. Frame mode cannot be mixed with reference/file/link mode.

POST /wan/videos
{
  "model": "wan3.0-video",
  "prompt": "Use image 1 to create a cinematic product video",
  "media": [{"type":"reference_image","url":"https://cdn.acedata.cloud/r9vsv9.png"}],
  "resolution": "720P",
  "ratio": "16:9",
  "duration": 5,
  "audio": true,
  "async": true
}

Wan 3 supports 2–30 seconds or -1 automatic duration. Poll with the existing /wan/tasks endpoint.

发现
标签

此技能尚未发布标签。

版本
最新版本元数据

版本

v2026.09.24

发布时间

2026年9月24日

分类

未分类

许可证

NOASSERTION

源路径

skills/wan-video

默认分支

main

最新提交

57cc298

Tree SHA

acaa402