Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
cinience avatar

Alicloud Ai Video Wan Video

  • 307 installs
  • 396 repo stars
  • Updated July 18, 2026
  • cinience/alicloud-skills

alicloud-ai-video-wan-video is a Claude Code skill that integrates Alibaba Cloud WAN text and prompt-to-video generation APIs into agents, backends, and media products for developers needing managed cloud video creation.

About

alicloud-ai-video-wan-video is an Alibaba Cloud skill from cinience/alicloud-skills for wiring WAN model text-to-video and prompt-to-video generation into application backends and AI agents. It targets developers building media features who want managed cloud video APIs instead of self-hosted diffusion pipelines. Use it when an agent or service must submit generation prompts, handle asynchronous video jobs, and return hosted video assets through Alibaba Cloud's WAN video product. The skill fits products that combine LLM prompt orchestration with managed media output where operational overhead of custom GPU infrastructure is undesirable.

  • WAN video generation API setup
  • Prompt and parameter configuration
  • Async job lifecycle handling
  • Output download and storage
  • Agent tool integration patterns

Alicloud Ai Video Wan Video by the numbers

  • 307 all-time installs (skills.sh)
  • Ranked #510 of 1,335 Generative Media skills by installs in the Skillselion catalog
  • Data as of Aug 5, 2026 (Skillselion catalog sync)
npx skills add https://github.com/cinience/alicloud-skills --skill alicloud-ai-video-wan-video

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs307
repo stars396
Last updatedJuly 18, 2026
Repositorycinience/alicloud-skills

How do you integrate Alibaba Cloud WAN video generation?

Integrate Alibaba Cloud WAN text or prompt-to-video generation into agents, backends, or media products that need managed cloud video creation APIs.

Who is it for?

Backend and AI engineers on Alibaba Cloud who need managed WAN text-to-video APIs inside agents, microservices, or media products.

Skip if: Teams requiring fully offline or self-hosted video models with no Alibaba Cloud account or WAN service access.

When should I use this skill?

User mentions Alibaba Cloud WAN video, prompt-to-video APIs, or integrating managed AI video generation into an agent backend.

What you get

WAN video API integrations, prompt-to-video job handlers, and cloud-hosted generated video assets in agent or backend services.

  • WAN video API client integration
  • prompt-to-video job handlers
  • hosted video asset URLs

Files

SKILL.mdMarkdownGitHub ↗

Category: provider

Model Studio Wan Video

Validation

mkdir -p output/alicloud-ai-video-wan-video
python -m py_compile skills/ai/video/alicloud-ai-video-wan-video/scripts/generate_video.py && echo "py_compile_ok" > output/alicloud-ai-video-wan-video/validate.txt

Pass criteria: command exits 0 and output/alicloud-ai-video-wan-video/validate.txt is generated.

Output And Evidence

  • Save task IDs, polling responses, and final video URLs to output/alicloud-ai-video-wan-video/.
  • Keep one end-to-end run log for troubleshooting.

Provide consistent video generation behavior for the video-agent pipeline by standardizing video.generate inputs/outputs and using DashScope SDK (Python) with the exact model name.

Critical model names

Use one of these exact model strings:

  • wan2.6-t2v
  • wan2.6-t2v-us
  • wan2.2-t2v-plus
  • wan2.2-t2v-flash
  • wan2.6-i2v-flash
  • wan2.6-i2v
  • wan2.6-i2v-us
  • wanx2.1-t2v-turbo

Prerequisites

  • Install SDK (recommended in a venv to avoid PEP 668 limits):
python3 -m venv .venv
. .venv/bin/activate
python -m pip install dashscope
  • Set DASHSCOPE_API_KEY in your environment, or add dashscope_api_key to ~/.alibabacloud/credentials (env takes precedence).

Normalized interface (video.generate)

Request

  • prompt (string, required)
  • negative_prompt (string, optional)
  • duration (number, required) seconds
  • fps (number, required)
  • size (string, required) e.g. 1280*720
  • seed (int, optional)
  • reference_image (string | bytes, optional for t2v, required for i2v family models)
  • motion_strength (number, optional)

Response

  • video_url (string)
  • duration (number)
  • fps (number)
  • seed (int)

Quick start (Python + DashScope SDK)

Video generation is usually asynchronous. Expect a task ID and poll until completion. Note: Wan i2v models require an input image; pure t2v models such as wan2.6-t2v can omit reference_image.

import os
from dashscope import VideoSynthesis

# Prefer env var for auth: export DASHSCOPE_API_KEY=...
# Or use ~/.alibabacloud/credentials with dashscope_api_key under [default].

def generate_video(req: dict) -> dict:
    payload = {
        "model": req.get("model", "wan2.6-i2v-flash"),
        "prompt": req["prompt"],
        "negative_prompt": req.get("negative_prompt"),
        "duration": req.get("duration", 4),
        "fps": req.get("fps", 24),
        "size": req.get("size", "1280*720"),
        "seed": req.get("seed"),
        "motion_strength": req.get("motion_strength"),
        "api_key": os.getenv("DASHSCOPE_API_KEY"),
    }

    if req.get("reference_image"):
        # DashScope expects img_url for i2v models; local files are auto-uploaded.
        payload["img_url"] = req["reference_image"]

    response = VideoSynthesis.call(**payload)

    # Some SDK versions require polling for the final result.
    # If a task_id is returned, poll until status is SUCCEEDED.
    result = response.output.get("results", [None])[0]

    return {
        "video_url": None if not result else result.get("url"),
        "duration": response.output.get("duration"),
        "fps": response.output.get("fps"),
        "seed": response.output.get("seed"),
    }

Async handling (polling)

import os
from dashscope import VideoSynthesis

task = VideoSynthesis.async_call(
    model=req.get("model", "wan2.6-i2v-flash"),
    prompt=req["prompt"],
    img_url=req["reference_image"],
    duration=req.get("duration", 4),
    fps=req.get("fps", 24),
    size=req.get("size", "1280*720"),
    api_key=os.getenv("DASHSCOPE_API_KEY"),
)

final = VideoSynthesis.wait(task)
video_url = final.output.get("video_url")

Operational guidance

  • Video generation can take minutes; expose progress and allow cancel/retry.
  • Cache by (prompt, negative_prompt, duration, fps, size, seed, reference_image hash, motion_strength).
  • Store video assets in object storage and persist only URLs in metadata.
  • reference_image can be a URL or local path; the SDK auto-uploads local files.
  • If you get Field required: input.img_url, the reference image is missing or not mapped.
  • wan2.6-t2v and wan2.6-t2v-us add multi-shot narrative support and optional audio input according to the official docs.

Size notes

  • Use WxH format (e.g. 1280*720).
  • Prefer common sizes; unsupported sizes can return 400.

Output location

  • Default output: output/alicloud-ai-video-wan-video/videos/
  • Override base dir with OUTPUT_DIR.

Anti-patterns

  • Do not invent model names or aliases; use official Wan i2v model IDs only.
  • Do not block the UI without progress updates.
  • Do not retry blindly on 4xx; handle validation failures explicitly.

Workflow

1) Confirm user intent, region, identifiers, and whether the operation is read-only or mutating. 2) Run one minimal read-only query first to verify connectivity and permissions. 3) Execute the target operation with explicit parameters and bounded scope. 4) Verify results and save output/evidence files.

References

  • See references/api_reference.md for DashScope SDK mapping and async handling notes.
  • Source list: references/sources.md

Related skills

How it compares

Pick this over generic media skills when the stack is Alibaba Cloud WAN APIs rather than open-source local diffusion models.

FAQ

What does alicloud-ai-video-wan-video integrate?

alicloud-ai-video-wan-video integrates Alibaba Cloud WAN text-to-video and prompt-to-video generation APIs. Developers embed managed cloud video creation into agents, backends, or media products instead of operating self-hosted diffusion infrastructure.

Who should use the WAN video Alibaba Cloud skill?

Backend and AI engineers on Alibaba Cloud should use alicloud-ai-video-wan-video when products need programmatic prompt-to-video generation, asynchronous job handling, and hosted video assets delivered through WAN managed APIs.

Generative Mediallmautomation

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.