Alibaba · 影片

OpenClaw x Wan

Wan 是 Alibaba 的全方位影片及圖片模型套件,是現有最完整的開放權重影片模型系列之一。透過 RunAPI,所有 Wan 版本共用同一個 key 及統一計費。

18 variants 由 $0.05 / 次請求 可商用

Prerequisite: npx runapi mcp install

Prompt

Prompt 模型

任務合適時使用 wan-2.2-a14b-image-to-video-turbo:基於 2.2 架構、以圖片為基礎的影片。

You have access to RunAPI task tools for Wan.

Available Wan models:
- wan-2.2-a14b-image-to-video-turbo: 基於 2.2 架構、以圖片為基礎的影片
  endpoints: /api/v1/wan/image_to_video
  request fields: model, prompt, first_frame_image_url
- wan-2.2-a14b-speech-to-video-turbo: 由音訊輸入驅動影片動態及對嘴
  endpoints: /api/v1/wan/speech_to_video
  request fields: model, source_image_url, source_audio_url, prompt
- wan-2.2-a14b-text-to-video-turbo: 開源 A14B MoE;基準質素
  endpoints: /api/v1/wan/text_to_video
  request fields: model, prompt
- wan-2.2-animate-move: 基於 2.2 的專門動態動畫
  endpoints: /api/v1/wan/animate
  request fields: model, source_image_url, reference_video_url
- wan-2.2-animate-replace: 基於 2.2 的主體替換動畫
  endpoints: /api/v1/wan/animate
  request fields: model, source_image_url, reference_video_url
- wan-2.5-image-to-video: 基於 2.5、以圖片為基礎並具原生音訊
  endpoints: /api/v1/wan/image_to_video
  request fields: model, prompt, first_frame_image_url, duration_seconds, output_resolution
- wan-2.5-text-to-video: 原生音訊生成 + 動態真實感比 2.2 更佳
  endpoints: /api/v1/wan/text_to_video
  request fields: model, prompt, output_resolution
- wan-2.6-edit-video: 影片編輯,具主體 + 鏡頭控制
  endpoints: /api/v1/wan/edit_video
  request fields: model, source_video_urls, prompt
- wan-2.6-flash-edit-video: 針對速度優化的 2.6 影片編輯
  endpoints: /api/v1/wan/edit_video
  request fields: model, source_video_urls, prompt, audio
- wan-2.6-flash-image-to-video: 針對速度優化的 2.6;成本更低,推理更快
  endpoints: /api/v1/wan/image_to_video
  request fields: model, prompt, first_frame_image_url, audio
- wan-2.6-image-to-video: 以圖片為基礎,具備 2.6 的場景連續性
  endpoints: /api/v1/wan/image_to_video
  request fields: model, prompt, first_frame_image_url
- wan-2.6-text-to-video: 場景連續性提升 + 支援複雜場景
  endpoints: /api/v1/wan/text_to_video
  request fields: model, prompt
- wan-2.7-edit-video: wan-2.7-edit-video
  endpoints: /api/v1/wan/edit_video
  request fields: model, source_video_urls, prompt, source_video_url
- wan-2.7-image: 基於 2.7 架構的標準圖片生成
  endpoints: /api/v1/wan/text_to_image
  request fields: model, prompt
- wan-2.7-image-pro: 3×3 多角度圖片網格;最高質素圖片輸出
  endpoints: /api/v1/wan/text_to_image
  request fields: model, prompt
- wan-2.7-image-to-video: 以圖片為基礎,具備 2.7 多參考支援
  endpoints: /api/v1/wan/image_to_video
  request fields: model, prompt, first_frame_image_url
- wan-2.7-r2v: R2V 文字生成影片,支援角色外觀 + 聲音參考
  endpoints: /api/v1/wan/text_to_video
  request fields: model, prompt
- wan-2.7-text-to-video: 首尾幀控制;同時支援 5 段參考影片
  endpoints: /api/v1/wan/text_to_video
  request fields: model, prompt

Use the model ID and endpoint that match the user's request.
Example prompt: 生成一段 5 秒影片:一隻貓跳上書架,自然室內光線,手持鏡頭感。
/api/v1/wan/image_to_video: Submit the task, poll for its status, then verify the completed output. POST /api/v1/wan/image_to_video,並儲存回傳的 Task ID。
GET /api/v1/wan/image_to_video/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/wan/speech_to_video: Submit the task, poll for its status, then verify the completed output. POST /api/v1/wan/speech_to_video,並儲存回傳的 Task ID。
GET /api/v1/wan/speech_to_video/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/wan/text_to_video: Submit the task, poll for its status, then verify the completed output. POST /api/v1/wan/text_to_video,並儲存回傳的 Task ID。
GET /api/v1/wan/text_to_video/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/wan/animate: Submit the task, poll for its status, then verify the completed output. POST /api/v1/wan/animate,並儲存回傳的 Task ID。
GET /api/v1/wan/animate/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/wan/edit_video: Submit the task, poll for its status, then verify the completed output. POST /api/v1/wan/edit_video,並儲存回傳的 Task ID。
GET /api/v1/wan/edit_video/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/wan/text_to_image: Submit the task, poll for its status, then verify the completed output. POST /api/v1/wan/text_to_image,並儲存回傳的 Task ID。
GET /api/v1/wan/text_to_image/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。

公開版本與端點

模型 ID 端點 起始價格 模型目錄
wan-2.2-a14b-image-to-video-turbo
/api/v1/wan/image_to_video
$0.40 / 次請求 模型詳情
wan-2.2-a14b-speech-to-video-turbo
/api/v1/wan/speech_to_video
$0.24 / 秒 模型詳情
wan-2.2-a14b-text-to-video-turbo
/api/v1/wan/text_to_video
$0.40 / 次請求 模型詳情
wan-2.2-animate-move
/api/v1/wan/animate
$0.13 / 秒 模型詳情
wan-2.2-animate-replace
/api/v1/wan/animate
$0.13 / 秒 模型詳情
wan-2.5-image-to-video
/api/v1/wan/image_to_video
$0.12 / 秒 模型詳情
wan-2.5-text-to-video
/api/v1/wan/text_to_video
$0.12 / 秒 模型詳情
wan-2.6-edit-video
/api/v1/wan/edit_video
$0.14 / 秒 模型詳情
wan-2.6-flash-edit-video
/api/v1/wan/edit_video
$0.30 / 次請求 模型詳情
wan-2.6-flash-image-to-video
/api/v1/wan/image_to_video
$9.44 / 次請求 模型詳情
wan-2.6-image-to-video
/api/v1/wan/image_to_video
$0.58 / 秒 模型詳情
wan-2.6-text-to-video
/api/v1/wan/text_to_video
$0.58 / 秒 模型詳情
wan-2.7-edit-video
/api/v1/wan/edit_video
$0.16 / 秒 模型詳情
wan-2.7-image
/api/v1/wan/text_to_image
$0.05 / 次請求 模型詳情
wan-2.7-image-pro
/api/v1/wan/text_to_image
$0.12 / 次請求 模型詳情
wan-2.7-image-to-video
/api/v1/wan/image_to_video
$0.16 / 秒 模型詳情
wan-2.7-r2v
/api/v1/wan/text_to_video
$0.16 / 秒 模型詳情
wan-2.7-text-to-video
/api/v1/wan/text_to_video
$0.16 / 秒 模型詳情

驗證

輪詢,直到 Task 進入終態

選擇 <model-id> 後產生驗證命令。

設定

指南端點: <endpoint>

選擇 <model-id> 後,依端點的公開輸入契約產生請求。
使用流程

三步開始使用

  1. 選擇模型 ID

    選擇公開模型 ID,並查看其端點與目前起始價格。

  2. 設定 RunAPI

    呼叫端點前設定 RUNAPI_API_KEY。

  3. 驗證結果

    對非同步端點輪詢同一路徑,直到 Task 進入終態。

使用 OpenClaw + Wan 可以建立甚麼

  • 從分鏡到影片

    用首幀和尾幀錨定每個片段,把分鏡畫面轉成鏡頭之間視覺連貫的影片序列。

  • 虛擬主持人與品牌吉祥物

    用一張人臉圖片和一段音訊生成口播影片,口型同步和頭部動作由模型完成。

  • 角色一致的多鏡頭序列

    製作同一角色出現在多個片段中的敘事內容,鏡頭之間臉部和服裝保持穩定。

為甚麼透過 RunAPI + OpenClaw 使用 Wan

  • 18 個變體,一個 API 金鑰

    透過一個 RunAPI 連線選擇可用的模型變體,無需更改現有整合。

  • 清晰的用量價格

    發送請求前查看目錄中的目前價格,無需訂閱或最低消費。

  • 自動化工作流程

    透過一致的工作流程提交工作、查詢狀態並收集非同步結果,無需編寫手動輪詢程式碼。

OpenClaw + Wan 常見問題

在 OpenClaw 中可以調用哪些 Wan 端點?

本頁列出的所有 Wan 端點,包括影片生成、圖片轉影片、語音轉影片、影片編輯和圖片生成。每個端點使用各自的模型版本。

Wan 的語音轉影片怎麼使用?

提供一段音訊和一張人臉圖片,Wan 會生成人物說出這段音訊、口型同步的影片。

Wan 是開源的嗎?

Wan 以開放權重發布。透過 RunAPI 使用時不必自行設定 GPU,一組 API 金鑰即可調用;如果需要自行託管的流程,同樣的權重也能在你自己的基礎架構上執行。

應該使用哪個模型 ID?

從版本表選擇公開模型 ID。每個 ID 可用的端點均列於相應資料列。

本指南會設定聊天模型嗎?

不會。此 Model Line 使用頁面所示的端點流程,不會被描述為 Agent 聊天模型。

透過 OpenClaw 開始使用 Wan

正在與團隊一起建立?

我們可以協助企業設定、整合和技術問題。

聯絡我們