DataEyesAI
Official SiteConsoleDocs Home
Getting StartedDeveloper ToolsAI Models API
Official SiteConsoleDocs Home
Getting StartedDeveloper ToolsAI Models API
  1. Wan Video Generation
  • Getting Started
    • Overview
    • Console (Getting Started)
    • API Key
    • Base URL
  • Developer Tool Integration
    • OpenClaw
    • Claude Code
    • Codex
    • Gemini CLI
    • Grok CLI
    • Other Tools
  • AI Models API
    • OpenAI format (supports major original models)
      • Chat (Response)
        • Create Network Search
        • Create Model Response GPT-5 Enable Thinking
        • Create Function Call
        • Create Model Response
        • Create Model Response (Streaming Return)
        • Create Model Response (Control Thinking Length)
      • ChatGPT Interface
        • Audio
          • Audio to text gpt-4o-transcribe
          • GPT-4o-audio
          • Audio to text whisper-1
          • Audio to text gpt-4o-transcribe
          • Create voice gpt-4o-mini-tts
        • Chat
          • Create chat-based image recognition (non-streaming)
          • Create chat-based image recognition (streaming)
          • Create chat-based image recognition (streaming) best64
          • Official N test
          • Create structured output
          • Control the effort level of the inference model
          • Create chat function call
          • deepseek-ocr recognition
          • Create chat completion (non-stream)
        • Completions
          • ChatGPT automatic completion
          • Create completion
      • Image
        • Edit image
        • Create chat completion (streaming)
        • Create chat completion (qwen-mt-turbo)
        • Create chat completion with deepseek v3.1 level of reasoning (streaming)
      • Audio
        • Speech recognition
        • Speech synthesis
        • Official Function Calling invocation
        • Create chat-generated images (non-streaming)
      • Embedding
        • Text embeddings
    • Anthropic format
      • Chat
      • Chat(prompt cache)
      • Streaming response
      • Chat (deep reasoning)
      • Tool invocation (function call)
      • Analyze image
    • Google Gemini interface
      • Native format
        • Text-to-image + control over aspect ratio + clarity
        • Generate image
        • Text generation
        • Text generation - stream
        • Text generation + reasoning - stream
        • Image generation
        • Formatted output
        • Function call
        • Document understanding
        • URL context [native format]
        • Code execution
        • Video understanding
        • URL context
        • Video understanding - url [native format]
        • Imagen 4
        • Audio understanding
        • Embeddings
        • Chat
        • Edit image
      • Image-to-image Base64 request method
        • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity
        • Image editing
        • Single image gemini-3-pro-image-preview, controlling aspect ratio and clarity.
        • Image generation( gemini-2.5-flash-image)
        • Image generation gemini-2.5-flash-image, controlling aspect ratio.
        • Image understanding
      • Image-to-image URL request returns URL request format OpenAI
        • Single image generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
        • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
        • Image understanding
    • NanoBanana
      • OpenAI request
        • Edit image
        • OpenAI image format
      • Gemini request
        • Generate image
        • Edit image
    • Midjourney format
      • Midjourney API Reference
      • Task query interface
      • Upload image
      • Get seed (Seed)
      • Submit Imagine task
      • Query tasks based on ID list
      • FaceSwap
      • Execute Action operation
      • /mj/submit/blend
      • Submit Describe task
      • Submit Modal
      • Refresh link
      • Edit image
      • Query task status by task ID
      • Get the seed of the task image
    • Doubao - Painting
      • doubao-seededit-3-0-i2i-250628
      • doubao-seedream-4-0-250828 - text-to-image
      • doubao-seedream-4-0-250828 - image-to-image
      • doubao-seedream-4-0-250828 - multi-image generation
    • Rerank Reordering Model
      • Rerank
    • Video Model
      • Grok Video Generation
        • 00-Overview
        • 01-Text-to-Video
        • 02-Image-to-Video
        • 03-Reference-to-Video
        • 04-Video-Editing
        • 05-Video-Extension
      • Seedance Video Generation
        • 00-Overview
        • 01-Create-Video-Generation-Task
        • 02-Query-Video-Generation-Task
        • 03-Query-Video-Generation-Task-List
        • 04-Cancel-or-Delete-Task
        • Seedance Private Asset Library API Documentation
      • MiniMax-H3 Video Generation
        • 00-Overview
        • 01-Create-Video-Generation
        • 02-Create-Video-Regeneration
        • 03-Create-H3-Context-IR
        • 04-Query-Task
        • 05-List-Tasks
        • 06-Cancel-or-Delete-Task
      • Hailuo Video Generation
        • 00-Overview
        • 01-Text-to-Video-T2V
        • 02-Image-to-Video-I2V
        • 03-First-Last-Frame-FL2V
        • 04-Subject-Reference-S2V
        • 05-Query-Task-Status
        • 06-Video-Download
        • 99-Appendix-Camera-Movement-and-Webhooks
      • Jimeng Video Generation
        • 00-Overview
        • 01-3.0-Pro-Video-Generation
        • 02-720P-Text-to-Video
        • 03-720P-Image-to-Video-First-Frame
        • 04-720P-Image-to-Video-Start-End-Frame
        • 05-720P-Image-to-Video-Camera
        • 06-1080P-Text-to-Video
        • 07-1080P-Image-to-Video-First-Frame
        • 08-1080P-Image-to-Video-Start-End-Frame
        • 09-Error-Codes
      • Kling AI Video Generation
        • 00-Overview
        • 01-Text-to-Video
        • 02-Image-to-Video
        • 03-Omni-Video
        • 04-Multi-Image-to-Video
        • 05-Motion-Control
        • 06-Multi-Elements
        • 07-Video-Extension
        • 08-Lip-Sync
        • 09-Avatar
        • 10-Text-to-Audio
        • 11-Video-to-Audio
        • 12-TTS
        • 13-Custom-Voices
        • 14-Image-Recognition
        • 15-Element-Management
        • 16-Video-Effects
      • Vidu Video Generation
        • 00-Overview
        • 01-Text-to-Video
        • 02-Image-to-Video
        • 03-Reference-to-Video
        • 04-Start-End-Frame
        • 05-Multi-Frame
        • 06-Scene-Template
        • 07-Template-Story
        • 08-Query-Tasks
      • HappyHorse
        • HappyHorse Text-to-Video
        • HappyHorse Image-to-Video (First Frame)
        • HappyHorse Reference-to-Video
        • HappyHorse Video Editing
      • Wan Video Generation
        • 00-Overview.md
        • 01-Text-to-Video
        • 02-Image-to-Video
        • 03-Reference-to-Video
        • 04-Video-Editing
        • 05-First-Last-Frame-to-Video
        • 06-Motion-Transfer-and-Character-Swap
        • 07-Digital-Human-Video
        • 08-VACE-Video-Editing
        • 09-Query-Task
    • Audio API
      • Audio API
      • Gemini TTS API
      • Google DeepMind Lyria API
      • Elevenlabs Speech to Text API Reference
      • Text-to-Music Suno
        • Task Submission
          • Generate Song (Inspiration Mode)
          • Generate Song (Custom Mode)
          • Generate Song (Continuation Mode)
          • Generate Song (Singer Style)
          • Generate Song (Secondary Creation from Uploaded Song)
          • Generate Song (Song Stitching)
          • Generate Lyrics
          • Song Stitching
        • Query Interface
          • Batch Retrieve Tasks
          • Query Single Task
  • Search / Reader Product
    • Web Reader API​​
      • Web Reader API
      • Web Reader API(HK)
    • Web Search API​​
      • Modal Card API
        • Weather
          • All City ID
          • Weather Query API
      • Web Search API
      • Video Search api
      • Trending Search API
    • Document OCR Parsing API
      • fiel upload
      • URL Parsing
  • Advanced & System API
    • Data Updates
    • System interface
      • API Key & Quota Query API
      • API Key Management API
    • API Reference​​
      • Error Codes
      • HTTP Notes
    • List models
      • Models
  1. Wan Video Generation

01-Text-to-Video

Text-to-Video#

Doc version:v1.0.0 | Last updated:2026-07-22
This platform fully supports the official Tongyi Wanxiang (Wan) video generation API. Requests and responses are transparently proxied; parameter semantics are identical to the official API.
Generate video from a text prompt. The wan2.7/wan2.6 series support automatic audio dubbing (or pass in an audio file for audio-visual sync) and multi-shot storytelling, up to 1080P and up to 15 seconds.
The official API splits Text-to-Video into two request protocols. The submission endpoint is the same; the request body parameters differ:
New protocol (wan2.7-t2v series): output specs are controlled by resolution (resolution tier) + ratio (aspect ratio);
Legacy protocol (wan2.6 and earlier models): output specs are controlled by size (exact pixels, "width*height").

Model Comparison#

ModelProtocolResolution controlDuration (seconds)AudioMulti-shotPrompt limit
wan2.7-t2v, wan2.7-t2v-2026-06-12Newresolution + ratio (720P/1080P)Integer 2–15, default 5With audio (auto dubbing or audio_url)Supported, controlled via natural language in prompt5000 characters
wan2.6-t2vLegacysize (720P/1080P tiers)Integer 2–15, default 5With audio (auto dubbing or audio_url)Supported, requires shot_type="multi"1500 characters
wan2.6-t2v-usLegacysize (720P/1080P tiers)5 or 10, default 5Same as wan2.6Same as wan2.61500 characters
wan2.5-t2v-previewLegacysize (480P/720P/1080P tiers)5 or 10, default 5With audio (auto dubbing or audio_url)Not supported1500 characters
wan2.2-t2v-plusLegacysize (480P/1080P tiers)Fixed 5, must not be passedSilentNot supported800 characters
wanx2.1-t2v-turboLegacysize (480P/720P tiers)Fixed 5, must not be passedSilentNot supported800 characters
wanx2.1-t2v-plusLegacysize (720P tier only)Fixed 5, must not be passedSilentNot supported800 characters

Create Task#

POST https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis

Request Headers#

ParameterTypeRequiredDefaultDescription
Content-TypestringYesapplication/jsonData exchange format
AuthorizationstringYesAuthentication, Bearer {API_KEY}
The X-DashScope-Async: enable header required by the official API is added automatically by the platform; you do not need to include it.

Request Body — New Protocol (wan2.7 Series)#

ParameterTypeRequiredDefaultDescription
modelstringYeswan2.7-t2v, wan2.7-t2v-2026-06-12
input.promptstringYesText prompt, Chinese or English, ≤5000 characters; excess is automatically truncated
input.negative_promptstringNoNegative prompt describing content you do not want, ≤500 characters
input.audio_urlstringNoAuto dubbing if omittedAudio file URL (publicly accessible). Format wav/mp3, duration 2–30 seconds, ≤15MB; audio longer than the video duration is truncated, and if shorter, the trailing segment is silent
parameters.resolutionstringNo1080PResolution tier: 720P, 1080P; directly affects cost
parameters.ratiostringNo16:9Aspect ratio: 16:9, 9:16, 1:1, 4:3, 3:4
parameters.durationintegerNo5Video duration in seconds, integer in [2, 15]; directly affects cost
parameters.prompt_extendbooleanNotrueIntelligent prompt rewriting; noticeably improves short prompts but increases processing time
parameters.watermarkbooleanNofalseWhen true, adds an "AI-generated" watermark at the bottom-right corner of the video
parameters.seedintegerNoRandomRandom seed, range [0, 2147483647]; identical seeds do not guarantee identical results
wan2.7 does not support the size and shot_type fields (they take no effect). Control single/multi-shot directly in the prompt with natural language, e.g. "generate a multi-shot video" or timestamped shot breakdowns like "Shot 1 [0–3s]…"; if unspecified, the model decides based on the prompt semantics.

Request Body — Legacy Protocol (wan2.6 / wan2.5 / wan2.2 / wanx2.1)#

ParameterTypeRequiredDefaultDescription
modelstringYesSee the model comparison table above, e.g. wan2.6-t2v
input.promptstringYesText prompt; wan2.6/wan2.5 series ≤1500 characters, wan2.2/wanx2.1 series ≤800 characters, excess is automatically truncated
input.negative_promptstringNoNegative prompt, ≤500 characters
input.audio_urlstringNoAuto dubbing if omittedSupported by wan2.6/wan2.5 series only. Format wav/mp3, duration 3–30 seconds, ≤15MB
parameters.sizestringNoModel-dependent (see table below)Output resolution; must be an exact pixel value such as 1280*720 — not 1:1 or 480P; valid values depend on the model and directly affect cost
parameters.durationintegerNo5Video duration in seconds. wan2.6-t2v: [2, 15]; wan2.6-t2v-us / wan2.5-t2v-preview: 5 or 10; wan2.2/wanx2.1 series are fixed at 5 seconds and must not be passed
parameters.shot_typestringNosinglewan2.6 only. single (single shot) / multi (multi-shot); only takes effect when prompt_extend=true, and takes precedence over shot descriptions in the prompt
parameters.prompt_extendbooleanNotrueSame as new protocol
parameters.watermarkbooleanNofalseSame as new protocol
parameters.seedintegerNoRandomSame as new protocol

Resolution Tiers and Pixel Mapping#

New protocol (wan2.7): resolution × ratio → output pixels
Tier16:99:161:14:33:4
720P1280*720720*1280960*9601104*832832*1104
1080P1920*10801080*19201440*14401648*12481248*1648
Legacy protocol: valid size values
TierValid size values (width*height: aspect ratio)
480P832*480 (16:9), 480*832 (9:16), 624*624 (1:1)
720P1280*720 (16:9), 720*1280 (9:16), 960*960 (1:1), 1088*832 (4:3), 832*1088 (3:4)
1080P1920*1080 (16:9), 1080*1920 (9:16), 1440*1440 (1:1), 1632*1248 (4:3), 1248*1632 (3:4)
Available tiers and defaults per model
ModelAvailable tiersDefault
wan2.7-t2v(-2026-06-12)720P, 1080P (resolution)1080P / 16:9
wan2.6-t2v, wan2.6-t2v-usAll sizes in the 720P and 1080P tiers1920*1080
wan2.5-t2v-previewAll sizes in the 480P, 720P and 1080P tiers1920*1080
wan2.2-t2v-plusAll sizes in the 480P and 1080P tiers1920*1080
wanx2.1-t2v-turboAll sizes in the 480P and 720P tiers1280*720
wanx2.1-t2v-plusAll sizes in the 720P tier only1280*720
Note: in the legacy 720P/1080P tiers, the 4:3 and 3:4 pixel values (1088*832, 1632*1248, etc.) differ from the new-protocol wan2.7 values (1104*832, 1648*1248, etc.); do not mix them up.

Request Example (wan2.7-t2v, New Protocol)#

curl --request POST \
  --url 'https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis' \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "wan2.7-t2v",
    "input": {
      "prompt": "At sunset, waves gently lap against the rocks while a sailboat drifts by in the distance"
    },
    "parameters": {
      "resolution": "720P",
      "ratio": "16:9",
      "duration": 5
    }
  }'

Request Example (wan2.6-t2v, Legacy Protocol)#

curl --request POST \
  --url 'https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis' \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "wan2.6-t2v",
    "input": {
      "prompt": "On a spring day, cherry blossoms drift down over a classical courtyard as a gentle breeze blows"
    },
    "parameters": {
      "size": "1280*720",
      "duration": 5
    }
  }'

Response Example#

{
  "output": {
    "task_status": "PENDING",
    "task_id": "0385dc79-5ff8-4d82-bcb6-xxxxxx"
  },
  "request_id": "4909100c-7b5a-9f92-bfe5-xxxxxx"
}

Response Fields#

FieldTypeDescription
output.task_idstringTask ID used for polling; valid for 24 hours
output.task_statusstringTask status; PENDING after successful creation
request_idstringUnique request identifier, useful for troubleshooting
code / messagestringError code and details, returned only when creation fails

Notes#

Fixed-duration models: wan2.2-t2v-plus, wanx2.1-t2v-turbo, and wanx2.1-t2v-plus always generate 5-second videos. Do not pass duration — doing so returns an error (duration customization is not supported).
Prompt over-length is auto-truncated: see the model comparison table for per-model prompt limits (5000/1500/800 characters); the excess is truncated silently without an error. negative_prompt is limited to 500 characters.
Multi-shot: wan2.7 removed the shot_type parameter — describe the shot structure directly in the prompt with natural language or timestamped shot breakdowns; wan2.6 requires shot_type: "multi" with prompt_extend: true.
Audio requirements: audio_url is supported by the wan2.7/wan2.6/wan2.5 series only; wav/mp3 format, duration 2–30 seconds for wan2.7 and 3–30 seconds for legacy models, size ≤15MB. If omitted, the model dubs the video automatically (wan2.2/wanx2.1 series generate silent videos only).
Billing: charged by the number of seconds of successfully generated video; resolution and duration directly affect cost. Inputs are not billed, and failed tasks are not billed.
Result validity: both the task_id and the generated video URL are valid for 24 hours; download or archive the video promptly.

Query Task#

After creating a task, poll the task status and results with the returned output.task_id:
GET https://platform.dataeyes.ai/ali/api/v1/tasks/{task_id}
See 09-Query-Task for details.
Previous
00-Overview.md
Next
02-Image-to-Video