DataEyesAI
Official SiteConsoleDocs Home
Getting StartedDeveloper ToolsAI Models APITerms & Policies
Official SiteConsoleDocs Home
Getting StartedDeveloper ToolsAI Models APITerms & Policies
  1. Wan Video Generation
  • Documentation
    • Getting Started
      • Overview
      • Console (Getting Started)
      • API Key
      • Base URL
    • Developer Tool Integration
      • OpenClaw
      • Claude Code
      • Codex
      • Gemini CLI
      • Grok CLI
      • Other Tools
    • AI Models API
      • OpenAI format (supports major original models)
        • Chat (Response)
          • Create Network Search
          • Create Model Response GPT-5 Enable Thinking
          • Create Function Call
          • Create Model Response
          • Create Model Response (Streaming Return)
          • Create Model Response (Control Thinking Length)
        • ChatGPT Interface
          • Audio
            • Audio to text gpt-4o-transcribe
            • GPT-4o-audio
            • Audio to text whisper-1
            • Audio to text gpt-4o-transcribe
            • Create voice gpt-4o-mini-tts
          • Chat
            • Create chat-based image recognition (non-streaming)
            • Create chat-based image recognition (streaming)
            • Create chat-based image recognition (streaming) best64
            • Official N test
            • Create structured output
            • Control the effort level of the inference model
            • Create chat function call
            • deepseek-ocr recognition
            • Create chat completion (non-stream)
          • Completions
            • ChatGPT automatic completion
            • Create completion
        • Image
          • Edit image
          • Create chat completion (streaming)
          • Create chat completion (qwen-mt-turbo)
          • Create chat completion with deepseek v3.1 level of reasoning (streaming)
        • Audio
          • Speech recognition
          • Speech synthesis
          • Official Function Calling invocation
          • Create chat-generated images (non-streaming)
        • Embedding
          • Text embeddings
      • Anthropic format
        • Chat
        • Chat(prompt cache)
        • Streaming response
        • Chat (deep reasoning)
        • Tool invocation (function call)
        • Analyze image
      • Google Gemini interface
        • Native format
          • Text-to-image + control over aspect ratio + clarity
          • Generate image
          • Text generation
          • Text generation - stream
          • Text generation + reasoning - stream
          • Image generation
          • Formatted output
          • Function call
          • Document understanding
          • URL context [native format]
          • Code execution
          • Video understanding
          • URL context
          • Video understanding - url [native format]
          • Imagen 4
          • Audio understanding
          • Embeddings
          • Chat
          • Edit image
        • Image-to-image Base64 request method
          • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity
          • Image editing
          • Single image gemini-3-pro-image-preview, controlling aspect ratio and clarity.
          • Image generation( gemini-2.5-flash-image)
          • Image generation gemini-2.5-flash-image, controlling aspect ratio.
          • Image understanding
        • Image-to-image URL request returns URL request format OpenAI
          • Single image generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
          • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
          • Image understanding
      • NanoBanana
        • OpenAI request
          • Edit image
          • OpenAI image format
        • Gemini request
          • Generate image
          • Edit image
      • Midjourney format
        • Midjourney API Reference
        • Task query interface
        • Upload image
        • Get seed (Seed)
        • Submit Imagine task
        • Query tasks based on ID list
        • FaceSwap
        • Execute Action operation
        • /mj/submit/blend
        • Submit Describe task
        • Submit Modal
        • Refresh link
        • Edit image
        • Query task status by task ID
        • Get the seed of the task image
      • Doubao - Painting
        • doubao-seededit-3-0-i2i-250628
        • doubao-seedream-4-0-250828 - text-to-image
        • doubao-seedream-4-0-250828 - image-to-image
        • doubao-seedream-4-0-250828 - multi-image generation
      • Rerank Reordering Model
        • Rerank
      • Video Model
        • Grok Video Generation
          • 00-Overview
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Reference-to-Video
          • 04-Video-Editing
          • 05-Video-Extension
        • Seedance Video Generation
          • 00-Overview
          • 01-Create-Video-Generation-Task
          • 02-Query-Video-Generation-Task
          • 03-Query-Video-Generation-Task-List
          • 04-Cancel-or-Delete-Task
          • Seedance Private Asset Library API Documentation
        • MiniMax-H3 Video Generation
          • 00-Overview
          • 01-Create-Video-Generation
          • 02-Create-Video-Regeneration
          • 03-Create-H3-Context-IR
          • 04-Query-Task
          • 05-List-Tasks
          • 06-Cancel-or-Delete-Task
        • Hailuo Video Generation
          • 00-Overview
          • 01-Text-to-Video-T2V
          • 02-Image-to-Video-I2V
          • 03-First-Last-Frame-FL2V
          • 04-Subject-Reference-S2V
          • 05-Query-Task-Status
          • 06-Video-Download
          • 99-Appendix-Camera-Movement-and-Webhooks
        • Jimeng Video Generation
          • 00-Overview
          • 01-3.0-Pro-Video-Generation
          • 02-720P-Text-to-Video
          • 03-720P-Image-to-Video-First-Frame
          • 04-720P-Image-to-Video-Start-End-Frame
          • 05-720P-Image-to-Video-Camera
          • 06-1080P-Text-to-Video
          • 07-1080P-Image-to-Video-First-Frame
          • 08-1080P-Image-to-Video-Start-End-Frame
          • 09-Error-Codes
        • Kling AI Video Generation
          • 00-Overview
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Omni-Video
          • 04-Multi-Image-to-Video
          • 05-Motion-Control
          • 06-Multi-Elements
          • 07-Video-Extension
          • 08-Lip-Sync
          • 09-Avatar
          • 10-Text-to-Audio
          • 11-Video-to-Audio
          • 12-TTS
          • 13-Custom-Voices
          • 14-Image-Recognition
          • 15-Element-Management
          • 16-Video-Effects
        • Vidu Video Generation
          • 00-Overview
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Reference-to-Video
          • 04-Start-End-Frame
          • 05-Multi-Frame
          • 06-Scene-Template
          • 07-Template-Story
          • 08-Query-Tasks
        • HappyHorse
          • HappyHorse Text-to-Video
          • HappyHorse Image-to-Video (First Frame)
          • HappyHorse Reference-to-Video
          • HappyHorse Video Editing
        • Wan Video Generation
          • 00-Overview.md
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Reference-to-Video
          • 04-Video-Editing
          • 05-First-Last-Frame-to-Video
          • 06-Motion-Transfer-and-Character-Swap
          • 07-Digital-Human-Video
          • 08-VACE-Video-Editing
          • 09-Query-Task
      • Audio API
        • Audio API
        • Gemini TTS API
        • Google DeepMind Lyria API
        • Elevenlabs Speech to Text API Reference
        • Text-to-Music Suno
          • Task Submission
            • Generate Song (Inspiration Mode)
            • Generate Song (Custom Mode)
            • Generate Song (Continuation Mode)
            • Generate Song (Singer Style)
            • Generate Song (Secondary Creation from Uploaded Song)
            • Generate Song (Song Stitching)
            • Generate Lyrics
            • Song Stitching
          • Query Interface
            • Batch Retrieve Tasks
            • Query Single Task
    • Search / Reader Product
      • Web Reader API​​
        • Web Reader API
        • Web Reader API(HK)
      • Web Search API​​
        • Modal Card API
          • Weather
            • All City ID
            • Weather Query API
        • Web Search API
        • Video Search api
        • Trending Search API
      • Document OCR Parsing API
        • fiel upload
        • URL Parsing
    • Advanced & System API
      • Data Updates
      • System interface
        • API Key & Quota Query API
        • API Key Management API
      • API Reference​​
        • Error Codes
        • HTTP Notes
      • List models
        • Models
  • Terms & Policies
    • DataEyesAI API Terms of Service
    • DataEyesAI Legal Notice and Privacy Policy
    • DataEyesAI Paid Services Agreement
    • Automatic Renewal Service Rules
  1. Wan Video Generation

01-Text-to-Video

Text-to-Video#

Doc version:v1.0.0 | Last updated:2026-07-22
This platform fully supports the official Tongyi Wanxiang (Wan) video generation API. Requests and responses are transparently proxied; parameter semantics are identical to the official API.
Generate video from a text prompt. The wan2.7/wan2.6 series support automatic audio dubbing (or pass in an audio file for audio-visual sync) and multi-shot storytelling, up to 1080P and up to 15 seconds.
The official API splits Text-to-Video into two request protocols. The submission endpoint is the same; the request body parameters differ:
New protocol (wan2.7-t2v series): output specs are controlled by resolution (resolution tier) + ratio (aspect ratio);
Legacy protocol (wan2.6 and earlier models): output specs are controlled by size (exact pixels, "width*height").

Model Comparison#

ModelProtocolResolution controlDuration (seconds)AudioMulti-shotPrompt limit
wan2.7-t2v, wan2.7-t2v-2026-06-12Newresolution + ratio (720P/1080P)Integer 2–15, default 5With audio (auto dubbing or audio_url)Supported, controlled via natural language in prompt5000 characters
wan2.6-t2vLegacysize (720P/1080P tiers)Integer 2–15, default 5With audio (auto dubbing or audio_url)Supported, requires shot_type="multi"1500 characters
wan2.6-t2v-usLegacysize (720P/1080P tiers)5 or 10, default 5Same as wan2.6Same as wan2.61500 characters
wan2.5-t2v-previewLegacysize (480P/720P/1080P tiers)5 or 10, default 5With audio (auto dubbing or audio_url)Not supported1500 characters
wan2.2-t2v-plusLegacysize (480P/1080P tiers)Fixed 5, must not be passedSilentNot supported800 characters
wanx2.1-t2v-turboLegacysize (480P/720P tiers)Fixed 5, must not be passedSilentNot supported800 characters
wanx2.1-t2v-plusLegacysize (720P tier only)Fixed 5, must not be passedSilentNot supported800 characters

Create Task#

POST https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis

Request Headers#

ParameterTypeRequiredDefaultDescription
Content-TypestringYesapplication/jsonData exchange format
AuthorizationstringYesAuthentication, Bearer {API_KEY}
The X-DashScope-Async: enable header required by the official API is added automatically by the platform; you do not need to include it.

Request Body — New Protocol (wan2.7 Series)#

ParameterTypeRequiredDefaultDescription
modelstringYeswan2.7-t2v, wan2.7-t2v-2026-06-12
input.promptstringYesText prompt, Chinese or English, ≤5000 characters; excess is automatically truncated
input.negative_promptstringNoNegative prompt describing content you do not want, ≤500 characters
input.audio_urlstringNoAuto dubbing if omittedAudio file URL (publicly accessible). Format wav/mp3, duration 2–30 seconds, ≤15MB; audio longer than the video duration is truncated, and if shorter, the trailing segment is silent
parameters.resolutionstringNo1080PResolution tier: 720P, 1080P; directly affects cost
parameters.ratiostringNo16:9Aspect ratio: 16:9, 9:16, 1:1, 4:3, 3:4
parameters.durationintegerNo5Video duration in seconds, integer in [2, 15]; directly affects cost
parameters.prompt_extendbooleanNotrueIntelligent prompt rewriting; noticeably improves short prompts but increases processing time
parameters.watermarkbooleanNofalseWhen true, adds an "AI-generated" watermark at the bottom-right corner of the video
parameters.seedintegerNoRandomRandom seed, range [0, 2147483647]; identical seeds do not guarantee identical results
wan2.7 does not support the size and shot_type fields (they take no effect). Control single/multi-shot directly in the prompt with natural language, e.g. "generate a multi-shot video" or timestamped shot breakdowns like "Shot 1 [0–3s]…"; if unspecified, the model decides based on the prompt semantics.

Request Body — Legacy Protocol (wan2.6 / wan2.5 / wan2.2 / wanx2.1)#

ParameterTypeRequiredDefaultDescription
modelstringYesSee the model comparison table above, e.g. wan2.6-t2v
input.promptstringYesText prompt; wan2.6/wan2.5 series ≤1500 characters, wan2.2/wanx2.1 series ≤800 characters, excess is automatically truncated
input.negative_promptstringNoNegative prompt, ≤500 characters
input.audio_urlstringNoAuto dubbing if omittedSupported by wan2.6/wan2.5 series only. Format wav/mp3, duration 3–30 seconds, ≤15MB
parameters.sizestringNoModel-dependent (see table below)Output resolution; must be an exact pixel value such as 1280*720 — not 1:1 or 480P; valid values depend on the model and directly affect cost
parameters.durationintegerNo5Video duration in seconds. wan2.6-t2v: [2, 15]; wan2.6-t2v-us / wan2.5-t2v-preview: 5 or 10; wan2.2/wanx2.1 series are fixed at 5 seconds and must not be passed
parameters.shot_typestringNosinglewan2.6 only. single (single shot) / multi (multi-shot); only takes effect when prompt_extend=true, and takes precedence over shot descriptions in the prompt
parameters.prompt_extendbooleanNotrueSame as new protocol
parameters.watermarkbooleanNofalseSame as new protocol
parameters.seedintegerNoRandomSame as new protocol

Resolution Tiers and Pixel Mapping#

New protocol (wan2.7): resolution × ratio → output pixels
Tier16:99:161:14:33:4
720P1280*720720*1280960*9601104*832832*1104
1080P1920*10801080*19201440*14401648*12481248*1648
Legacy protocol: valid size values
TierValid size values (width*height: aspect ratio)
480P832*480 (16:9), 480*832 (9:16), 624*624 (1:1)
720P1280*720 (16:9), 720*1280 (9:16), 960*960 (1:1), 1088*832 (4:3), 832*1088 (3:4)
1080P1920*1080 (16:9), 1080*1920 (9:16), 1440*1440 (1:1), 1632*1248 (4:3), 1248*1632 (3:4)
Available tiers and defaults per model
ModelAvailable tiersDefault
wan2.7-t2v(-2026-06-12)720P, 1080P (resolution)1080P / 16:9
wan2.6-t2v, wan2.6-t2v-usAll sizes in the 720P and 1080P tiers1920*1080
wan2.5-t2v-previewAll sizes in the 480P, 720P and 1080P tiers1920*1080
wan2.2-t2v-plusAll sizes in the 480P and 1080P tiers1920*1080
wanx2.1-t2v-turboAll sizes in the 480P and 720P tiers1280*720
wanx2.1-t2v-plusAll sizes in the 720P tier only1280*720
Note: in the legacy 720P/1080P tiers, the 4:3 and 3:4 pixel values (1088*832, 1632*1248, etc.) differ from the new-protocol wan2.7 values (1104*832, 1648*1248, etc.); do not mix them up.

Request Example (wan2.7-t2v, New Protocol)#

curl --request POST \
  --url 'https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis' \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "wan2.7-t2v",
    "input": {
      "prompt": "At sunset, waves gently lap against the rocks while a sailboat drifts by in the distance"
    },
    "parameters": {
      "resolution": "720P",
      "ratio": "16:9",
      "duration": 5
    }
  }'

Request Example (wan2.6-t2v, Legacy Protocol)#

curl --request POST \
  --url 'https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis' \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "wan2.6-t2v",
    "input": {
      "prompt": "On a spring day, cherry blossoms drift down over a classical courtyard as a gentle breeze blows"
    },
    "parameters": {
      "size": "1280*720",
      "duration": 5
    }
  }'

Response Example#

{
  "output": {
    "task_status": "PENDING",
    "task_id": "0385dc79-5ff8-4d82-bcb6-xxxxxx"
  },
  "request_id": "4909100c-7b5a-9f92-bfe5-xxxxxx"
}

Response Fields#

FieldTypeDescription
output.task_idstringTask ID used for polling; valid for 24 hours
output.task_statusstringTask status; PENDING after successful creation
request_idstringUnique request identifier, useful for troubleshooting
code / messagestringError code and details, returned only when creation fails

Notes#

Fixed-duration models: wan2.2-t2v-plus, wanx2.1-t2v-turbo, and wanx2.1-t2v-plus always generate 5-second videos. Do not pass duration — doing so returns an error (duration customization is not supported).
Prompt over-length is auto-truncated: see the model comparison table for per-model prompt limits (5000/1500/800 characters); the excess is truncated silently without an error. negative_prompt is limited to 500 characters.
Multi-shot: wan2.7 removed the shot_type parameter — describe the shot structure directly in the prompt with natural language or timestamped shot breakdowns; wan2.6 requires shot_type: "multi" with prompt_extend: true.
Audio requirements: audio_url is supported by the wan2.7/wan2.6/wan2.5 series only; wav/mp3 format, duration 2–30 seconds for wan2.7 and 3–30 seconds for legacy models, size ≤15MB. If omitted, the model dubs the video automatically (wan2.2/wanx2.1 series generate silent videos only).
Billing: charged by the number of seconds of successfully generated video; resolution and duration directly affect cost. Inputs are not billed, and failed tasks are not billed.
Result validity: both the task_id and the generated video URL are valid for 24 hours; download or archive the video promptly.

Query Task#

After creating a task, poll the task status and results with the returned output.task_id:
GET https://platform.dataeyes.ai/ali/api/v1/tasks/{task_id}
See 09-Query-Task for details.
Previous
00-Overview.md
Next
02-Image-to-Video