DataEyesAI
Official SiteConsoleDocs HomeGetting StartedDeveloper ToolsAI Models API
Official SiteConsoleDocs HomeGetting StartedDeveloper ToolsAI Models API
  1. Wan Video Generation
  • OpenAI format (supports major original models)
    • Chat (Response)
      • Create Network Search
      • Create Model Response GPT-5 Enable Thinking
      • Create Function Call
      • Create Model Response
      • Create Model Response (Streaming Return)
      • Create Model Response (Control Thinking Length)
    • ChatGPT Interface
      • Audio
        • Audio to text gpt-4o-transcribe
        • GPT-4o-audio
        • Audio to text whisper-1
        • Audio to text gpt-4o-transcribe
        • Create voice gpt-4o-mini-tts
      • Chat
        • Create chat-based image recognition (non-streaming)
        • Create chat-based image recognition (streaming)
        • Create chat-based image recognition (streaming) best64
        • Official N test
        • Create structured output
        • Control the effort level of the inference model
        • Create chat function call
        • deepseek-ocr recognition
        • Create chat completion (non-stream)
      • Completions
        • ChatGPT automatic completion
        • Create completion
    • Image
      • Edit image
      • Create chat completion (streaming)
      • Create chat completion (qwen-mt-turbo)
      • Create chat completion with deepseek v3.1 level of reasoning (streaming)
    • Audio
      • Speech recognition
      • Speech synthesis
      • Official Function Calling invocation
      • Create chat-generated images (non-streaming)
    • Embedding
      • Text embeddings
  • Anthropic format
    • Chat
    • Chat(prompt cache)
    • Streaming response
    • Chat (deep reasoning)
    • Tool invocation (function call)
    • Analyze image
  • Google Gemini interface
    • Native format
      • Text-to-image + control over aspect ratio + clarity
      • Generate image
      • Text generation
      • Text generation - stream
      • Text generation + reasoning - stream
      • Image generation
      • Formatted output
      • Function call
      • Document understanding
      • URL context [native format]
      • Code execution
      • Video understanding
      • URL context
      • Video understanding - url [native format]
      • Imagen 4
      • Audio understanding
      • Embeddings
      • Chat
      • Edit image
    • Image-to-image Base64 request method
      • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity
      • Image editing
      • Single image gemini-3-pro-image-preview, controlling aspect ratio and clarity.
      • Image generation( gemini-2.5-flash-image)
      • Image generation gemini-2.5-flash-image, controlling aspect ratio.
      • Image understanding
    • Image-to-image URL request returns URL request format OpenAI
      • Single image generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
      • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
      • Image understanding
  • NanoBanana
    • OpenAI request
      • Edit image
      • OpenAI image format
    • Gemini request
      • Generate image
      • Edit image
  • Midjourney format
    • Midjourney API Reference
    • Task query interface
    • Upload image
    • Get seed (Seed)
    • Submit Imagine task
    • Query tasks based on ID list
    • FaceSwap
    • Execute Action operation
    • /mj/submit/blend
    • Submit Describe task
    • Submit Modal
    • Refresh link
    • Edit image
    • Query task status by task ID
    • Get the seed of the task image
  • Doubao - Painting
    • doubao-seededit-3-0-i2i-250628
    • doubao-seedream-4-0-250828 - text-to-image
    • doubao-seedream-4-0-250828 - image-to-image
    • doubao-seedream-4-0-250828 - multi-image generation
  • Rerank Reordering Model
    • Rerank
  • Video Model
    • Grok Video Generation
      • 00-Overview
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Reference-to-Video
      • 04-Video-Editing
      • 05-Video-Extension
    • Seedance Video Generation
      • 00-Overview
      • 01-Create-Video-Generation-Task
      • 02-Query-Video-Generation-Task
      • 03-Query-Video-Generation-Task-List
      • 04-Cancel-or-Delete-Task
      • Seedance Private Asset Library API Documentation
    • MiniMax-H3 Video Generation
      • 00-Overview
      • 01-Create-Video-Generation
      • 02-Create-Video-Regeneration
      • 03-Create-H3-Context-IR
      • 04-Query-Task
      • 05-List-Tasks
      • 06-Cancel-or-Delete-Task
    • Hailuo Video Generation
      • 00-Overview
      • 01-Text-to-Video-T2V
      • 02-Image-to-Video-I2V
      • 03-First-Last-Frame-FL2V
      • 04-Subject-Reference-S2V
      • 05-Query-Task-Status
      • 06-Video-Download
      • 99-Appendix-Camera-Movement-and-Webhooks
    • Jimeng Video Generation
      • 00-Overview
      • 01-3.0-Pro-Video-Generation
      • 02-720P-Text-to-Video
      • 03-720P-Image-to-Video-First-Frame
      • 04-720P-Image-to-Video-Start-End-Frame
      • 05-720P-Image-to-Video-Camera
      • 06-1080P-Text-to-Video
      • 07-1080P-Image-to-Video-First-Frame
      • 08-1080P-Image-to-Video-Start-End-Frame
      • 09-Error-Codes
    • Kling AI Video Generation
      • 00-Overview
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Omni-Video
      • 04-Multi-Image-to-Video
      • 05-Motion-Control
      • 06-Multi-Elements
      • 07-Video-Extension
      • 08-Lip-Sync
      • 09-Avatar
      • 10-Text-to-Audio
      • 11-Video-to-Audio
      • 12-TTS
      • 13-Custom-Voices
      • 14-Image-Recognition
      • 15-Element-Management
      • 16-Video-Effects
    • Vidu Video Generation
      • 00-Overview
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Reference-to-Video
      • 04-Start-End-Frame
      • 05-Multi-Frame
      • 06-Scene-Template
      • 07-Template-Story
      • 08-Query-Tasks
    • HappyHorse
      • HappyHorse Text-to-Video
      • HappyHorse Image-to-Video (First Frame)
      • HappyHorse Reference-to-Video
      • HappyHorse Video Editing
    • Wan Video Generation
      • 00-Overview.md
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Reference-to-Video
      • 04-Video-Editing
      • 05-First-Last-Frame-to-Video
      • 06-Motion-Transfer-and-Character-Swap
      • 07-Digital-Human-Video
      • 08-VACE-Video-Editing
      • 09-Query-Task
  • Audio API
    • Audio API
    • Gemini TTS API
    • Google DeepMind Lyria API
    • Elevenlabs Speech to Text API Reference
    • Text-to-Music Suno
      • Task Submission
        • Generate Song (Inspiration Mode)
        • Generate Song (Custom Mode)
        • Generate Song (Continuation Mode)
        • Generate Song (Singer Style)
        • Generate Song (Secondary Creation from Uploaded Song)
        • Generate Song (Song Stitching)
        • Generate Lyrics
        • Song Stitching
      • Query Interface
        • Batch Retrieve Tasks
        • Query Single Task
  1. Wan Video Generation

01-Text-to-Video

Text-to-Video#

Doc version:v1.0.0 | Last updated:2026-07-22
This platform fully supports the official Tongyi Wanxiang (Wan) video generation API. Requests and responses are transparently proxied; parameter semantics are identical to the official API.
Generate video from a text prompt. The wan2.7/wan2.6 series support automatic audio dubbing (or pass in an audio file for audio-visual sync) and multi-shot storytelling, up to 1080P and up to 15 seconds.
The official API splits Text-to-Video into two request protocols. The submission endpoint is the same; the request body parameters differ:
New protocol (wan2.7-t2v series): output specs are controlled by resolution (resolution tier) + ratio (aspect ratio);
Legacy protocol (wan2.6 and earlier models): output specs are controlled by size (exact pixels, "width*height").

Model Comparison#

ModelProtocolResolution controlDuration (seconds)AudioMulti-shotPrompt limit
wan2.7-t2v, wan2.7-t2v-2026-06-12Newresolution + ratio (720P/1080P)Integer 2–15, default 5With audio (auto dubbing or audio_url)Supported, controlled via natural language in prompt5000 characters
wan2.6-t2vLegacysize (720P/1080P tiers)Integer 2–15, default 5With audio (auto dubbing or audio_url)Supported, requires shot_type="multi"1500 characters
wan2.6-t2v-usLegacysize (720P/1080P tiers)5 or 10, default 5Same as wan2.6Same as wan2.61500 characters
wan2.5-t2v-previewLegacysize (480P/720P/1080P tiers)5 or 10, default 5With audio (auto dubbing or audio_url)Not supported1500 characters
wan2.2-t2v-plusLegacysize (480P/1080P tiers)Fixed 5, must not be passedSilentNot supported800 characters
wanx2.1-t2v-turboLegacysize (480P/720P tiers)Fixed 5, must not be passedSilentNot supported800 characters
wanx2.1-t2v-plusLegacysize (720P tier only)Fixed 5, must not be passedSilentNot supported800 characters

Create Task#

POST https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis

Request Headers#

ParameterTypeRequiredDefaultDescription
Content-TypestringYesapplication/jsonData exchange format
AuthorizationstringYesAuthentication, Bearer {API_KEY}
The X-DashScope-Async: enable header required by the official API is added automatically by the platform; you do not need to include it.

Request Body — New Protocol (wan2.7 Series)#

ParameterTypeRequiredDefaultDescription
modelstringYeswan2.7-t2v, wan2.7-t2v-2026-06-12
input.promptstringYesText prompt, Chinese or English, ≤5000 characters; excess is automatically truncated
input.negative_promptstringNoNegative prompt describing content you do not want, ≤500 characters
input.audio_urlstringNoAuto dubbing if omittedAudio file URL (publicly accessible). Format wav/mp3, duration 2–30 seconds, ≤15MB; audio longer than the video duration is truncated, and if shorter, the trailing segment is silent
parameters.resolutionstringNo1080PResolution tier: 720P, 1080P; directly affects cost
parameters.ratiostringNo16:9Aspect ratio: 16:9, 9:16, 1:1, 4:3, 3:4
parameters.durationintegerNo5Video duration in seconds, integer in [2, 15]; directly affects cost
parameters.prompt_extendbooleanNotrueIntelligent prompt rewriting; noticeably improves short prompts but increases processing time
parameters.watermarkbooleanNofalseWhen true, adds an "AI-generated" watermark at the bottom-right corner of the video
parameters.seedintegerNoRandomRandom seed, range [0, 2147483647]; identical seeds do not guarantee identical results
wan2.7 does not support the size and shot_type fields (they take no effect). Control single/multi-shot directly in the prompt with natural language, e.g. "generate a multi-shot video" or timestamped shot breakdowns like "Shot 1 [0–3s]…"; if unspecified, the model decides based on the prompt semantics.

Request Body — Legacy Protocol (wan2.6 / wan2.5 / wan2.2 / wanx2.1)#

ParameterTypeRequiredDefaultDescription
modelstringYesSee the model comparison table above, e.g. wan2.6-t2v
input.promptstringYesText prompt; wan2.6/wan2.5 series ≤1500 characters, wan2.2/wanx2.1 series ≤800 characters, excess is automatically truncated
input.negative_promptstringNoNegative prompt, ≤500 characters
input.audio_urlstringNoAuto dubbing if omittedSupported by wan2.6/wan2.5 series only. Format wav/mp3, duration 3–30 seconds, ≤15MB
parameters.sizestringNoModel-dependent (see table below)Output resolution; must be an exact pixel value such as 1280*720 — not 1:1 or 480P; valid values depend on the model and directly affect cost
parameters.durationintegerNo5Video duration in seconds. wan2.6-t2v: [2, 15]; wan2.6-t2v-us / wan2.5-t2v-preview: 5 or 10; wan2.2/wanx2.1 series are fixed at 5 seconds and must not be passed
parameters.shot_typestringNosinglewan2.6 only. single (single shot) / multi (multi-shot); only takes effect when prompt_extend=true, and takes precedence over shot descriptions in the prompt
parameters.prompt_extendbooleanNotrueSame as new protocol
parameters.watermarkbooleanNofalseSame as new protocol
parameters.seedintegerNoRandomSame as new protocol

Resolution Tiers and Pixel Mapping#

New protocol (wan2.7): resolution × ratio → output pixels
Tier16:99:161:14:33:4
720P1280*720720*1280960*9601104*832832*1104
1080P1920*10801080*19201440*14401648*12481248*1648
Legacy protocol: valid size values
TierValid size values (width*height: aspect ratio)
480P832*480 (16:9), 480*832 (9:16), 624*624 (1:1)
720P1280*720 (16:9), 720*1280 (9:16), 960*960 (1:1), 1088*832 (4:3), 832*1088 (3:4)
1080P1920*1080 (16:9), 1080*1920 (9:16), 1440*1440 (1:1), 1632*1248 (4:3), 1248*1632 (3:4)
Available tiers and defaults per model
ModelAvailable tiersDefault
wan2.7-t2v(-2026-06-12)720P, 1080P (resolution)1080P / 16:9
wan2.6-t2v, wan2.6-t2v-usAll sizes in the 720P and 1080P tiers1920*1080
wan2.5-t2v-previewAll sizes in the 480P, 720P and 1080P tiers1920*1080
wan2.2-t2v-plusAll sizes in the 480P and 1080P tiers1920*1080
wanx2.1-t2v-turboAll sizes in the 480P and 720P tiers1280*720
wanx2.1-t2v-plusAll sizes in the 720P tier only1280*720
Note: in the legacy 720P/1080P tiers, the 4:3 and 3:4 pixel values (1088*832, 1632*1248, etc.) differ from the new-protocol wan2.7 values (1104*832, 1648*1248, etc.); do not mix them up.

Request Example (wan2.7-t2v, New Protocol)#

curl --request POST \
  --url 'https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis' \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "wan2.7-t2v",
    "input": {
      "prompt": "At sunset, waves gently lap against the rocks while a sailboat drifts by in the distance"
    },
    "parameters": {
      "resolution": "720P",
      "ratio": "16:9",
      "duration": 5
    }
  }'

Request Example (wan2.6-t2v, Legacy Protocol)#

curl --request POST \
  --url 'https://platform.dataeyes.ai/ali/api/v1/services/aigc/video-generation/video-synthesis' \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "wan2.6-t2v",
    "input": {
      "prompt": "On a spring day, cherry blossoms drift down over a classical courtyard as a gentle breeze blows"
    },
    "parameters": {
      "size": "1280*720",
      "duration": 5
    }
  }'

Response Example#

{
  "output": {
    "task_status": "PENDING",
    "task_id": "0385dc79-5ff8-4d82-bcb6-xxxxxx"
  },
  "request_id": "4909100c-7b5a-9f92-bfe5-xxxxxx"
}

Response Fields#

FieldTypeDescription
output.task_idstringTask ID used for polling; valid for 24 hours
output.task_statusstringTask status; PENDING after successful creation
request_idstringUnique request identifier, useful for troubleshooting
code / messagestringError code and details, returned only when creation fails

Notes#

Fixed-duration models: wan2.2-t2v-plus, wanx2.1-t2v-turbo, and wanx2.1-t2v-plus always generate 5-second videos. Do not pass duration — doing so returns an error (duration customization is not supported).
Prompt over-length is auto-truncated: see the model comparison table for per-model prompt limits (5000/1500/800 characters); the excess is truncated silently without an error. negative_prompt is limited to 500 characters.
Multi-shot: wan2.7 removed the shot_type parameter — describe the shot structure directly in the prompt with natural language or timestamped shot breakdowns; wan2.6 requires shot_type: "multi" with prompt_extend: true.
Audio requirements: audio_url is supported by the wan2.7/wan2.6/wan2.5 series only; wav/mp3 format, duration 2–30 seconds for wan2.7 and 3–30 seconds for legacy models, size ≤15MB. If omitted, the model dubs the video automatically (wan2.2/wanx2.1 series generate silent videos only).
Billing: charged by the number of seconds of successfully generated video; resolution and duration directly affect cost. Inputs are not billed, and failed tasks are not billed.
Result validity: both the task_id and the generated video URL are valid for 24 hours; download or archive the video promptly.

Query Task#

After creating a task, poll the task status and results with the returned output.task_id:
GET https://platform.dataeyes.ai/ali/api/v1/tasks/{task_id}
See 09-Query-Task for details.
Previous
00-Overview.md
Next
02-Image-to-Video