DataEyesAI
Official SiteConsoleDocs HomeGetting StartedDeveloper ToolsAI Models API
Official SiteConsoleDocs HomeGetting StartedDeveloper ToolsAI Models API
  1. Vidu Video Generation
  • OpenAI format (supports major original models)
    • Chat (Response)
      • Create Network Search
      • Create Model Response GPT-5 Enable Thinking
      • Create Function Call
      • Create Model Response
      • Create Model Response (Streaming Return)
      • Create Model Response (Control Thinking Length)
    • ChatGPT Interface
      • Audio
        • Audio to text gpt-4o-transcribe
        • GPT-4o-audio
        • Audio to text whisper-1
        • Audio to text gpt-4o-transcribe
        • Create voice gpt-4o-mini-tts
      • Chat
        • Create chat-based image recognition (non-streaming)
        • Create chat-based image recognition (streaming)
        • Create chat-based image recognition (streaming) best64
        • Official N test
        • Create structured output
        • Control the effort level of the inference model
        • Create chat function call
        • deepseek-ocr recognition
        • Create chat completion (non-stream)
      • Completions
        • ChatGPT automatic completion
        • Create completion
    • Image
      • Edit image
      • Create chat completion (streaming)
      • Create chat completion (qwen-mt-turbo)
      • Create chat completion with deepseek v3.1 level of reasoning (streaming)
    • Audio
      • Speech recognition
      • Speech synthesis
      • Official Function Calling invocation
      • Create chat-generated images (non-streaming)
    • Embedding
      • Text embeddings
  • Anthropic format
    • Chat
    • Chat(prompt cache)
    • Streaming response
    • Chat (deep reasoning)
    • Tool invocation (function call)
    • Analyze image
  • Google Gemini interface
    • Native format
      • Text-to-image + control over aspect ratio + clarity
      • Generate image
      • Text generation
      • Text generation - stream
      • Text generation + reasoning - stream
      • Image generation
      • Formatted output
      • Function call
      • Document understanding
      • URL context [native format]
      • Code execution
      • Video understanding
      • URL context
      • Video understanding - url [native format]
      • Imagen 4
      • Audio understanding
      • Embeddings
      • Chat
      • Edit image
    • Image-to-image Base64 request method
      • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity
      • Image editing
      • Single image gemini-3-pro-image-preview, controlling aspect ratio and clarity.
      • Image generation( gemini-2.5-flash-image)
      • Image generation gemini-2.5-flash-image, controlling aspect ratio.
      • Image understanding
    • Image-to-image URL request returns URL request format OpenAI
      • Single image generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
      • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
      • Image understanding
  • NanoBanana
    • OpenAI request
      • Edit image
      • OpenAI image format
    • Gemini request
      • Generate image
      • Edit image
  • Midjourney format
    • Midjourney API Reference
    • Task query interface
    • Upload image
    • Get seed (Seed)
    • Submit Imagine task
    • Query tasks based on ID list
    • FaceSwap
    • Execute Action operation
    • /mj/submit/blend
    • Submit Describe task
    • Submit Modal
    • Refresh link
    • Edit image
    • Query task status by task ID
    • Get the seed of the task image
  • Doubao - Painting
    • doubao-seededit-3-0-i2i-250628
    • doubao-seedream-4-0-250828 - text-to-image
    • doubao-seedream-4-0-250828 - image-to-image
    • doubao-seedream-4-0-250828 - multi-image generation
  • Rerank Reordering Model
    • Rerank
  • Video Model
    • Grok Video Generation
      • 00-Overview
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Reference-to-Video
      • 04-Video-Editing
      • 05-Video-Extension
    • Seedance Video Generation
      • 00-Overview
      • 01-Create-Video-Generation-Task
      • 02-Query-Video-Generation-Task
      • 03-Query-Video-Generation-Task-List
      • 04-Cancel-or-Delete-Task
      • Seedance Private Asset Library API Documentation
    • MiniMax-H3 Video Generation
      • 00-Overview
      • 01-Create-Video-Generation
      • 02-Create-Video-Regeneration
      • 03-Create-H3-Context-IR
      • 04-Query-Task
      • 05-List-Tasks
      • 06-Cancel-or-Delete-Task
    • Hailuo Video Generation
      • 00-Overview
      • 01-Text-to-Video-T2V
      • 02-Image-to-Video-I2V
      • 03-First-Last-Frame-FL2V
      • 04-Subject-Reference-S2V
      • 05-Query-Task-Status
      • 06-Video-Download
      • 99-Appendix-Camera-Movement-and-Webhooks
    • Jimeng Video Generation
      • 00-Overview
      • 01-3.0-Pro-Video-Generation
      • 02-720P-Text-to-Video
      • 03-720P-Image-to-Video-First-Frame
      • 04-720P-Image-to-Video-Start-End-Frame
      • 05-720P-Image-to-Video-Camera
      • 06-1080P-Text-to-Video
      • 07-1080P-Image-to-Video-First-Frame
      • 08-1080P-Image-to-Video-Start-End-Frame
      • 09-Error-Codes
    • Kling AI Video Generation
      • 00-Overview
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Omni-Video
      • 04-Multi-Image-to-Video
      • 05-Motion-Control
      • 06-Multi-Elements
      • 07-Video-Extension
      • 08-Lip-Sync
      • 09-Avatar
      • 10-Text-to-Audio
      • 11-Video-to-Audio
      • 12-TTS
      • 13-Custom-Voices
      • 14-Image-Recognition
      • 15-Element-Management
      • 16-Video-Effects
    • Vidu Video Generation
      • 00-Overview
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Reference-to-Video
      • 04-Start-End-Frame
      • 05-Multi-Frame
      • 06-Scene-Template
      • 07-Template-Story
      • 08-Query-Tasks
    • HappyHorse
      • HappyHorse Text-to-Video
      • HappyHorse Image-to-Video (First Frame)
      • HappyHorse Reference-to-Video
      • HappyHorse Video Editing
    • Wan Video Generation
      • 00-Overview.md
      • 01-Text-to-Video
      • 02-Image-to-Video
      • 03-Reference-to-Video
      • 04-Video-Editing
      • 05-First-Last-Frame-to-Video
      • 06-Motion-Transfer-and-Character-Swap
      • 07-Digital-Human-Video
      • 08-VACE-Video-Editing
      • 09-Query-Task
  • Audio API
    • Audio API
    • Gemini TTS API
    • Google DeepMind Lyria API
    • Elevenlabs Speech to Text API Reference
    • Text-to-Music Suno
      • Task Submission
        • Generate Song (Inspiration Mode)
        • Generate Song (Custom Mode)
        • Generate Song (Continuation Mode)
        • Generate Song (Singer Style)
        • Generate Song (Secondary Creation from Uploaded Song)
        • Generate Song (Song Stitching)
        • Generate Lyrics
        • Song Stitching
      • Query Interface
        • Batch Retrieve Tasks
        • Query Single Task
  1. Vidu Video Generation

00-Overview

Vidu Video Generation · Overview#

Document Version: v1.0.0 | Last Updated: 2026-06-11
This platform fully supports the official Vidu video generation APIs. Requests and responses are transparently proxied with identical parameter semantics.

Capabilities#

Vidu video generation supports the following task types, each submitted through a dedicated endpoint:
Task TypeEndpointDescription
Text-to-Video (T2V)POST /vidu/ent/v2/text2videoGenerate video from a text prompt
Image-to-Video (I2V)POST /vidu/ent/v2/img2videoGenerate video driven by a first-frame image
Reference-to-VideoPOST /vidu/ent/v2/reference2videoGenerate video with subject consistency from reference images/videos; supports subject libraries
Start-End FramePOST /vidu/ent/v2/start-end2videoGenerate a transition video from specified start and end frame images
Smart Multi-FramePOST /vidu/ent/v2/multiframeGenerate long videos from multiple keyframe images
Scene Effect TemplatePOST /vidu/ent/v2/templateGenerate effect videos based on preset scene templates
Template StoryPOST /vidu/ent/v2/template-storyGenerate a complete video from a story template in one click
Video generation is asynchronous: after submission, a task_id is returned immediately. Use the Query Tasks endpoint to poll the task status. Once the task succeeds, the response contains creations.url (video download URL, valid for 24 hours) — please download or transfer the file promptly.

Standard Workflow#

Create Task (POST /vidu/ent/v2/{endpoint})
        │  Returns task_id + state: "created"
        ▼
Query Status (GET /vidu/ent/v2/tasks?task_ids={task_id})   ← Poll until success or failed
        │  Returns creations.url (valid for 24 hours)
        ▼
Download Video File

Endpoints#

APIMethodURL
Text-to-VideoPOSThttps://platform.dataeyes.ai/vidu/ent/v2/text2video
Image-to-VideoPOSThttps://platform.dataeyes.ai/vidu/ent/v2/img2video
Reference-to-VideoPOSThttps://platform.dataeyes.ai/vidu/ent/v2/reference2video
Start-End FramePOSThttps://platform.dataeyes.ai/vidu/ent/v2/start-end2video
Smart Multi-FramePOSThttps://platform.dataeyes.ai/vidu/ent/v2/multiframe
Scene Effect TemplatePOSThttps://platform.dataeyes.ai/vidu/ent/v2/template
Template StoryPOSThttps://platform.dataeyes.ai/vidu/ent/v2/template-story
Query Task ListGEThttps://platform.dataeyes.ai/vidu/ent/v2/tasks
Path Rule: Append the channel prefix /vidu after the platform domain, then concatenate the original Vidu API path.
For example, official https://api.vidu.cn/ent/v2/text2video → platform https://platform.dataeyes.ai/vidu/ent/v2/text2video.

Authentication#

All requests are authenticated via the HTTP Authorization header with a Bearer Token:
Authorization: Bearer {API_KEY}
{API_KEY} is the API key created in your platform console.
Note: The official Vidu API uses the Token prefix, but this platform uniformly uses the Bearer prefix. Do not send the Token prefix — it will be rejected with an "invalid token" error.

Request Headers#

HeaderRequiredDescription
AuthorizationYesBearer {API_KEY}
Content-TypeYesFixed as application/json

Model Overview#

ModelText-to-VideoImage-to-VideoReferenceStart-End FrameSmart Multi-FrameAudio-Video OutputDuration RangeDefault Resolution
viduq3-pro✅✅✅✅—✅ (enabled by default)1–16s720p
viduq3-turbo✅✅✅✅—✅ (enabled by default)1–16s720p
viduq3-pro-fast—✅———✅ (enabled by default)1–16s720p
viduq3-mix——✅ (non-subject only)——✅3–16s720p
viduq2-pro—✅✅✅✅Optional1–10s720p
viduq2-pro-fast—✅—✅—Optional1–10s720p
viduq2-turbo—✅—✅✅Optional1–10s720p
viduq2✅—✅———1–10s720p
viduq1✅✅✅✅——5s1080p
viduq1-classic—✅—✅——5s1080p
vidu2.0—✅✅✅——4s/8s360p/720p
The supported models may vary by endpoint. Refer to each endpoint's documentation for the specific model list.

Task States#

StateDescription
createdCreated successfully
queueingWaiting in queue
processingProcessing
successGeneration succeeded
failedGeneration failed

Callback Notifications#

All task creation endpoints support the callback_url parameter. When set, Vidu will send a POST callback request to the specified URL whenever the task state changes. The callback body structure is identical to the Query Tasks response. Callback states include processing, success, and failed, with up to 3 retries on failure.

Related Documentation#

Text-to-Video
Image-to-Video
Reference-to-Video
Start-End Frame
Smart Multi-Frame
Scene Effect Template
Template Story
Query Tasks
Previous
16-Video-Effects
Next
01-Text-to-Video