DataEyesAI
Official SiteConsoleDocs Home
Getting StartedDeveloper ToolsAI Models APITerms & Policies
Official SiteConsoleDocs Home
Getting StartedDeveloper ToolsAI Models APITerms & Policies
  1. Kling AI Video Generation
  • Documentation
    • Getting Started
      • Overview
      • Console (Getting Started)
      • API Key
      • Base URL
    • Developer Tool Integration
      • OpenClaw
      • Claude Code
      • Codex
      • Gemini CLI
      • Grok CLI
      • Other Tools
    • AI Models API
      • OpenAI format (supports major original models)
        • Chat (Response)
          • Create Network Search
          • Create Model Response GPT-5 Enable Thinking
          • Create Function Call
          • Create Model Response
          • Create Model Response (Streaming Return)
          • Create Model Response (Control Thinking Length)
        • ChatGPT Interface
          • Audio
            • Audio to text gpt-4o-transcribe
            • GPT-4o-audio
            • Audio to text whisper-1
            • Audio to text gpt-4o-transcribe
            • Create voice gpt-4o-mini-tts
          • Chat
            • Create chat-based image recognition (non-streaming)
            • Create chat-based image recognition (streaming)
            • Create chat-based image recognition (streaming) best64
            • Official N test
            • Create structured output
            • Control the effort level of the inference model
            • Create chat function call
            • deepseek-ocr recognition
            • Create chat completion (non-stream)
          • Completions
            • ChatGPT automatic completion
            • Create completion
        • Image
          • Edit image
          • Create chat completion (streaming)
          • Create chat completion (qwen-mt-turbo)
          • Create chat completion with deepseek v3.1 level of reasoning (streaming)
        • Audio
          • Speech recognition
          • Speech synthesis
          • Official Function Calling invocation
          • Create chat-generated images (non-streaming)
        • Embedding
          • Text embeddings
      • Anthropic format
        • Chat
        • Chat(prompt cache)
        • Streaming response
        • Chat (deep reasoning)
        • Tool invocation (function call)
        • Analyze image
      • Google Gemini interface
        • Native format
          • Text-to-image + control over aspect ratio + clarity
          • Generate image
          • Text generation
          • Text generation - stream
          • Text generation + reasoning - stream
          • Image generation
          • Formatted output
          • Function call
          • Document understanding
          • URL context [native format]
          • Code execution
          • Video understanding
          • URL context
          • Video understanding - url [native format]
          • Imagen 4
          • Audio understanding
          • Embeddings
          • Chat
          • Edit image
        • Image-to-image Base64 request method
          • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity
          • Image editing
          • Single image gemini-3-pro-image-preview, controlling aspect ratio and clarity.
          • Image generation( gemini-2.5-flash-image)
          • Image generation gemini-2.5-flash-image, controlling aspect ratio.
          • Image understanding
        • Image-to-image URL request returns URL request format OpenAI
          • Single image generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
          • Multi-image fusion slice generation with gemini-3-pro-image-preview, controlling aspect ratio and clarity.
          • Image understanding
      • NanoBanana
        • OpenAI request
          • Edit image
          • OpenAI image format
        • Gemini request
          • Generate image
          • Edit image
      • Midjourney format
        • Midjourney API Reference
        • Task query interface
        • Upload image
        • Get seed (Seed)
        • Submit Imagine task
        • Query tasks based on ID list
        • FaceSwap
        • Execute Action operation
        • /mj/submit/blend
        • Submit Describe task
        • Submit Modal
        • Refresh link
        • Edit image
        • Query task status by task ID
        • Get the seed of the task image
      • Doubao - Painting
        • doubao-seededit-3-0-i2i-250628
        • doubao-seedream-4-0-250828 - text-to-image
        • doubao-seedream-4-0-250828 - image-to-image
        • doubao-seedream-4-0-250828 - multi-image generation
      • Rerank Reordering Model
        • Rerank
      • Video Model
        • Grok Video Generation
          • 00-Overview
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Reference-to-Video
          • 04-Video-Editing
          • 05-Video-Extension
        • Seedance Video Generation
          • 00-Overview
          • 01-Create-Video-Generation-Task
          • 02-Query-Video-Generation-Task
          • 03-Query-Video-Generation-Task-List
          • 04-Cancel-or-Delete-Task
          • Seedance Private Asset Library API Documentation
        • MiniMax-H3 Video Generation
          • 00-Overview
          • 01-Create-Video-Generation
          • 02-Create-Video-Regeneration
          • 03-Create-H3-Context-IR
          • 04-Query-Task
          • 05-List-Tasks
          • 06-Cancel-or-Delete-Task
        • Hailuo Video Generation
          • 00-Overview
          • 01-Text-to-Video-T2V
          • 02-Image-to-Video-I2V
          • 03-First-Last-Frame-FL2V
          • 04-Subject-Reference-S2V
          • 05-Query-Task-Status
          • 06-Video-Download
          • 99-Appendix-Camera-Movement-and-Webhooks
        • Jimeng Video Generation
          • 00-Overview
          • 01-3.0-Pro-Video-Generation
          • 02-720P-Text-to-Video
          • 03-720P-Image-to-Video-First-Frame
          • 04-720P-Image-to-Video-Start-End-Frame
          • 05-720P-Image-to-Video-Camera
          • 06-1080P-Text-to-Video
          • 07-1080P-Image-to-Video-First-Frame
          • 08-1080P-Image-to-Video-Start-End-Frame
          • 09-Error-Codes
        • Kling AI Video Generation
          • 00-Overview
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Omni-Video
          • 04-Multi-Image-to-Video
          • 05-Motion-Control
          • 06-Multi-Elements
          • 07-Video-Extension
          • 08-Lip-Sync
          • 09-Avatar
          • 10-Text-to-Audio
          • 11-Video-to-Audio
          • 12-TTS
          • 13-Custom-Voices
          • 14-Image-Recognition
          • 15-Element-Management
          • 16-Video-Effects
        • Vidu Video Generation
          • 00-Overview
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Reference-to-Video
          • 04-Start-End-Frame
          • 05-Multi-Frame
          • 06-Scene-Template
          • 07-Template-Story
          • 08-Query-Tasks
        • HappyHorse
          • HappyHorse Text-to-Video
          • HappyHorse Image-to-Video (First Frame)
          • HappyHorse Reference-to-Video
          • HappyHorse Video Editing
        • Wan Video Generation
          • 00-Overview.md
          • 01-Text-to-Video
          • 02-Image-to-Video
          • 03-Reference-to-Video
          • 04-Video-Editing
          • 05-First-Last-Frame-to-Video
          • 06-Motion-Transfer-and-Character-Swap
          • 07-Digital-Human-Video
          • 08-VACE-Video-Editing
          • 09-Query-Task
      • Audio API
        • Audio API
        • Gemini TTS API
        • Google DeepMind Lyria API
        • Elevenlabs Speech to Text API Reference
        • Text-to-Music Suno
          • Task Submission
            • Generate Song (Inspiration Mode)
            • Generate Song (Custom Mode)
            • Generate Song (Continuation Mode)
            • Generate Song (Singer Style)
            • Generate Song (Secondary Creation from Uploaded Song)
            • Generate Song (Song Stitching)
            • Generate Lyrics
            • Song Stitching
          • Query Interface
            • Batch Retrieve Tasks
            • Query Single Task
    • Search / Reader Product
      • Web Reader API​​
        • Web Reader API
        • Web Reader API(HK)
      • Web Search API​​
        • Modal Card API
          • Weather
            • All City ID
            • Weather Query API
        • Web Search API
        • Video Search api
        • Trending Search API
      • Document OCR Parsing API
        • fiel upload
        • URL Parsing
    • Advanced & System API
      • Data Updates
      • System interface
        • API Key & Quota Query API
        • API Key Management API
      • API Reference​​
        • Error Codes
        • HTTP Notes
      • List models
        • Models
  • Terms & Policies
    • DataEyesAI API Terms of Service
    • DataEyesAI Legal Notice and Privacy Policy
    • DataEyesAI Paid Services Agreement
    • Automatic Renewal Service Rules
  1. Kling AI Video Generation

12-TTS

TTS#

Doc version: v1.0.0 | Last updated: 2026-06-11
This platform fully supports the Kling AI official video generation API. Requests and responses are transparently proxied; parameter semantics are identical to the official API.

Create Task#

POST https://platform.dataeyes.ai/kling/v1/audio/tts
Text-to-Speech synthesis API for generating audio from text.

Request Headers#

ParameterTypeRequiredDefaultDescription
Content-TypestringYesapplication/jsonData Exchange Format
AuthorizationstringYesAuthentication information, refer to API authentication

Request Body#

ParameterTypeRequiredDefaultDescription
textstringYesText Content for Audio Synthesis
- The maximum length of the text content is 1000 characters; content that is too long will return an error code and other information.
- The system will validate the text content; if there are issues, it will return an error code and other information.
voice_idstringYesVoice ID
- The system offers a variety of voice options to choose from. For specific voice effects, voice IDs, and corresponding voice languages, see Voice Guide. Voice previews do not support custom scripts.
- Voice preview file naming convention: Voice Name#Voice ID#Voice Language
voice_languagestringYeszhVoice Language
Options: zh, en
- The voice language corresponds to the Voice ID, as detailed above.
voice_speedfloatNo1.0Speech Rate
- Valid range: [0.8, 2.0], accurate to one decimal place; values outside this range will be automatically rounded.

Request Example#

curl --request POST \
  --url https://platform.dataeyes.ai/kling/v1/audio/tts \
  --header 'Authorization: Bearer <token>' \
  --header 'Content-Type: application/json' \
  --data '{
    "text": "Throughout my time in college, several memorable event left a significant impact on my life",
    "voice_id": "oversea_male1",
    "voice_language": "en",
    "voice_speed": 1
  }'

Response Example#

{
  "code": 0, // Error codes; Specific definitions can be found in Error codes
  "message": "string", // Error information
  "request_id": "string", // Request ID, generated by the system, is used to track requests and troubleshoot problems
  "data": {
    "task_id": "string", // Task ID, generated by the system
    "task_status": "string", // Task status, Enum values: submitted, processing, succeed, failed
    "task_status_msg": "string", // Task status information, displaying the failure reason when the task fails (such as triggering the content risk control of the platform, etc.)
    "task_result": {
      "audios": [
        {
          "id": "string", // Generated sound ID; globally unique, will be cleared after 30 days
          "url": "string", // URL for generating sounds,such as https://p1.a.kwimgs.com/bs2/upload-ylab-stunt/special-effect/output/HB1_PROD_ai_web_46554461/-2878350957757294165/output.mp3(To ensure information security, generated images/videos will be cleared after 30 days. Please make sure to save them promptly.)
          "duration": "string" // Total audio duration, unit: s (seconds)
        }
      ]
    },
    "final_unit_deduction": "string", // The deduction units of task
    "created_at": 1722769557708, // Task creation time, Unix timestamp, unit: ms
    "updated_at": 1722769557708 // Task update time, Unix timestamp, unit: ms
  }
}
Previous
11-Video-to-Audio
Next
13-Custom-Voices