Skip to main content
POST
Kling v3.0 4K Text-to-Video
Kling v3.0 4K text-to-video generates 4K ultra-high-definition videos from text prompts, with richer visual details, natural motion, and smooth scene dynamics. It supports flexible durations from 3 to 15 seconds, synchronized audio and video generation, and multi-shot video generation.
This is an asynchronous API and will only return the asynchronous task’s task_id. You should use this task_id to request the Get task result API to retrieve the generated result.

Request Headers

string
required
Enum value: application/json
string
required
Bearer authentication format: Bearer {{API Key}}.

Request Body

boolean
default:false
Whether to generate audio at the same time when generating the video.
string
required
The positive prompt text for generating the video. It can describe scene motion, camera movement, actions, voice style, atmosphere, and sound effects. Must not exceed 2500 characters. Mutually exclusive with multi_prompt.Length limit: 0 - 2500
integer
default:5
The duration of the generated video (in seconds). Supports flexible durations from 3 to 15 seconds.Value range: [3, 15]
number
Controls the flexibility of video generation. The higher the value, the more closely the model-generated content adheres to the prompt; the lower the value, the more natural the motion effects.Value range: [0, 1]
string
default:"16:9"
The aspect ratio of the generated video.Optional values: 16:9, 9:16, 1:1
string[]
A list of prompts for multi-shot video generation. Divides the video into multiple shots. Mutually exclusive with prompt.
string
Negative prompt, specifying elements to avoid in the visuals and audio. The length must not exceed 2500 characters.Length limit: 0 - 2500

Response

string
required
Use task_id to request the Get task result API to retrieve the generated output.