Generate shots from text descriptions
Use action=text2video, explicitly specify model=kling-v3-turbo, and describe the subject, action, composition, and sound requirements in the prompt.
Use kling-v3-turbo to generate 3–15 second video clips. Choose std 720p or pro 1080p output, and structure prompts around the subject, action, scene, and sound.
Use action=text2video, explicitly specify model=kling-v3-turbo, and describe the subject, action, composition, and sound requirements in the prompt.
Use action=image2video and start_image_url to make the input image the starting point of the shot, then describe the subject's movement and scene changes.
This model includes native audio and does not provide a switch to disable it. Omit generate_audio or set it to true; after generation, you should still verify that the audio meets delivery requirements.
Fit one clear action into 3–15 seconds, validate the visuals and sound first, then combine them into a longer work.
Use images of people, products, or scenes as the first frame, and describe changes around the subjects already present in the image.
Use real business prompts to compare quality, motion, and audio before deciding whether to use the generated result. The API does not guarantee that a single generation is ready for direct delivery.
When the task only requires text or a single first-frame image, and native audio is acceptable, you can start with kling-v3-turbo. std and pro correspond to different output resolutions; see Pricing for actual billing.
If you need an end frame, camera movement objects, an independent negative prompt, or cfg_scale, you should not apply these parameters directly to Turbo. For multi-image references and existing video editing, choose the appropriate specialized model and operation, and verify against that endpoint's documentation.
Determine the shot content and integer duration; for image-to-video, prepare an accessible start_image_url.
Set model=kling-v3-turbo, choose std or pro as needed, and use the corresponding action.
You can set async=true to obtain a task_id and then query the result, or configure callback_url to receive completion notifications. Successful submission does not mean video generation is complete.
Submit action=text2video, model=kling-v3-turbo, and prompt to POST /kling/videos, and set a valid quality and integer duration.
Use action=image2video and provide start_image_url. This model supports first-frame guidance but does not support end_image_url end frames.
Duration must be an integer from 3–15 seconds; std is 720p and pro is 1080p.
This model includes native audio and does not provide a toggle to disable it. Omit generate_audio or set it to true; for muted output, post-process the completed video.
camera_control, negative_prompt, and cfg_scale are not supported. Describe shot requirements and content to avoid in the prompt.
After setting async=true, first obtain the task_id, then query the status and final video result through the task API, or use callback_url to receive completion notifications.
Platform API contract · Updated: 2026-10-01.
Verify parameters and pricing before calling, and check the visuals and audio after generation.
Pricing
$0.11–$0.13· per second
You pay$13,584for 142,800 credits · $0.0951 / credits46% OFF
You're billed in credits. Pick a top-up tier to see its per-credit rate — the larger the top-up, the cheaper each credit.
This is only a basic call example. See the full docs for more parameters and advanced usage.
View Docs