How can I ensure V2.5 Turbo is used when making a request?
Explicitly specify model=kling-v2-5-turbo in the /kling/videos or /kling/talking-photo request; do not rely on the default model. For video creation, choose either text2video or image2video; image-to-video requires a first frame, while talking photos require an image and audio.
What is the difference between std and pro for start and end frames?
V2.5 Turbo std does not support end frames, while pro supports start-and-end-frame guidance. When you need to specify the ending shot, choose image2video and pro, and submit both start_image_url and end_image_url; submitting only an end frame is not valid, nor can an end frame replace a first frame.
Can V2.5 Turbo directly generate videos with sound?
Standard video generation does not support generate_audio. If you already have voice-over audio, you can use the talking photo feature to combine a portrait image and audio into a lip-sync video; this is not the model creating sound on its own. For video tasks that need synchronized audio generation, choose a Kling version that supports this capability.
Can I use multiple reference images or reference videos to edit the scene?
This model is suited for text, first-frame, and pro start-and-end-frame creation, and should not be submitted with multiple images or reference videos according to the Omni workflow. When you need to reference multiple subjects, borrow video characteristics, or modify an existing video, choose kling-o1 or kling-v3-omni and organize the task according to the corresponding asset rules.
How do I receive the completed generated video?
Completed results provide video_url, video_id, task_id, and status information. For batch or background tasks, you can use async=true to obtain a task ID and then query the result, or configure callback_url to receive completion notifications; talking photo results additionally provide source_video_url for the intermediate animation.