How can I ensure V2.5 Turbo is used when making a call?
Explicitly specify model=kling-v2-5-turbo in /kling/videos or /kling/talking-photo requests; do not rely on the default model. For video creation, also choose text2video or image2video; image-to-video requires a first frame, while talking photo requires an image and audio.
What is the difference between std and pro for start and end frames?
V2.5 Turbo std does not support end frames, while pro supports start and end frame guidance. When you need to specify the ending image, choose image2video and pro, and submit both start_image_url and end_image_url; submitting only an end frame is not valid, and an end frame cannot replace a start frame.
Can V2.5 Turbo directly generate videos with sound?
Regular video generation does not support generate_audio. If you already have voiceover audio, you can use the talking photo option to combine a portrait image and audio into a lip-sync video; this does not mean the model creates sound itself. For video tasks that need synchronized audio generation, choose a Kling version that supports this capability.
Can I use multiple reference images or reference videos to edit the scene?
This model is suited for text, start-frame, and pro start-and-end-frame creation, and should not be given multiple images or reference videos according to the Omni workflow. When you need to reference multiple subjects, use video features, or modify an existing video, choose kling-o1 or kling-v3-omni and organize the task according to the corresponding asset rules.
How can I receive the generated video after it is complete?
Completed results provide video_url, video_id, task_id, and status information. For batch or background tasks, you can use async=true to obtain a task ID and then query the result, or configure callback_url to receive completion notifications; talking photo results additionally provide source_video_url for the intermediate animation.