How should I choose between std and pro in V2.6?
If you only create clips without synchronized audio and do not need to specify an end frame, you can choose std. If you need synchronized audio or start/end frame control, use pro. Both modes support 5-second or 10-second videos; these pro features must be enabled explicitly and are not all automatically turned on simply by selecting this mode.
How do I enable synchronized audio in V2.6?
In the /kling/videos request, explicitly set model to kling-v2-6, mode to pro, and generate_audio to true. This switch is off by default. Synchronized audio and submitting existing audio for photo lip-sync are two different approaches; choose based on whether you already have a recording.
Can I generate a video using only an end-frame image?
You cannot use an end frame alone as input for this image-to-video workflow. Use action=image2video, provide start_image_url, and add end_image_url as needed in pro mode. If you only have one image, use it as the start frame and describe the desired action in the prompt.
How do I make a photo speak according to a recording with V2.6?
Submit image_url and audio_url to /kling/talking-photo, explicitly specify model=kling-v2-6, and choose 5 seconds or 10 seconds. The prompt can describe the movements and expressions during the photo animation stage. The final video_url is the lip-synced video, while source_video_url is the intermediate photo animation video.
How do I receive generation results without keeping the connection waiting?
You can set async=true, save the returned task_id first, then retrieve the status and result through task queries; you can also configure callback_url to receive completion notifications. After obtaining video_url, download it or proceed to the editing workflow. Both generation endpoints should explicitly specify kling-v2-6 to avoid using other default models.