Tạo một tác vụ image-to-video của Kling
Chuyển đổi một hình ảnh đơn thành video Kling thông qua CometAPI POST /kling/v1/videos/image2video, với việc tạo tác vụ bất đồng bộ và theo dõi tiến độ bằng task ID.
Trước khi gọi
- Cung cấp một URL hình ảnh công khai hoặc một chuỗi hình ảnh base64
- Sử dụng hình ảnh đáp ứng yêu cầu pixel của Kling; các ảnh thumbnail quá nhỏ sẽ bị tác vụ tạo từ chối
- Bắt đầu với ví dụ mặc định
kling-v2-6, sau đó chọnmodel_namekhác từ enum trong OpenAPI khi bạn cần một nhánh model khác - Giữ request đầu tiên ở mức đơn giản: một hình ảnh đầu vào, một prompt, không có khung hình cuối, không có motion mask
- Sử dụng
dynamic_masksdưới dạng một mảng các đối tượng mask-và-quỹ đạo khi bạn cần kiểm soát chuyển động cục bộ - Sử dụng
sound: offcho request đầu tiên không âm thanh có tính xác định trên các nhánh model hỗ trợ tạo âm thanh
Đặt tên model
Sử dụng các model ID video Kling thông thường trên endpoint này. Giữ Omni model ID cho Omni Video.Luồng tác vụ
Gửi request image-to-video
task_id của Kling được trả về.Poll tác vụ
Lưu kết quả
Ủy quyền
Bearer token authentication. Use your CometAPI key.
Nội dung
Source image URL or base64 image string. Use an image that meets Kling pixel requirements; very small thumbnails are rejected. For base64 input, send the encoded image string as the field value.
Text prompt describing the desired motion. Maximum 500 characters.
Elements to exclude from the video. Maximum 200 characters.
Webhook URL to receive task status updates when the task completes.
Generation mode. std for standard (faster), pro for professional (higher quality).
std, pro Model ID for this image-to-video request. Use kling-v3 for new requests. Use Omni model IDs only with the Omni Video endpoint.
kling-v1, kling-v1-5, kling-v1-6, kling-v2-master, kling-v2-1, kling-v2-1-master, kling-v2-5-turbo, kling-v2-6, kling-v3 Tail-frame reference image as a Base64 string or public URL. Same format requirements as image. Controls the last frame of the generated video.
Prompt adherence strength. Higher values follow the prompt more closely. Range: 0–1.
Output video length in seconds. Use 5 or 10; omit to use 5.
Static brush mask image as a Base64 string or public URL. White areas are frozen in place during video generation. Must match the aspect ratio and resolution of the input image.
Optional motion masks. Each entry contains a mask image and an ordered trajectory list.
Custom task id for your own tracking. Does not replace the system-generated task id but can be used to query tasks. Must be unique per user.
Optional generated-audio switch for models that support video sound. Use on or off, or omit the field for the model default.
on, off