Criar uma tarefa de image-to-video do Kling
Converta uma única imagem em um vídeo do Kling via CometAPI POST /kling/v1/videos/image2video, com criação assíncrona de tarefas e acompanhamento do progresso por task ID.
Antes de chamar
- Forneça uma URL pública de imagem ou uma string de imagem em base64
- Use uma imagem que atenda aos requisitos de pixels do Kling; miniaturas muito pequenas são rejeitadas pela tarefa de geração
- Comece com o exemplo padrão
kling-v2-6e depois escolha outromodel_namea partir do enum do OpenAPI quando precisar de uma trilha de modelo diferente - Mantenha a primeira solicitação simples: uma imagem de entrada, um prompt, sem frame final, sem máscaras de movimento
- Use
dynamic_maskscomo um array de objetos de máscara e trajetória quando precisar de movimento local controlado - Use
sound: offpara uma primeira solicitação determinística sem áudio em trilhas de modelo que oferecem suporte a som gerado
Nomenclatura de modelos
Use model IDs comuns de vídeo do Kling neste endpoint. Mantenha os model IDs Omni para Omni Video.Fluxo da tarefa
Envie a solicitação de image-to-video
task_id retornado pelo Kling.Consulte a tarefa
Armazene o resultado
Autorizações
Bearer token authentication. Use your CometAPI key.
Corpo
Source image URL or base64 image string. Use an image that meets Kling pixel requirements; very small thumbnails are rejected. For base64 input, send the encoded image string as the field value.
Text prompt describing the desired motion. Maximum 500 characters.
Elements to exclude from the video. Maximum 200 characters.
Webhook URL to receive task status updates when the task completes.
Generation mode. std for standard (faster), pro for professional (higher quality).
std, pro Model ID for this image-to-video request. Use kling-v3 for new requests. Use Omni model IDs only with the Omni Video endpoint.
kling-v1, kling-v1-5, kling-v1-6, kling-v2-master, kling-v2-1, kling-v2-1-master, kling-v2-5-turbo, kling-v2-6, kling-v3 Tail-frame reference image as a Base64 string or public URL. Same format requirements as image. Controls the last frame of the generated video.
Prompt adherence strength. Higher values follow the prompt more closely. Range: 0–1.
Output video length in seconds. Use 5 or 10; omit to use 5.
Static brush mask image as a Base64 string or public URL. White areas are frozen in place during video generation. Must match the aspect ratio and resolution of the input image.
Optional motion masks. Each entry contains a mask image and an ordered trajectory list.
Custom task id for your own tracking. Does not replace the system-generated task id but can be used to query tasks. Must be unique per user.
Optional generated-audio switch for models that support video sound. Use on or off, or omit the field for the model default.
on, off