> ## Documentation Index
> Fetch the complete documentation index at: https://apidoc.cometapi.com/llms.txt
> Use this file to discover all available pages before exploring further.

# Créer une vidéo MiniMax H3

> Créez une tâche vidéo MiniMax H3 ou H3-Max avec du texte, des médias de référence ou des première et dernière images via CometAPI

Créez une vidéo MiniMax H3 ou H3-Max à partir de texte, de médias de référence ou de première et dernière images. Stockez le `id` renvoyé pour interroger la tâche et télécharger la vidéo terminée.

`POST /v1/videos` utilise `multipart/form-data`. Envoyez les contrôles sous forme de champs de formulaire et répétez chaque champ de référence une fois par entrée média.

## Choisir un modèle

Définissez `model` sur l’un des ID de modèle suivants. Choisissez une taille de requête dans le tableau correspondant au modèle dans [Définir la durée et la taille](#set-duration-and-size).

| ID de modèle     | Niveaux de résolution |
| ---------------- | --------------------- |
| `minimax-h3`     | `768P`, `2K`          |
| `minimax-h3-max` | `480P`, `768P`        |

## Choisir un mode d’entrée

Les deux modèles utilisent les champs ci-dessous. Définissez `model` sur l’ID du modèle que vous avez sélectionné.

| Objectif                                             | Champs obligatoires                                                             | Champs facultatifs                      |
| ---------------------------------------------------- | ------------------------------------------------------------------------------- | --------------------------------------- |
| Texte vers vidéo                                     | `model`, `prompt`                                                               | `seconds`, `size`                       |
| Images de référence                                  | `model`, `prompt`, une ou plusieurs `input_reference` URL ou fichiers           | `seconds`, `size`                       |
| URL, URI de données ou fichier de vidéo de référence | `model`, `prompt`, un ou plusieurs champs `reference_videos`                    | `seconds`, `size`                       |
| Audio de référence avec une image                    | `model`, `prompt`, `input_reference`, un ou plusieurs champs `reference_audios` | `seconds`, `size`                       |
| URL de la première et de la dernière image           | `model`, `prompt`, `first_frame`                                                | `last_frame`, `seconds`, `size`         |
| Première et dernière images importées                | `model`, `prompt`, `first_frame` fichier                                        | `last_frame` fichier, `seconds`, `size` |

Les exemples de requêtes H3-Max incluent la génération de vidéo à partir de texte aux deux niveaux de résolution, des images de référence, des références combinées d’image, de vidéo et d’audio, ainsi que des première et dernière images.

## Utiliser des images de référence

Envoyez des images de référence via `input_reference` sous forme d’URL HTTPS accessibles publiquement ou de fichiers importés. Pour plusieurs images, répétez le champ dans l’ordre des images, avec une URL ou un fichier par champ.

Utilisez uniquement des URL ou uniquement des fichiers dans ce champ. Envoyez chaque URL séparément au lieu de combiner les URL dans une valeur séparée par des virgules.

Vous pouvez envoyer jusqu’à neuf images de référence PNG ou JPEG, que vous utilisiez des URL ou des fichiers importés. Limitez chaque image importée à 20 Mio ou moins.

Dans le prompt, décrivez les détails visuels que la vidéo générée doit préserver.

## Utiliser une vidéo et un fichier audio de référence

Envoyez `reference_videos` et `reference_audios` sous forme de valeurs textuelles multipart ou de fichiers importés, en utilisant les formats ci-dessous.
Répétez chaque champ une fois par entrée, en utilisant uniquement des valeurs textuelles ou uniquement des fichiers dans ce champ. Ne regroupez pas les entrées dans une valeur séparée par des virgules.

Différents champs peuvent utiliser différents types d'entrée, par exemple une URL d'image avec un fichier audio importé.
Pour combiner des références d'image, de vidéo et d'audio, envoyez `input_reference`, `reference_videos` et `reference_audios` dans la même requête.

| Champ              | Entrée                                                                         | Format                                                     |
| ------------------ | ------------------------------------------------------------------------------ | ---------------------------------------------------------- |
| `reference_videos` | Une URL HTTPS MP4 directe et publiquement accessible ou un fichier MP4 importé | MP4 avec vidéo H.264 ; une piste audio AAC est facultative |
| `reference_videos` | Un URI de données Base64 complet dans un champ texte                           | `data:video/mp4;base64,<BASE64_MP4>`                       |
| `reference_audios` | Une URL HTTPS MP3 directe et publiquement accessible ou un fichier MP3 importé | MP3                                                        |

Envoyez au maximum 12 entrées de référence au total, dont trois vidéos et trois extraits audio au maximum.
Chaque clip vidéo ou audio doit durer de 2 à 15 secondes. La durée cumulée des vidéos et celle des fichiers audio doivent chacune être inférieures ou égales à 15 secondes.

Associez l'audio de référence à au moins une image via `input_reference`.
Décrivez le mouvement, la synchronisation ou le son que la vidéo générée doit suivre.

La taille de chaque MP4 importé ne doit pas dépasser 50 MiB et celle de chaque MP3 importé ne doit pas dépasser 15 MiB. Un MiB correspond à 1024 × 1024 octets.

Les exemples de requêtes en haut de cette page montrent des URL MP4, des URI de données MP4 et des URL MP3 associés à une image.

### Importer une vidéo de référence

Importez un fichier MP4 pour guider la structure du plan et le mouvement :

```bash theme={null}
curl \
  https://api.cometapi.com/v1/videos \
  -H "Authorization: Bearer $COMETAPI_KEY" \
  --form-string 'model=minimax-h3' \
  --form-string 'prompt=Follow the uploaded reference video shot structure and motion. No text.' \
  --form-string 'seconds=5' \
  --form-string 'size=1344x768' \
  --form 'reference_videos=@/path/to/reference.mp4;type=video/mp4'
```

### Importer un fichier audio de référence avec une image

Associez une URL d'image de référence à un fichier MP3 importé :

```bash theme={null}
curl \
  https://api.cometapi.com/v1/videos \
  -H "Authorization: Bearer $COMETAPI_KEY" \
  --form-string 'model=minimax-h3' \
  --form-string 'prompt=Animate the reference image and synchronize the motion with the uploaded reference audio. No text.' \
  --form-string 'seconds=5' \
  --form-string 'size=1344x768' \
  --form-string 'input_reference=https://your-image-host/reference.png' \
  --form 'reference_audios=@/path/to/reference.mp3;type=audio/mpeg'
```

## Définir la première et la dernière image

Envoyez `first_frame` pour l'image d'ouverture et, éventuellement, un `last_frame` pour l'image de fin.
Chaque champ accepte une URL HTTPS PNG directe et publiquement accessible ou un fichier PNG importé. La taille de chaque PNG importé ne doit pas dépasser 20 MiB (20 × 1024 × 1024 octets).

Ne combinez pas les champs d'image avec les champs d'image, de vidéo ou d'audio de référence. Utilisez une requête distincte pour la génération basée sur des images de début et de fin.

Avec des images de début et de fin, `size` sélectionne le niveau de résolution. Le service peut normaliser les dimensions finales de sortie en fonction des images fournies.
Continuez à envoyer `size` sous la forme `WxH`.

<Note>
  Enregistrez `id` renvoyé. Interrogez la tâche jusqu'à ce que `status` soit `completed` ou
  `failed`. Lorsque la tâche est terminée, téléchargez le contenu vidéo.
</Note>

## Définir la durée et la taille

Pour les deux modèles, définissez `seconds` sur un entier compris entre `5` et `15`. La valeur par défaut est de `5` secondes.

Choisissez une résolution et un rapport d’aspect dans le tableau correspondant à votre modèle, puis envoyez la valeur correspondante de **Taille de la requête** au format `size` dans `WxH`.
Pour les requêtes de texte et de médias de référence, `size` sélectionne le niveau de résolution et le rapport d’aspect du cadrage.
Le service peut normaliser les dimensions finales de sortie selon la taille demandée et les médias de référence fournis.

### Tailles pour `minimax-h3`

Utilisez ces tailles de requête avec `model=minimax-h3` :

| Résolution | Rapport d’aspect | Taille de la requête |
| ---------- | ---------------- | -------------------- |
| `768P`     | `21:9`           | `1536x672`           |
| `768P`     | `16:9`           | `1344x768`           |
| `768P`     | `4:3`            | `1024x768`           |
| `768P`     | `1:1`            | `768x768`            |
| `768P`     | `3:4`            | `768x1024`           |
| `768P`     | `9:16`           | `768x1344`           |
| `2K`       | `21:9`           | `2912x1280`          |
| `2K`       | `16:9`           | `2544x1456`          |
| `2K`       | `4:3`            | `1920x1440`          |
| `2K`       | `1:1`            | `1440x1440`          |
| `2K`       | `3:4`            | `1440x1920`          |
| `2K`       | `9:16`           | `1440x2560`          |

### Tailles pour `minimax-h3-max`

Utilisez ces tailles de requête avec `model=minimax-h3-max` :

| Résolution | Rapport d’aspect | Taille de la requête |
| ---------- | ---------------- | -------------------- |
| `480P`     | `21:9`           | `1120x480`           |
| `480P`     | `16:9`           | `864x480`            |
| `480P`     | `4:3`            | `640x480`            |
| `480P`     | `1:1`            | `480x480`            |
| `480P`     | `3:4`            | `480x640`            |
| `480P`     | `9:16`           | `480x864`            |
| `768P`     | `21:9`           | `1792x768`           |
| `768P`     | `16:9`           | `1360x768`           |
| `768P`     | `4:3`            | `1024x768`           |
| `768P`     | `1:1`            | `768x768`            |
| `768P`     | `3:4`            | `768x1024`           |
| `768P`     | `9:16`           | `768x1360`           |

## Flux de tâches

<Steps>
  <Step title="Créer la tâche">
    Envoyez la requête de formulaire multipart et stockez la valeur `id` renvoyée.
  </Step>

  <Step title="Interroger la tâche">
    Appelez [Récupérer une vidéo MiniMax H3](./retrieve) jusqu’à ce que `status` soit `completed` ou `failed`.
  </Step>

  <Step title="Télécharger le résultat">
    Lorsque la tâche est `completed`, appelez [Télécharger le contenu vidéo MiniMax H3](./retrieve-content) pour enregistrer le fichier MP4.
  </Step>
</Steps>


## OpenAPI

````yaml api/openapi/video/minimax-h3/post-create.openapi.json POST /v1/videos
openapi: 3.1.0
info:
  title: MiniMax H3 Video Create API
  version: 1.0.0
  description: >-
    Create an asynchronous MiniMax H3 or H3-Max video task from text, reference
    media, or first and last frames through CometAPI. Send the request as
    multipart/form-data, then use the returned id to retrieve the task and
    download its completed MP4 file.
servers:
  - url: https://api.cometapi.com
security:
  - bearerAuth: []
paths:
  /v1/videos:
    post:
      summary: Create a MiniMax H3 video task
      description: >-
        Create a MiniMax H3 or H3-Max video task from text, reference images,
        reference video or audio, or first and last frames. Repeat
        input_reference with one publicly accessible HTTPS image URL or uploaded
        image file per field, up to nine images. Use all URLs or all files
        within input_reference. Send size in WxH form using the table for the
        selected model. The service can normalize the final output dimensions.
      operationId: minimax_h3_create_video
      requestBody:
        required: true
        content:
          multipart/form-data:
            schema:
              $ref: '#/components/schemas/MiniMaxH3CreateRequest'
            encoding:
              input_reference:
                contentType: text/plain, image/png, image/jpeg
                style: form
                explode: true
              reference_videos:
                style: form
                explode: true
                contentType: text/plain, video/mp4
              reference_audios:
                style: form
                explode: true
                contentType: text/plain, audio/mpeg
              first_frame:
                style: form
              last_frame:
                style: form
            examples:
              text_2544x1456:
                summary: Text-to-video at 2544x1456
                value:
                  model: minimax-h3
                  prompt: >-
                    A paper boat glides across a still pond, daylight, locked
                    camera, no text.
                  seconds: 5
                  size: 2544x1456
              text_1344x768:
                summary: Text-to-video at 1344x768
                value:
                  model: minimax-h3
                  prompt: >-
                    A cinematic view of clouds moving over green mountains,
                    locked camera, no text.
                  seconds: 5
                  size: 1344x768
              url_reference_image:
                summary: HTTPS reference images at 1344x768
                value:
                  model: minimax-h3
                  prompt: >-
                    The referenced characters walk together through a city
                    street. Preserve their appearance and use natural motion. No
                    text.
                  seconds: 5
                  size: 1344x768
                  input_reference:
                    - https://your-image-host/reference-1.png
                    - https://your-image-host/reference-2.jpg
              uploaded_reference_images:
                summary: Uploaded reference images
                value:
                  model: minimax-h3
                  prompt: >-
                    The uploaded reference characters walk together through a
                    city street. Preserve their appearance and use natural
                    motion. No text.
                  seconds: 5
                  size: 2544x1456
                  input_reference:
                    - '@/path/to/reference-1.png'
                    - '@/path/to/reference-2.jpg'
              https_reference_video:
                summary: HTTPS MP4 reference video
                value:
                  model: minimax-h3
                  prompt: >-
                    Follow the reference clip's shot structure and motion. No
                    text.
                  seconds: 5
                  size: 1344x768
                  reference_videos:
                    - https://your-media-host/reference.mp4
              https_reference_audio:
                summary: HTTPS MP3 reference audio with an image
                value:
                  model: minimax-h3
                  prompt: >-
                    Animate the reference image and synchronize the motion with
                    the reference audio. No text.
                  seconds: 5
                  size: 1344x768
                  reference_audios:
                    - https://your-media-host/reference.mp3
                  input_reference: https://your-image-host/reference.png
              https_first_last_frames:
                summary: HTTPS first and last frames
                value:
                  model: minimax-h3
                  prompt: >-
                    Move naturally from the opening frame to the ending frame.
                    No text.
                  seconds: 5
                  size: 1344x768
                  first_frame: https://your-image-host/first-frame.png
                  last_frame: https://your-image-host/last-frame.png
              uploaded_first_last_frames:
                summary: Uploaded first and last frames
                value:
                  model: minimax-h3
                  prompt: >-
                    Move naturally from the uploaded opening frame to the
                    uploaded ending frame. No text.
                  seconds: 5
                  size: 1344x768
                  first_frame: <binary PNG file>
                  last_frame: <binary PNG file>
              uploaded_reference_video:
                summary: Uploaded MP4 reference video
                value:
                  model: minimax-h3
                  prompt: >-
                    Follow the uploaded reference video shot structure and
                    motion. No text.
                  seconds: 5
                  size: 1344x768
                  reference_videos:
                    - '@/path/to/reference.mp4'
              uploaded_reference_audio_with_image:
                summary: Uploaded MP3 reference audio with an image
                value:
                  model: minimax-h3
                  prompt: >-
                    Animate the reference image and synchronize the motion with
                    the uploaded reference audio. No text.
                  seconds: 5
                  size: 1344x768
                  reference_audios:
                    - '@/path/to/reference.mp3'
                  input_reference: https://your-image-host/reference.png
              h3_max_text_480p:
                summary: 'H3-Max: text-to-video at 480P'
                value:
                  model: minimax-h3-max
                  prompt: A paper boat crosses a pond. Fixed camera. No text.
                  seconds: 5
                  size: 864x480
              h3_max_text_768p:
                summary: 'H3-Max: text-to-video at 768P'
                value:
                  model: minimax-h3-max
                  prompt: A paper boat crosses a pond. Fixed camera. No text.
                  seconds: 5
                  size: 1360x768
              h3_max_reference_images:
                summary: 'H3-Max: HTTPS reference images'
                value:
                  model: minimax-h3-max
                  prompt: The two reference characters walk through a park. No text.
                  seconds: 5
                  size: 1360x768
                  input_reference:
                    - https://your-image-host/reference-1.png
                    - https://your-image-host/reference-2.jpg
              h3_max_mixed_references:
                summary: 'H3-Max: image, video, and audio references'
                value:
                  model: minimax-h3-max
                  prompt: >-
                    Use the image for the subject, video for motion, and audio
                    for sound.
                  seconds: 5
                  size: 1360x768
                  input_reference:
                    - https://your-image-host/reference.png
                  reference_videos:
                    - https://your-media-host/reference.mp4
                  reference_audios:
                    - https://your-media-host/reference.mp3
              h3_max_first_last_frames:
                summary: 'H3-Max: first and last frames'
                value:
                  model: minimax-h3-max
                  prompt: Move from the opening frame to the ending frame. No text.
                  seconds: 5
                  size: 1360x768
                  first_frame: https://your-image-host/first-frame.png
                  last_frame: https://your-image-host/last-frame.png
      responses:
        '200':
          description: >-
            Task created. Store the returned id and use it in retrieve and
            content requests.
          content:
            application/json:
              schema:
                $ref: '#/components/schemas/MiniMaxH3VideoTask'
              example:
                id: <task_id>
                task_id: <task_id>
                object: video
                model: minimax-h3
                status: queued
                progress: 0
                created_at: 1779938152
      security:
        - bearerAuth: []
      x-codeSamples:
        - lang: Shell
          label: Text-to-video at 2544x1456
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=A paper boat glides across a still pond, daylight, locked camera, no text.' \
              --form-string 'seconds=5' \
              --form-string 'size=2544x1456'
        - lang: Python
          label: Text-to-video at 2544x1456
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "A paper boat glides across a still pond, daylight, "
                        "locked camera, no text.",
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "2544x1456")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: Text-to-video at 2544x1456
          source: |
            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "A paper boat glides across a still pond, daylight, " +
                "locked camera, no text.",
            );
            form.append("seconds", "5");
            form.append("size", "2544x1456");

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: Text-to-video at 1344x768
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=A cinematic view of clouds moving over green mountains, locked camera, no text.' \
              --form-string 'seconds=5' \
              --form-string 'size=1344x768'
        - lang: Python
          label: Text-to-video at 1344x768
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "A cinematic view of clouds moving over green mountains, "
                        "locked camera, no text.",
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1344x768")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: Text-to-video at 1344x768
          source: |
            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "A cinematic view of clouds moving over green mountains, " +
                "locked camera, no text.",
            );
            form.append("seconds", "5");
            form.append("size", "1344x768");

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: HTTPS reference images at 1344x768
          source: |
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string model=minimax-h3 \
              --form-string 'prompt=The referenced characters walk together through a city street. Preserve their appearance and use natural motion. No text.' \
              --form-string seconds=5 \
              --form-string size=1344x768 \
              --form-string input_reference=https://your-image-host/reference-1.png \
              --form-string input_reference=https://your-image-host/reference-2.jpg
        - lang: Python
          label: HTTPS reference images at 1344x768
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "The referenced characters walk together through a city street. "
                        "Preserve their appearance and use natural motion. No text."
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1344x768")),
                ("input_reference", (None, "https://your-image-host/reference-1.png")),
                ("input_reference", (None, "https://your-image-host/reference-2.jpg")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )

            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: HTTPS reference images at 1344x768
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3");

            form.append(
              "prompt",
              "The referenced characters walk together through a city street. " +
                "Preserve their appearance and use natural motion. No text.",
            );

            form.append("seconds", "5");

            form.append("size", "1344x768");

            form.append("input_reference",
            "https://your-image-host/reference-1.png");

            form.append("input_reference",
            "https://your-image-host/reference-2.jpg");


            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: Uploaded reference images
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=The uploaded reference characters walk together through a city street. Preserve their appearance and use natural motion. No text.' \
              --form-string 'seconds=5' \
              --form-string 'size=2544x1456' \
              --form 'input_reference=@/path/to/reference-1.png;type=image/png' \
              --form 'input_reference=@/path/to/reference-2.jpg;type=image/jpeg'
        - lang: Python
          label: Uploaded reference images
          source: |
            import os
            from contextlib import ExitStack

            import requests

            prompt = (
                "The uploaded reference characters walk together through a city street. "
                "Preserve their appearance and use natural motion. No text."
            )

            with ExitStack() as stack:
                files = [
                    (
                        "input_reference",
                        (
                            "reference-1.png",
                            stack.enter_context(
                                open("/path/to/reference-1.png", "rb")
                            ),
                            "image/png",
                        ),
                    ),
                    (
                        "input_reference",
                        (
                            "reference-2.jpg",
                            stack.enter_context(
                                open("/path/to/reference-2.jpg", "rb")
                            ),
                            "image/jpeg",
                        ),
                    ),
                ]
                response = requests.post(
                    "https://api.cometapi.com/v1/videos",
                    headers={
                        "Authorization": "Bearer "
                        + os.environ["COMETAPI_KEY"]
                    },
                    data={
                        "model": "minimax-h3",
                        "prompt": prompt,
                        "seconds": "5",
                        "size": "2544x1456",
                    },
                    files=files,
                    timeout=120,
                )

            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: Uploaded reference images
          source: |
            import { readFile } from "node:fs/promises";

            const firstReference = await readFile("/path/to/reference-1.png");
            const secondReference = await readFile("/path/to/reference-2.jpg");

            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "The uploaded reference characters walk together through a city street. " +
                "Preserve their appearance and use natural motion. No text.",
            );
            form.append("seconds", "5");
            form.append("size", "2544x1456");
            form.append(
              "input_reference",
              new Blob([firstReference], { type: "image/png" }),
              "reference-1.png",
            );
            form.append(
              "input_reference",
              new Blob([secondReference], { type: "image/jpeg" }),
              "reference-2.jpg",
            );

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: HTTPS MP4 reference video
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=Follow the reference clip shot structure and motion. No text.' \
              --form-string 'seconds=5' \
              --form-string 'size=1344x768' \
              --form-string 'reference_videos=https://your-media-host/reference.mp4'
        - lang: Python
          label: HTTPS MP4 reference video
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "Follow the reference clip shot structure and motion. "
                        "No text.",
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1344x768")),
                (
                    "reference_videos",
                    (None, "https://your-media-host/reference.mp4"),
                ),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: HTTPS MP4 reference video
          source: |
            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "Follow the reference clip shot structure and motion. No text.",
            );
            form.append("seconds", "5");
            form.append("size", "1344x768");
            form.append(
              "reference_videos",
              "https://your-media-host/reference.mp4",
            );

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: MP4 reference video data URI
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=Follow the reference clip shot structure and motion. No text.' \
              --form-string 'seconds=5' \
              --form-string 'size=1344x768' \
              --form 'reference_videos=<-;type=text/plain' \
              < <(printf 'data:video/mp4;base64,'; openssl base64 -A -in /path/to/reference.mp4)
        - lang: Python
          label: MP4 reference video data URI
          source: |
            import base64
            import os
            from pathlib import Path

            import requests

            video_base64 = base64.b64encode(
                Path("/path/to/reference.mp4").read_bytes()
            ).decode("ascii")
            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "Follow the reference clip shot structure and motion. "
                        "No text.",
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1344x768")),
                (
                    "reference_videos",
                    (None, "data:video/mp4;base64," + video_base64),
                ),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: MP4 reference video data URI
          source: |
            import { readFile } from "node:fs/promises";

            const videoBase64 = (
              await readFile("/path/to/reference.mp4")
            ).toString("base64");

            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "Follow the reference clip shot structure and motion. No text.",
            );
            form.append("seconds", "5");
            form.append("size", "1344x768");
            form.append(
              "reference_videos",
              `data:video/mp4;base64,${videoBase64}`,
            );

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: HTTPS MP3 reference audio with an image
          source: |
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string model=minimax-h3 \
              --form-string 'prompt=Animate the reference image and synchronize the motion with the reference audio. No text.' \
              --form-string seconds=5 \
              --form-string size=1344x768 \
              --form-string reference_audios=https://your-media-host/reference.mp3 \
              --form-string input_reference=https://your-image-host/reference.png
        - lang: Python
          label: HTTPS MP3 reference audio with an image
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "Animate the reference image and synchronize the motion with the "
                        "reference audio. No text."
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1344x768")),
                ("reference_audios", (None, "https://your-media-host/reference.mp3")),
                ("input_reference", (None, "https://your-image-host/reference.png")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )

            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: HTTPS MP3 reference audio with an image
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3");

            form.append(
              "prompt",
              "Animate the reference image and synchronize the motion with the " +
              "reference audio. No text.",
            );

            form.append("seconds", "5");

            form.append("size", "1344x768");

            form.append("reference_audios",
            "https://your-media-host/reference.mp3");

            form.append("input_reference",
            "https://your-image-host/reference.png");


            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: HTTPS first and last frames
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=Move naturally from the opening frame to the ending frame. No text.' \
              --form-string 'seconds=5' \
              --form-string 'size=1344x768' \
              --form-string 'first_frame=https://your-image-host/first-frame.png' \
              --form-string 'last_frame=https://your-image-host/last-frame.png'
        - lang: Python
          label: HTTPS first and last frames
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3")),
                (
                    "prompt",
                    (
                        None,
                        "Move naturally from the opening frame to the ending "
                        "frame. No text.",
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1344x768")),
                (
                    "first_frame",
                    (None, "https://your-image-host/first-frame.png"),
                ),
                (
                    "last_frame",
                    (None, "https://your-image-host/last-frame.png"),
                ),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: HTTPS first and last frames
          source: |
            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "Move naturally from the opening frame to the ending frame. No text.",
            );
            form.append("seconds", "5");
            form.append("size", "1344x768");
            form.append(
              "first_frame",
              "https://your-image-host/first-frame.png",
            );
            form.append(
              "last_frame",
              "https://your-image-host/last-frame.png",
            );

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: Uploaded first and last frames
          source: |-
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=Move naturally from the uploaded opening frame to the uploaded ending frame. No text.' \
              --form-string 'seconds=5' \
              --form-string 'size=1344x768' \
              --form 'first_frame=@/path/to/first-frame.png;type=image/png' \
              --form 'last_frame=@/path/to/last-frame.png;type=image/png'
        - lang: Python
          label: Uploaded first and last frames
          source: |
            import os
            from contextlib import ExitStack

            import requests

            prompt = (
                "Move naturally from the uploaded opening frame to the uploaded "
                "ending frame. No text."
            )

            with ExitStack() as stack:
                files = {
                    "first_frame": (
                        "first-frame.png",
                        stack.enter_context(
                            open("/path/to/first-frame.png", "rb")
                        ),
                        "image/png",
                    ),
                    "last_frame": (
                        "last-frame.png",
                        stack.enter_context(
                            open("/path/to/last-frame.png", "rb")
                        ),
                        "image/png",
                    ),
                }
                response = requests.post(
                    "https://api.cometapi.com/v1/videos",
                    headers={
                        "Authorization": "Bearer "
                        + os.environ["COMETAPI_KEY"]
                    },
                    data={
                        "model": "minimax-h3",
                        "prompt": prompt,
                        "seconds": "5",
                        "size": "1344x768",
                    },
                    files=files,
                    timeout=120,
                )

            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: Uploaded first and last frames
          source: |
            import { readFile } from "node:fs/promises";

            const firstFrame = await readFile("/path/to/first-frame.png");
            const lastFrame = await readFile("/path/to/last-frame.png");

            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "Move naturally from the uploaded opening frame to the uploaded " +
                "ending frame. No text.",
            );
            form.append("seconds", "5");
            form.append("size", "1344x768");
            form.append(
              "first_frame",
              new Blob([firstFrame], { type: "image/png" }),
              "first-frame.png",
            );
            form.append(
              "last_frame",
              new Blob([lastFrame], { type: "image/png" }),
              "last-frame.png",
            );

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: Uploaded MP4 reference video
          source: |
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3' \
              --form-string 'prompt=Follow the uploaded reference video shot structure and motion. No text.' \
              --form-string 'seconds=5' \
              --form-string 'size=1344x768' \
              --form 'reference_videos=@/path/to/reference.mp4;type=video/mp4'
        - lang: Python
          label: Uploaded MP4 reference video
          source: |
            import os
            from contextlib import ExitStack

            import requests

            prompt = (
                "Follow the uploaded reference video shot structure and motion. No "
                "text."
            )

            with ExitStack() as stack:
                fields = [
                    ("model", (None, "minimax-h3")),
                    ("prompt", (None, prompt)),
                    ("seconds", (None, "5")),
                    ("size", (None, "1344x768")),
                    (
                        "reference_videos",
                        (
                            "reference.mp4",
                            stack.enter_context(
                                open("/path/to/reference.mp4", "rb")
                            ),
                            "video/mp4",
                        ),
                    ),
                ]

                response = requests.post(
                    "https://api.cometapi.com/v1/videos",
                    headers={
                        "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                    },
                    files=fields,
                    timeout=120,
                )

            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: Uploaded MP4 reference video
          source: |
            import { readFile } from "node:fs/promises";

            const form = new FormData();
            form.append("model", "minimax-h3");
            form.append(
              "prompt",
              "Follow the uploaded reference video shot structure and motion. No " +
              "text.",
            );
            form.append("seconds", "5");
            form.append("size", "1344x768");
            form.append(
              "reference_videos",
              new Blob([await readFile("/path/to/reference.mp4")], {
                type: "video/mp4",
              }),
              "reference.mp4",
            );

            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });

            if (!response.ok) {
              throw new Error(await response.text());
            }

            console.log(await response.json());
        - lang: Shell
          label: Uploaded MP3 reference audio with an image
          source: |
            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string model=minimax-h3 \
              --form-string 'prompt=Animate the reference image and synchronize the motion with the uploaded reference audio. No text.' \
              --form-string seconds=5 \
              --form-string size=1344x768 \
              --form 'reference_audios=@/path/to/reference.mp3;type=audio/mpeg' \
              --form-string input_reference=https://your-image-host/reference.png
        - lang: Python
          label: Uploaded MP3 reference audio with an image
          source: |
            import os
            from contextlib import ExitStack

            import requests

            with ExitStack() as stack:
                fields = [
                    ("model", (None, "minimax-h3")),
                    (
                        "prompt",
                        (
                            None,
                            "Animate the reference image and synchronize the motion with "
                            "the uploaded reference audio. No text."
                        ),
                    ),
                    ("seconds", (None, "5")),
                    ("size", (None, "1344x768")),
                    (
                        "reference_audios",
                        (
                            "reference.mp3",
                            stack.enter_context(open("/path/to/reference.mp3", "rb")),
                            "audio/mpeg",
                        ),
                    ),
                    ("input_reference", (None, "https://your-image-host/reference.png")),
                ]

                response = requests.post(
                    "https://api.cometapi.com/v1/videos",
                    headers={
                        "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                    },
                    files=fields,
                    timeout=120,
                )

            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: Uploaded MP3 reference audio with an image
          source: >
            import { readFile } from "node:fs/promises";


            const form = new FormData();

            form.append("model", "minimax-h3");

            form.append(
              "prompt",
              "Animate the reference image and synchronize the motion with the " +
              "uploaded reference audio. No text.",
            );

            form.append("seconds", "5");

            form.append("size", "1344x768");

            form.append(
              "reference_audios",
              new Blob([await readFile("/path/to/reference.mp3")], {
                type: "audio/mpeg",
              }),
              "reference.mp3",
            );

            form.append("input_reference",
            "https://your-image-host/reference.png");


            const response = await fetch("https://api.cometapi.com/v1/videos", {
              method: "POST",
              headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
              body: form,
            });


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: 'H3-Max: text-to-video at 480P'
          source: |
            PROMPT='A paper boat crosses a pond. Fixed camera. No '
            PROMPT="$PROMPT"'text.'

            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3-max' \
              --form-string "prompt=$PROMPT" \
              --form-string 'seconds=5' \
              --form-string 'size=864x480'
        - lang: Python
          label: 'H3-Max: text-to-video at 480P'
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3-max")),
                ("prompt", (None, "A paper boat crosses a pond. Fixed camera. No text.")),
                ("seconds", (None, "5")),
                ("size", (None, "864x480")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: 'H3-Max: text-to-video at 480P'
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3-max");

            form.append("prompt", "A paper boat crosses a pond. Fixed camera. No
            text.");

            form.append("seconds", "5");

            form.append("size", "864x480");


            const response = await fetch(
              "https://api.cometapi.com/v1/videos",
              {
                method: "POST",
                headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
                body: form,
              },
            );


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: 'H3-Max: text-to-video at 768P'
          source: |
            PROMPT='A paper boat crosses a pond. Fixed camera. No '
            PROMPT="$PROMPT"'text.'

            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3-max' \
              --form-string "prompt=$PROMPT" \
              --form-string 'seconds=5' \
              --form-string 'size=1360x768'
        - lang: Python
          label: 'H3-Max: text-to-video at 768P'
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3-max")),
                ("prompt", (None, "A paper boat crosses a pond. Fixed camera. No text.")),
                ("seconds", (None, "5")),
                ("size", (None, "1360x768")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: 'H3-Max: text-to-video at 768P'
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3-max");

            form.append("prompt", "A paper boat crosses a pond. Fixed camera. No
            text.");

            form.append("seconds", "5");

            form.append("size", "1360x768");


            const response = await fetch(
              "https://api.cometapi.com/v1/videos",
              {
                method: "POST",
                headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
                body: form,
              },
            );


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: 'H3-Max: HTTPS reference images'
          source: |
            PROMPT='The two reference characters walk through a '
            PROMPT="$PROMPT"'park. No text.'
            IMAGE_URL_1="https://your-image-host/reference-1.png"
            IMAGE_URL_2="https://your-image-host/reference-2.jpg"

            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3-max' \
              --form-string "prompt=$PROMPT" \
              --form-string 'seconds=5' \
              --form-string 'size=1360x768' \
              --form-string "input_reference=$IMAGE_URL_1" \
              --form-string "input_reference=$IMAGE_URL_2"
        - lang: Python
          label: 'H3-Max: HTTPS reference images'
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3-max")),
                (
                    "prompt",
                    (
                        None,
                        "The two reference characters walk through a park. No text."
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1360x768")),
                ("input_reference", (None, "https://your-image-host/reference-1.png")),
                ("input_reference", (None, "https://your-image-host/reference-2.jpg")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: 'H3-Max: HTTPS reference images'
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3-max");

            form.append(
              "prompt",
              "The two reference characters walk through a park. No text.",
            );

            form.append("seconds", "5");

            form.append("size", "1360x768");

            form.append("input_reference",
            "https://your-image-host/reference-1.png");

            form.append("input_reference",
            "https://your-image-host/reference-2.jpg");


            const response = await fetch(
              "https://api.cometapi.com/v1/videos",
              {
                method: "POST",
                headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
                body: form,
              },
            );


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: 'H3-Max: image, video, and audio references'
          source: |
            PROMPT='Use the image for the subject, video for motion, '
            PROMPT="$PROMPT"'and audio for sound.'
            IMAGE_URL="https://your-image-host/reference.png"
            VIDEO_URL="https://your-media-host/reference.mp4"
            AUDIO_URL="https://your-media-host/reference.mp3"

            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3-max' \
              --form-string "prompt=$PROMPT" \
              --form-string 'seconds=5' \
              --form-string 'size=1360x768' \
              --form-string "input_reference=$IMAGE_URL" \
              --form-string "reference_videos=$VIDEO_URL" \
              --form-string "reference_audios=$AUDIO_URL"
        - lang: Python
          label: 'H3-Max: image, video, and audio references'
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3-max")),
                (
                    "prompt",
                    (
                        None,
                        "Use the image for the subject, video for motion, and audio "
                        "for sound."
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1360x768")),
                ("input_reference", (None, "https://your-image-host/reference.png")),
                ("reference_videos", (None, "https://your-media-host/reference.mp4")),
                ("reference_audios", (None, "https://your-media-host/reference.mp3")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: 'H3-Max: image, video, and audio references'
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3-max");

            form.append(
              "prompt",
              "Use the image for the subject, video for motion, and audio for sound.",
            );

            form.append("seconds", "5");

            form.append("size", "1360x768");

            form.append("input_reference",
            "https://your-image-host/reference.png");

            form.append("reference_videos",
            "https://your-media-host/reference.mp4");

            form.append("reference_audios",
            "https://your-media-host/reference.mp3");


            const response = await fetch(
              "https://api.cometapi.com/v1/videos",
              {
                method: "POST",
                headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
                body: form,
              },
            );


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
        - lang: Shell
          label: 'H3-Max: first and last frames'
          source: |
            PROMPT='Move from the opening frame to the ending frame. '
            PROMPT="$PROMPT"'No text.'
            FIRST_FRAME_URL="https://your-image-host/first-frame.png"
            LAST_FRAME_URL="https://your-image-host/last-frame.png"

            curl \
              https://api.cometapi.com/v1/videos \
              -H "Authorization: Bearer $COMETAPI_KEY" \
              --form-string 'model=minimax-h3-max' \
              --form-string "prompt=$PROMPT" \
              --form-string 'seconds=5' \
              --form-string 'size=1360x768' \
              --form-string "first_frame=$FIRST_FRAME_URL" \
              --form-string "last_frame=$LAST_FRAME_URL"
        - lang: Python
          label: 'H3-Max: first and last frames'
          source: |
            import os

            import requests

            fields = [
                ("model", (None, "minimax-h3-max")),
                (
                    "prompt",
                    (
                        None,
                        "Move from the opening frame to the ending frame. No text."
                    ),
                ),
                ("seconds", (None, "5")),
                ("size", (None, "1360x768")),
                ("first_frame", (None, "https://your-image-host/first-frame.png")),
                ("last_frame", (None, "https://your-image-host/last-frame.png")),
            ]

            response = requests.post(
                "https://api.cometapi.com/v1/videos",
                headers={
                    "Authorization": "Bearer " + os.environ["COMETAPI_KEY"]
                },
                files=fields,
                timeout=120,
            )
            response.raise_for_status()
            print(response.json())
        - lang: JavaScript
          label: 'H3-Max: first and last frames'
          source: >
            const form = new FormData();

            form.append("model", "minimax-h3-max");

            form.append(
              "prompt",
              "Move from the opening frame to the ending frame. No text.",
            );

            form.append("seconds", "5");

            form.append("size", "1360x768");

            form.append("first_frame",
            "https://your-image-host/first-frame.png");

            form.append("last_frame", "https://your-image-host/last-frame.png");


            const response = await fetch(
              "https://api.cometapi.com/v1/videos",
              {
                method: "POST",
                headers: { Authorization: `Bearer ${process.env.COMETAPI_KEY}` },
                body: form,
              },
            );


            if (!response.ok) {
              throw new Error(await response.text());
            }


            console.log(await response.json());
components:
  schemas:
    MiniMaxH3CreateRequest:
      type: object
      required:
        - model
        - prompt
      properties:
        model:
          type: string
          default: minimax-h3
          description: >-
            Model ID for the video task. Use minimax-h3 for 768P or 2K, or
            minimax-h3-max for 480P or 768P. Choose size from the matching model
            table.
          enum:
            - minimax-h3
            - minimax-h3-max
        prompt:
          type: string
          minLength: 1
          description: >-
            Text that describes the scene, motion, camera behavior, and visual
            details that the video should preserve.
          default: >-
            A paper boat glides across a still pond, daylight, locked camera, no
            text.
        seconds:
          type: integer
          minimum: 5
          maximum: 15
          default: 5
          description: >-
            Requested clip duration in whole seconds. Use an integer from 5
            through 15. The default is 5.
        size:
          type: string
          pattern: ^[1-9][0-9]*x[1-9][0-9]*$
          examples:
            - 1536x672
            - 1344x768
            - 1024x768
            - 768x768
            - 768x1024
            - 768x1344
            - 2912x1280
            - 2544x1456
            - 1920x1440
            - 1440x1440
            - 1440x1920
            - 1440x2560
            - 1120x480
            - 864x480
            - 640x480
            - 480x480
            - 480x640
            - 480x864
            - 1792x768
            - 1360x768
            - 768x1360
          description: >-
            Requested size in WxH form. For text and reference-media requests,
            size selects the resolution tier and framing aspect ratio. With
            first or last frames, size selects the resolution tier. Use the
            matching model table: minimax-h3 uses 768P or 2K; minimax-h3-max
            uses 480P or 768P. For example, use 1344x768 for H3 at 768P, 864x480
            for H3-Max at 480P, or 1360x768 for H3-Max at 768P. The service can
            normalize the final output dimensions.
        input_reference:
          description: >-
            PNG or JPEG reference images. Repeat input_reference with one
            publicly accessible HTTPS image URL or uploaded image file per
            field. Send up to nine reference images, whether you use URLs or
            uploaded files. Use all URLs or all files within this field. Do not
            combine URLs into a comma-separated value. Keep each uploaded image
            at or below 20 MiB. Do not combine this field with first_frame or
            last_frame. Across images, videos, and audio, send at most 12
            reference inputs. Each reference video or audio clip must be 2–15
            seconds; keep the combined video duration and combined audio
            duration at or below 15 seconds each. Submit images in the order you
            want to reference them in the prompt.
          anyOf:
            - title: Single image URL
              type: string
              format: uri
              pattern: ^https://
              description: One publicly accessible HTTPS image URL.
            - title: Image URLs
              type: array
              minItems: 1
              maxItems: 9
              items:
                type: string
                format: uri
                pattern: ^https://
                description: One publicly accessible HTTPS reference image URL.
            - title: Image files
              type: array
              minItems: 1
              maxItems: 9
              items:
                type: string
                format: binary
                description: One uploaded reference image file.
        reference_videos:
          type: array
          minItems: 1
          maxItems: 3
          items:
            anyOf:
              - type: string
                format: uri
                pattern: ^(https://|data:video/mp4;base64,)
                example: https://your-media-host/reference.mp4
              - type: string
                format: binary
          description: >-
            Reference video inputs. Repeat reference_videos with a direct
            publicly accessible HTTPS MP4 URL or an uploaded MP4 file. Use only
            text values or only files within this field. Each uploaded file can
            be up to 50 MiB. Send no more than three video inputs. A complete
            data:video/mp4;base64,... URI is also accepted as a text value. Use
            MP4 with H.264 video; an AAC audio track is optional. Do not combine
            this field with first_frame or last_frame. Across images, videos,
            and audio, send at most 12 reference inputs. Each reference video or
            audio clip must be 2–15 seconds; keep the combined video duration
            and combined audio duration at or below 15 seconds each.
        reference_audios:
          type: array
          minItems: 1
          maxItems: 3
          items:
            anyOf:
              - type: string
                format: uri
                pattern: ^https://
                example: https://your-media-host/reference.mp3
              - type: string
                format: binary
          description: >-
            Reference audio inputs. Repeat reference_audios with a direct
            publicly accessible HTTPS MP3 URL or an uploaded MP3 file. Use only
            text values or only files within this field. Each uploaded file can
            be up to 15 MiB. Send no more than three audio inputs. Pair
            reference audio with at least one reference image. Do not combine
            this field with first_frame or last_frame. Across images, videos,
            and audio, send at most 12 reference inputs. Each reference video or
            audio clip must be 2–15 seconds; keep the combined video duration
            and combined audio duration at or below 15 seconds each.
        first_frame:
          description: >-
            Opening frame input as one direct publicly accessible HTTPS PNG URL
            or one uploaded PNG file at or below 20 MiB (20 x 1024 x 1024
            bytes). Do not combine frame fields with reference media fields.
          anyOf:
            - type: string
              format: uri
              pattern: ^https://
              example: https://your-image-host/first-frame.png
            - type: string
              format: binary
        last_frame:
          description: >-
            Optional ending frame input, used with first_frame, as one direct
            publicly accessible HTTPS PNG URL or one uploaded PNG file at or
            below 20 MiB (20 x 1024 x 1024 bytes). Do not combine frame fields
            with reference media fields.
          anyOf:
            - type: string
              format: uri
              pattern: ^https://
              example: https://your-image-host/last-frame.png
            - type: string
              format: binary
      dependentRequired:
        last_frame:
          - first_frame
      additionalProperties: false
    MiniMaxH3VideoTask:
      type: object
      required:
        - id
        - object
        - model
        - status
        - progress
        - created_at
      properties:
        id:
          type: string
          description: Task ID. Use this value as task_id in retrieve and content requests.
          example: <task_id>
        task_id:
          type: string
          description: >-
            Compatibility alias for id. This field can be omitted from retrieve
            responses.
          example: <task_id>
        object:
          type: string
          const: video
          description: Object type for the asynchronous video task.
        model:
          type: string
          description: Model ID that the task uses.
          enum:
            - minimax-h3
            - minimax-h3-max
        status:
          type: string
          enum:
            - queued
            - in_progress
            - completed
            - failed
          description: Task lifecycle status. Poll until the value is completed or failed.
        progress:
          type: integer
          minimum: 0
          maximum: 100
          description: Task progress as a coarse percentage.
        created_at:
          type: integer
          format: int64
          description: Task creation time as a Unix timestamp in seconds.
        completed_at:
          type: integer
          format: int64
          description: >-
            Unix timestamp returned by the platform. Use status, not this field,
            to decide when polling can stop.
        expires_at:
          type: integer
          format: int64
          description: >-
            Result expiration time as a Unix timestamp in seconds when the task
            provides one.
        video_url:
          type: string
          format: uri
          description: Video delivery URL. This field appears on completed tasks.
          example: https://your-media-host/result.mp4
        error:
          type: object
          description: Failure details. This field appears when the task fails.
          properties:
            message:
              type: string
              description: Human-readable failure description.
            code:
              type: string
              description: Failure code when the task provides one.
          additionalProperties: true
      additionalProperties: true
  securitySchemes:
    bearerAuth:
      type: http
      scheme: bearer
      description: Bearer authentication. Use your CometAPI API key.

````