> ## Documentation Index
> Fetch the complete documentation index at: https://docs.aiid.edu.kg/llms.txt
> Use this file to discover all available pages before exploring further.

# Create video

> OpenAI Sora video generation API in a compatible format. Supports text-to-video and image/video reference generation modes.

- `POST /v1/videos`: Create video task
- `GET /v1/videos/{video_id}`: Query task status
- `GET /v1/videos/{video_id}/content`: Retrieve the video binary content

Currently not available externally:
- `GET /v1/videos`
- `POST /v1/videos/{video_id}/remix`

Standard authentication headers:
- `Content-Type: application/json`

1. Use `POST /v1/videos` to create a task and obtain `id`
2. Poll `GET /v1/videos/{id}` until `completed` or `failed`
3. After completion, prefer using `video_url` returned in the response
4. For unified download, call `GET /v1/videos/{id}/content`

Core parameters:
- `model: string`: Required, public model name
- `prompt: string`: Recommended, video generation description
- `seconds: number|string`: Optional, target duration
- `duration: number|string`: Optional, duration alias
- `size: string`: Optional, output resolution
- `mode: string`: Optional, common values: `t2v` `i2v` `i2v_first_last` `reference_material`

Common compatibility parameters:
- `aspect_ratio`
- `ratio`
- `quality`
- `resolution`
- `fps`
- `image`
- `image_url`
- `image_urls`
- `images`
- `reference_images`
- `input_reference`
- `end_image_url`
- `last_image_url`
- `video_urls`
- `audio_urls`
- `function_mode`
- `content`
- `callback_url`
- `external_task_id`

Kling common extended parameters:
- `model_name`
- `negative_prompt`
- `cfg_scale`
- `sound`
- `camera_control`
- `image_list`
- `video_id`
- `task_id`
- `watermark_info`

### 4.1 Veo series

Note: `stable`, `official`, and similar labels indicate different groups, not external `model` suffixes; when calling, keep `model` as the base model name, and let the account or token group determine which group is actually used.

#### Veo 3.x main model

- `veo3`
- `veo3-fast`
- `veo3-fast-frames`
- `veo3-frames`
- `veo3-pro`
- `veo3-pro-frames`
- `veo3.1`
- `veo3.1-fast`
- `veo3.1-pro`
- `veo3.1-components`
- `veo3.1-4k`
- `veo3.1-pro-4k`

Recommendations:
- For text-to-video only, prefer the `veo3*` / `veo3.1*` base models
- For image-to-video or reference-to-video scenarios, it is recommended to explicitly pass `mode + image_url/reference_images`

### 4.2 Sora series

Base models:
- `sora-2`
- `sora-2-pro`

Recommendations:
- `sora-2*` is suitable for general-purpose use
- To use groups such as `stable` / `official`, keep `model` as the base model name and do not append any suffix to the model name.

Request example:

```json
{
  "size": "1280x720",
  "images": [
    "https://example.png"
  ],
  "model": "sora-2",
  "prompt": "this is an example",
  "duration": 12,
  "fps": "24",
  "seed": "20231024",
  "mode": "i2v"
}
```

### 4.3 HappyHorse series

Public model:
- `happyhorse-1.0`

Recommendations:
- `happyhorse-1.0` are distinguished by `mode` into `t2v`, `i2v`, `r2v`, and `video_edit`.
- In Sora-compatible format, it is recommended to pass media items in the `content` array for reference images.
- Use `type=text` for text items and `type=image_url` for media items, and optionally use `role` and `name` to indicate the asset's purpose.

Request example:

```json
{
  "size": "720x1280",
  "model": "happyhorse-1.0",
  "mode": "r2v",
  "prompt": "this is an example",
  "duration": 15,
  "content": [
    {
      "type": "text",
      "text": "this is an example"
    },
    {
      "type": "image_url",
      "image_url": {
        "url": "https://example.png"
      },
      "role": "reference_image",
      "name": "image1"
    }
  ],
  "parameters": {
    "resolution": "720P",
    "ratio": "9:16",
    "duration": 15,
    "seed": 1
  }
}
```

### 4.4 Seedance Series

Public model:
- `doubao-seedance-1-0-lite-t2v-250428`
- `doubao-seedance-1-0-lite-i2v-250428`
- `doubao-seedance-1-0-pro-250528`
- `doubao-seedance-1-0-pro-fast-251015`
- `doubao-seedance-1-5-pro-251215`
- `doubao-seedance-2-0-260128`
- `doubao-seedance-2-0-fast-260128`

Recommendations:
- For text-to-video, prefer `*-t2v-*` or `pro / fast`
- For image-to-video, prefer `*-i2v-*`
- For scenarios with complex reference materials, prefer `content`

Request example:

```json
{
  "size": "1280x720",
  "model": "doubao-seedance-2-0-fast-260128",
  "prompt": "this is an example",
  "duration": 15,
  "fps": "24",
  "seed": "20231024",
  "mode": "reference_material",
  "content": [
    {
      "type": "text",
      "text": "this is an example"
    },
    {
      "type": "image_url",
      "image_url": {
        "url": "https://example.png"
      },
      "role": "reference_image",
      "name": "image1"
    },
    {
      "type": "audio_url",
      "audio_url": {
        "url": "https://example.mp3"
      },
      "role": "reference_audio",
      "name": "audio1"
    }
  ]
}
```

### 4.5 Grok Video Series

Public model:
- `grok-imagine-1.0-video`
- `grok-imagine-video-1.5-preview`
- `grok-video-3`

Common parameters:
- `prompt`
- `ratio` / `aspect_ratio`
- `resolution` / `size`
- `seconds` / `duration`
- `image` / `image_url` / `input_reference`
- `reference_images`

Example:

```json
{
  "model": "grok-imagine-1.0-video",
  "prompt": "雨夜霓虹街道上的电影感推镜，光影丰富，运动自然",
  "reference_images": [
    "https://example.com/ref-1.jpg"
  ],
  "seconds": 10,
  "aspect_ratio": "16:9",
  "resolution": "720P"
}
```

### 4.6 Kling Video Main Model

Public base model:
- `kling-video`

Required:
- `model`
- `model_name`

Supported `model_name`:
- `kling-v1`
- `kling-v1-5`
- `kling-v1-6`
- `kling-v2-master`
- `kling-v2-1`
- `kling-v2-1-master`
- `kling-v2-5-turbo`
- `kling-v2-6`
- `kling-v3`

Common `mode`:
- `t2v`
- `i2v`
- `multi_i2v`
- `extend`

Minimum input parameters:
- Text-to-video: `model + model_name + prompt + mode=t2v`
- Image-to-video: `model + model_name + prompt + mode=i2v + image`
- Multi-image reference: `model + model_name + prompt + mode=multi_i2v + image_list`
- Video extension: `model + model_name + mode=extend + video_id`

Example:

```json
{
  "model": "kling-video",
  "model_name": "kling-v2-6",
  "mode": "t2v",
  "prompt": "海边日落镜头，电影感，风吹长发",
  "duration": 5,
  "aspect_ratio": "16:9"
}
```

### 4.7 MiniMax / Hailuo Video Series

Public model:
- `minimax-h3`
- `hailuo-2.3`

Recommendations:
- `minimax-h3` supports text-to-video, image references, and audio references. We recommend mixing text and assets via `content`; you can also use compatible fields such as `image_urls` and `audio_urls`.
- `hailuo-2.3` requires a first-frame image. We recommend passing the `image_url` field of `role=first_frame` in `content`, or using `image_urls` to pass the first-frame image.
- `minimax-h3` commonly uses `duration=5~15` and `resolution=1440p`; `hailuo-2.3` commonly uses `duration=6` or `10` and `resolution=768p`.

`minimax-h3` Example:

```json
{
  "model": "minimax-h3",
  "prompt": "A cinematic product video with smooth camera motion",
  "content": [
    {
      "type": "text",
      "text": "Use the product image as reference and match the rhythm of the audio."
    },
    {
      "type": "image_url",
      "image_url": {
        "url": "https://example.com/product.jpg"
      },
      "role": "reference_image",
      "name": "product"
    },
    {
      "type": "audio_url",
      "audio_url": {
        "url": "https://example.com/music.mp3"
      },
      "role": "reference_audio",
      "name": "music"
    }
  ],
  "duration": 8,
  "resolution": "1440p",
  "aspect_ratio": "16:9",
  "generate_audio": true
}
```

`hailuo-2.3` Example:

```json
{
  "model": "hailuo-2.3",
  "prompt": "The camera slowly pushes in, natural motion and cinematic lighting.",
  "content": [
    {
      "type": "text",
      "text": "The camera slowly pushes in, natural motion and cinematic lighting."
    },
    {
      "type": "image_url",
      "image_url": {
        "url": "https://example.com/first-frame.jpg"
      },
      "role": "first_frame",
      "name": "first"
    }
  ],
  "duration": 10,
  "resolution": "768p",
  "aspect_ratio": "16:9"
}
```

- `mode=t2v`
- Minimum input: `model + prompt`

- `mode=i2v`
- Minimum input: `model + prompt + image_url`

- `mode=i2v_first_last`
- Minimum input: `model + prompt + image_url + end_image_url`

- `mode=reference_images`
- Minimum input: `model + prompt + reference_images`

- `mode=reference_material`
- Minimum input: `model + prompt + (image_urls/video_urls/audio_urls 至少一种)`

### Gemini Omni usage instructions

- `gemini-omni` is the publicly exposed video model name of new-api, and can be called directly via `POST /v1/videos`.
- `mode=t2v`: text-to-video, with the minimum input parameter being `model + prompt`.
- `mode=r2v`: reference image/reference material generation; images can be placed in `image`, `image_url`, `images`, `image_urls`, `reference_images`, `input_reference`, or `content`.
- `mode=edit`: video editing; videos can be placed in `video`, `video_url`, `videos`, or `content`, and reference images can also be provided at the same time.
- The duration field can use `seconds` or `duration`, and will be automatically mapped to the 4 / 6 / 8 / 10 second options.



## OpenAPI

````yaml api-reference/en/openapi.json POST /v1/videos
openapi: 3.0.0
info:
  title: Overseas Expansion Camp API Reference Documentation
  version: 1.0.0
  description: Public AI Gateway API Reference
servers:
  - url: https://api.aiid.edu.kg
security:
  - BearerAuth: []
tags:
  - name: OpenAI format (Chat)
  - name: OpenAI Format (Responses)
  - name: Gemini Format for Image Generation
  - name: Image generationOpenAI DALL-E format
  - name: List models
  - name: Video generationHappyHorse and Wan
  - name: Video generation Kling format
  - name: Video generation Omni and Veo formats
  - name: Video generation Seedance
  - name: Video generation Sora compatible formats
  - name: Video generation Vidu
  - name: Music generation task format
paths:
  /v1/videos:
    post:
      tags:
        - Video generation Sora compatible formats
      summary: Create video
      description: >-
        OpenAI Sora video generation API in a compatible format. Supports
        text-to-video and image/video reference generation modes.


        - `POST /v1/videos`: Create video task

        - `GET /v1/videos/{video_id}`: Query task status

        - `GET /v1/videos/{video_id}/content`: Retrieve the video binary content


        Currently not available externally:

        - `GET /v1/videos`

        - `POST /v1/videos/{video_id}/remix`


        Standard authentication headers:

        - `Content-Type: application/json`


        1. Use `POST /v1/videos` to create a task and obtain `id`

        2. Poll `GET /v1/videos/{id}` until `completed` or `failed`

        3. After completion, prefer using `video_url` returned in the response

        4. For unified download, call `GET /v1/videos/{id}/content`


        Core parameters:

        - `model: string`: Required, public model name

        - `prompt: string`: Recommended, video generation description

        - `seconds: number|string`: Optional, target duration

        - `duration: number|string`: Optional, duration alias

        - `size: string`: Optional, output resolution

        - `mode: string`: Optional, common values: `t2v` `i2v` `i2v_first_last`
        `reference_material`


        Common compatibility parameters:

        - `aspect_ratio`

        - `ratio`

        - `quality`

        - `resolution`

        - `fps`

        - `image`

        - `image_url`

        - `image_urls`

        - `images`

        - `reference_images`

        - `input_reference`

        - `end_image_url`

        - `last_image_url`

        - `video_urls`

        - `audio_urls`

        - `function_mode`

        - `content`

        - `callback_url`

        - `external_task_id`


        Kling common extended parameters:

        - `model_name`

        - `negative_prompt`

        - `cfg_scale`

        - `sound`

        - `camera_control`

        - `image_list`

        - `video_id`

        - `task_id`

        - `watermark_info`


        ### 4.1 Veo series


        Note: `stable`, `official`, and similar labels indicate different
        groups, not external `model` suffixes; when calling, keep `model` as the
        base model name, and let the account or token group determine which
        group is actually used.


        #### Veo 3.x main model


        - `veo3`

        - `veo3-fast`

        - `veo3-fast-frames`

        - `veo3-frames`

        - `veo3-pro`

        - `veo3-pro-frames`

        - `veo3.1`

        - `veo3.1-fast`

        - `veo3.1-pro`

        - `veo3.1-components`

        - `veo3.1-4k`

        - `veo3.1-pro-4k`


        Recommendations:

        - For text-to-video only, prefer the `veo3*` / `veo3.1*` base models

        - For image-to-video or reference-to-video scenarios, it is recommended
        to explicitly pass `mode + image_url/reference_images`


        ### 4.2 Sora series


        Base models:

        - `sora-2`

        - `sora-2-pro`


        Recommendations:

        - `sora-2*` is suitable for general-purpose use

        - To use groups such as `stable` / `official`, keep `model` as the base
        model name and do not append any suffix to the model name.


        Request example:


        ```json

        {
          "size": "1280x720",
          "images": [
            "https://example.png"
          ],
          "model": "sora-2",
          "prompt": "this is an example",
          "duration": 12,
          "fps": "24",
          "seed": "20231024",
          "mode": "i2v"
        }

        ```


        ### 4.3 HappyHorse series


        Public model:

        - `happyhorse-1.0`


        Recommendations:

        - `happyhorse-1.0` are distinguished by `mode` into `t2v`, `i2v`, `r2v`,
        and `video_edit`.

        - In Sora-compatible format, it is recommended to pass media items in
        the `content` array for reference images.

        - Use `type=text` for text items and `type=image_url` for media items,
        and optionally use `role` and `name` to indicate the asset's purpose.


        Request example:


        ```json

        {
          "size": "720x1280",
          "model": "happyhorse-1.0",
          "mode": "r2v",
          "prompt": "this is an example",
          "duration": 15,
          "content": [
            {
              "type": "text",
              "text": "this is an example"
            },
            {
              "type": "image_url",
              "image_url": {
                "url": "https://example.png"
              },
              "role": "reference_image",
              "name": "image1"
            }
          ],
          "parameters": {
            "resolution": "720P",
            "ratio": "9:16",
            "duration": 15,
            "seed": 1
          }
        }

        ```


        ### 4.4 Seedance Series


        Public model:

        - `doubao-seedance-1-0-lite-t2v-250428`

        - `doubao-seedance-1-0-lite-i2v-250428`

        - `doubao-seedance-1-0-pro-250528`

        - `doubao-seedance-1-0-pro-fast-251015`

        - `doubao-seedance-1-5-pro-251215`

        - `doubao-seedance-2-0-260128`

        - `doubao-seedance-2-0-fast-260128`


        Recommendations:

        - For text-to-video, prefer `*-t2v-*` or `pro / fast`

        - For image-to-video, prefer `*-i2v-*`

        - For scenarios with complex reference materials, prefer `content`


        Request example:


        ```json

        {
          "size": "1280x720",
          "model": "doubao-seedance-2-0-fast-260128",
          "prompt": "this is an example",
          "duration": 15,
          "fps": "24",
          "seed": "20231024",
          "mode": "reference_material",
          "content": [
            {
              "type": "text",
              "text": "this is an example"
            },
            {
              "type": "image_url",
              "image_url": {
                "url": "https://example.png"
              },
              "role": "reference_image",
              "name": "image1"
            },
            {
              "type": "audio_url",
              "audio_url": {
                "url": "https://example.mp3"
              },
              "role": "reference_audio",
              "name": "audio1"
            }
          ]
        }

        ```


        ### 4.5 Grok Video Series


        Public model:

        - `grok-imagine-1.0-video`

        - `grok-imagine-video-1.5-preview`

        - `grok-video-3`


        Common parameters:

        - `prompt`

        - `ratio` / `aspect_ratio`

        - `resolution` / `size`

        - `seconds` / `duration`

        - `image` / `image_url` / `input_reference`

        - `reference_images`


        Example:


        ```json

        {
          "model": "grok-imagine-1.0-video",
          "prompt": "雨夜霓虹街道上的电影感推镜，光影丰富，运动自然",
          "reference_images": [
            "https://example.com/ref-1.jpg"
          ],
          "seconds": 10,
          "aspect_ratio": "16:9",
          "resolution": "720P"
        }

        ```


        ### 4.6 Kling Video Main Model


        Public base model:

        - `kling-video`


        Required:

        - `model`

        - `model_name`


        Supported `model_name`:

        - `kling-v1`

        - `kling-v1-5`

        - `kling-v1-6`

        - `kling-v2-master`

        - `kling-v2-1`

        - `kling-v2-1-master`

        - `kling-v2-5-turbo`

        - `kling-v2-6`

        - `kling-v3`


        Common `mode`:

        - `t2v`

        - `i2v`

        - `multi_i2v`

        - `extend`


        Minimum input parameters:

        - Text-to-video: `model + model_name + prompt + mode=t2v`

        - Image-to-video: `model + model_name + prompt + mode=i2v + image`

        - Multi-image reference: `model + model_name + prompt + mode=multi_i2v +
        image_list`

        - Video extension: `model + model_name + mode=extend + video_id`


        Example:


        ```json

        {
          "model": "kling-video",
          "model_name": "kling-v2-6",
          "mode": "t2v",
          "prompt": "海边日落镜头，电影感，风吹长发",
          "duration": 5,
          "aspect_ratio": "16:9"
        }

        ```


        ### 4.7 MiniMax / Hailuo Video Series


        Public model:

        - `minimax-h3`

        - `hailuo-2.3`


        Recommendations:

        - `minimax-h3` supports text-to-video, image references, and audio
        references. We recommend mixing text and assets via `content`; you can
        also use compatible fields such as `image_urls` and `audio_urls`.

        - `hailuo-2.3` requires a first-frame image. We recommend passing the
        `image_url` field of `role=first_frame` in `content`, or using
        `image_urls` to pass the first-frame image.

        - `minimax-h3` commonly uses `duration=5~15` and `resolution=1440p`;
        `hailuo-2.3` commonly uses `duration=6` or `10` and `resolution=768p`.


        `minimax-h3` Example:


        ```json

        {
          "model": "minimax-h3",
          "prompt": "A cinematic product video with smooth camera motion",
          "content": [
            {
              "type": "text",
              "text": "Use the product image as reference and match the rhythm of the audio."
            },
            {
              "type": "image_url",
              "image_url": {
                "url": "https://example.com/product.jpg"
              },
              "role": "reference_image",
              "name": "product"
            },
            {
              "type": "audio_url",
              "audio_url": {
                "url": "https://example.com/music.mp3"
              },
              "role": "reference_audio",
              "name": "music"
            }
          ],
          "duration": 8,
          "resolution": "1440p",
          "aspect_ratio": "16:9",
          "generate_audio": true
        }

        ```


        `hailuo-2.3` Example:


        ```json

        {
          "model": "hailuo-2.3",
          "prompt": "The camera slowly pushes in, natural motion and cinematic lighting.",
          "content": [
            {
              "type": "text",
              "text": "The camera slowly pushes in, natural motion and cinematic lighting."
            },
            {
              "type": "image_url",
              "image_url": {
                "url": "https://example.com/first-frame.jpg"
              },
              "role": "first_frame",
              "name": "first"
            }
          ],
          "duration": 10,
          "resolution": "768p",
          "aspect_ratio": "16:9"
        }

        ```


        - `mode=t2v`

        - Minimum input: `model + prompt`


        - `mode=i2v`

        - Minimum input: `model + prompt + image_url`


        - `mode=i2v_first_last`

        - Minimum input: `model + prompt + image_url + end_image_url`


        - `mode=reference_images`

        - Minimum input: `model + prompt + reference_images`


        - `mode=reference_material`

        - Minimum input: `model + prompt + (image_urls/video_urls/audio_urls
        至少一种)`


        ### Gemini Omni usage instructions


        - `gemini-omni` is the publicly exposed video model name of new-api, and
        can be called directly via `POST /v1/videos`.

        - `mode=t2v`: text-to-video, with the minimum input parameter being
        `model + prompt`.

        - `mode=r2v`: reference image/reference material generation; images can
        be placed in `image`, `image_url`, `images`, `image_urls`,
        `reference_images`, `input_reference`, or `content`.

        - `mode=edit`: video editing; videos can be placed in `video`,
        `video_url`, `videos`, or `content`, and reference images can also be
        provided at the same time.

        - The duration field can use `seconds` or `duration`, and will be
        automatically mapped to the 4 / 6 / 8 / 10 second options.
      operationId: createVideo
      parameters: []
      requestBody:
        required: true
        content:
          application/json:
            schema:
              type: object
              properties:
                model:
                  type: string
                  description: >-
                    Required, external model name. `gemini-omni` Can be called
                    via `/v1/videos`.
                  enum:
                    - sora-2
                    - sora-2-pro
                    - gemini-omni
                    - happyhorse-1.0
                    - happyhorse-1.0-i2v
                    - happyhorse-1.0-t2v
                    - happyhorse-1.0-r2v
                    - happyhorse-1.0-video-edit
                    - doubao-seedance-1-0-lite-i2v-250428
                    - doubao-seedance-1-0-lite-t2v-250428
                    - doubao-seedance-1-0-pro-250528
                    - doubao-seedance-1-0-pro-fast-251015
                    - doubao-seedance-1-5-pro-251215
                    - doubao-seedance-2-0-260128
                    - doubao-seedance-2-0-fast-260128
                    - veo3
                    - veo3-fast
                    - veo3-fast-frames
                    - veo3-frames
                    - veo3-pro
                    - veo3-pro-frames
                    - veo3.1
                    - veo3.1-4k
                    - veo3.1-components
                    - veo3.1-fast
                    - veo3.1-pro
                    - veo3.1-pro-4k
                    - kling-video
                    - grok-imagine-1.0-video
                    - grok-imagine-video-1.5-preview
                    - grok-video-3
                    - minimax-h3
                    - hailuo-2.3
                    - viduq3-turbo
                  example: sora-2
                prompt:
                  type: string
                  description: Unified prompt entry point. Required for most models.
                image:
                  type: string
                  format: binary
                  description: >-
                    Single image entry point; will be mapped to `images` in
                    certain compatibility modes.
                duration:
                  type: integer
                  description: >-
                    Unified duration parameter. Some models also accept
                    `seconds`.
                width:
                  type: integer
                  example: 512
                  description: Video width
                height:
                  type: integer
                  example: 512
                  description: Video height
                fps:
                  type: integer
                  example: 30
                  description: Video frame rate
                seed:
                  type: integer
                  example: 20231234
                  description: Random seed
                'n':
                  type: integer
                  example: 1
                  description: Number of videos to generate
                response_format:
                  type: string
                  example: url
                  description: Response format
                user:
                  type: string
                  example: user-1234
                  description: User identifier
                metadata:
                  type: object
                  additionalProperties: true
                  description: >-
                    Dynamic extension field container. Numerous model-specific
                    fields are deserialized from here.
                mode:
                  type: string
                  description: >-
                    Optional, video generation mode. `gemini-omni` Supports
                    `t2v`, `r2v`, and `edit`.
                images:
                  type: array
                  description: >-
                    Unified multi-image entry point. Used by Vidu, Seedance,
                    Veo, etc.
                  items:
                    type: string
                content:
                  type: array
                  items:
                    type: object
                    additionalProperties: true
                size:
                  type: string
                  description: >-
                    Unified size input, which maps to resolution/aspect_ratio
                    and related fields.
                seconds:
                  type: string
                  description: >-
                    Sora compatibility entry point; at runtime, it falls back to
                    duration.
                input_reference:
                  type: string
                  description: >-
                    Sora/Veo compatible reference image input, can be a
                    multipart file or a compatible object.
                parameters:
                  type: object
                  additionalProperties: true
              required:
                - model
                - prompt
            examples:
              openai_i2v_json:
                summary: OpenAI Sora i2v (application/json)
                value:
                  size: 1280x720
                  images:
                    - https://example.png
                  model: sora-2
                  prompt: this is an example
                  duration: 12
                  fps: '24'
                  seed: '20231024'
                  mode: i2v
              openai_i2v_input_reference:
                summary: OpenAI i2v (input_reference)
                value:
                  model: sora-2
                  prompt: A cinematic drone shot over snowy mountains at sunrise
                  input_reference:
                    image_url: https://example.com/input.jpg
                  seconds: '8'
                  size: 1280x720
                  response_format: url
              openai_alias_compat:
                summary: Alias兼容 (image_url/duration/aspect_ratio)
                value:
                  model: sora-2
                  prompt: A cinematic drone shot over snowy mountains at sunrise
                  image_url: https://example.com/input.jpg
                  duration: '8'
                  aspect_ratio: '16:9'
                  response_format: url
              gemini_omni_sora_t2v:
                summary: Gemini Omni Sora t2v
                value:
                  model: gemini-omni
                  mode: t2v
                  prompt: A cinematic product video with smooth camera motion
                  duration: 8
                  size: 1280x720
              gemini_omni_sora_r2v:
                summary: Gemini Omni Sora r2v
                value:
                  model: gemini-omni
                  mode: r2v
                  content:
                    - type: text
                      text: Keep the product consistent and create a short ad video
                    - type: image_url
                      image_url:
                        url: https://example.com/product.jpg
                  duration: 8
                  size: 1280x720
              gemini_omni_sora_edit:
                summary: Gemini Omni Sora edit
                value:
                  model: gemini-omni
                  mode: edit
                  content:
                    - type: text
                      text: Change the outfit color while keeping the motion natural
                    - type: video_url
                      video_url:
                        url: https://example.com/input.mp4
                    - type: image_url
                      image_url:
                        url: https://example.com/style.jpg
                  duration: 8
                  size: 1280x720
              happyhorse_sora_t2v:
                summary: HappyHorse Sora t2v
                value:
                  model: happyhorse-1.0
                  mode: t2v
                  content:
                    - type: text
                      text: >-
                        A cinematic white horse running through a neon city
                        street
                  duration: 5
                  resolution: 720P
                  ratio: '16:9'
              happyhorse_sora_i2v:
                summary: HappyHorse Sora i2v
                value:
                  model: happyhorse-1.0
                  mode: i2v
                  content:
                    - type: text
                      text: The camera slowly pushes in, natural motion
                    - type: image_url
                      image_url: https://example.com/first-frame.jpg
                  duration: 5
                  resolution: 720P
              happyhorse_sora_r2v:
                summary: HappyHorse Sora r2v
                value:
                  size: 720x1280
                  model: happyhorse-1.0
                  mode: r2v
                  prompt: this is an example
                  duration: 15
                  content:
                    - type: text
                      text: this is an example
                    - type: image_url
                      image_url:
                        url: https://example.png
                      role: reference_image
                      name: image1
                  parameters:
                    resolution: 720P
                    ratio: '9:16'
                    duration: 15
                    seed: 1
              happyhorse_sora_video_edit:
                summary: HappyHorse Sora video_edit
                value:
                  model: happyhorse-1.0
                  mode: video_edit
                  content:
                    - type: text
                      text: >-
                        Change the jacket to a striped sweater and keep the
                        motion natural
                    - type: video_url
                      video_url: https://example.com/input.mp4
                    - type: image_url
                      image_url: https://example.com/reference.jpg
                  resolution: 720P
              seedance_i2v_first:
                summary: Seedance i2v_first
                value:
                  model: doubao-seedance-1-0-lite-i2v-250428
                  mode: i2v_first
                  prompt: A panda skateboarding in a city street, cinematic
                  image_urls:
                    - https://example.com/first.jpg
                  seconds: '5'
                  size: 1280x720
              seedance_i2v_first_last:
                summary: Seedance i2v_first_last
                value:
                  model: doubao-seedance-1-0-lite-i2v-250428
                  mode: i2v_first_last
                  prompt: A camera moves from first frame to last frame naturally
                  image_urls:
                    - https://example.com/first.jpg
                  end_image_url: https://example.com/last.jpg
                  seconds: '5'
                  size: 1280x720
              seedance_reference_images:
                summary: Seedance reference_images
                value:
                  model: doubao-seedance-1-0-lite-i2v-250428
                  mode: reference_images
                  prompt: Stylized short video keeping character consistency
                  reference_images:
                    - https://example.com/ref-1.jpg
                    - https://example.com/ref-2.jpg
                  seconds: '5'
                  size: 1280x720
              seedance_reference_material:
                summary: Seedance All-in-one reference mode (reference_material)
                value:
                  model: doubao-seedance-1-0-lite-i2v-250428
                  mode: reference_material
                  prompt: >-
                    Blend image and short reference video into a coherent
                    cinematic clip
                  reference_images:
                    - https://example.com/ref-1.jpg
                    - https://example.com/ref-2.jpg
                  video_urls:
                    - https://example.com/ref-video.mp4
                  audio_urls:
                    - https://example.com/ref-audio.mp3
                  seconds: '5'
                  size: 1280x720
              seedance_volc_content:
                summary: Seedance Volcano Ark content mode
                value:
                  size: 1280x720
                  model: doubao-seedance-2-0-fast-260128
                  prompt: this is an example
                  duration: 15
                  fps: '24'
                  seed: '20231024'
                  mode: reference_material
                  content:
                    - type: text
                      text: this is an example
                    - type: image_url
                      image_url:
                        url: https://example.png
                      role: reference_image
                      name: image1
                    - type: audio_url
                      audio_url:
                        url: https://example.mp3
                      role: reference_audio
                      name: audio1
              veo_first_last_frame:
                summary: Veo First and last frame mode
                value:
                  model: veo3.1-fast
                  prompt: >-
                    A smooth transition from first frame to last frame,
                    cinematic motion
                  image_urls:
                    - https://example.com/first-frame.jpg
                  end_image_url: https://example.com/last-frame.jpg
                  seconds: '8'
                  size: 1280x720
              veo_reference_images_1to3:
                summary: Veo 1–3 reference images mode
                value:
                  model: veo3.1-fast
                  prompt: >-
                    Use reference style and character consistency for a short ad
                    clip
                  reference_images:
                    - https://example.com/ref-1.jpg
                  seconds: '8'
                  size: 1280x720
              kling_text2video:
                summary: Kling Text-to-Video
                value:
                  model: kling-video
                  prompt: A cat playing piano in the garden
                  duration: 5
                  size: 1280x720
                  metadata:
                    negative_prompt: blurry, low quality
                    mode: std
              kling_image2video:
                summary: Kling Image-to-Video
                value:
                  model: kling-video
                  prompt: A cat playing piano in the garden
                  image: https://example.com/input.jpg
                  duration: 5
                  size: 1280x720
                  metadata:
                    negative_prompt: blurry, low quality
                    mode: std
              grok_reference_video:
                summary: Grok Video Mode
                value:
                  model: grok-imagine-1.0-video
                  prompt: A dramatic cinematic shot with fast motion and rich lighting
                  reference_images:
                    - https://example.com/ref-1.jpg
                  seconds: 10
                  aspect_ratio: '16:9'
                  resolution: 720P
              minimax_h3_sora_reference:
                summary: MiniMax H3 multi-source video
                value:
                  model: minimax-h3
                  prompt: A cinematic product video with smooth camera motion
                  content:
                    - type: text
                      text: >-
                        Use the product image as reference and match the rhythm
                        of the audio.
                    - type: image_url
                      image_url:
                        url: https://example.com/product.jpg
                      role: reference_image
                      name: product
                    - type: audio_url
                      audio_url:
                        url: https://example.com/music.mp3
                      role: reference_audio
                      name: music
                  duration: 8
                  resolution: 1440p
                  aspect_ratio: '16:9'
                  generate_audio: true
              hailuo_2_3_sora_first_frame:
                summary: Hailuo 2.3 First-frame image-to-video
                value:
                  model: hailuo-2.3
                  prompt: >-
                    The camera slowly pushes in, natural motion and cinematic
                    lighting.
                  content:
                    - type: text
                      text: >-
                        The camera slowly pushes in, natural motion and
                        cinematic lighting.
                    - type: image_url
                      image_url:
                        url: https://example.com/first-frame.jpg
                      role: first_frame
                      name: first
                  duration: 10
                  resolution: 768p
                  aspect_ratio: '16:9'
              vidu_sora_t2v:
                summary: Vidu Sora t2v
                value:
                  model: viduq3-turbo
                  mode: t2v
                  content:
                    - type: text
                      text: A stylish product ad clip with cinematic lighting
                  duration: 5
                  resolution: 720p
                  aspect_ratio: '16:9'
              vidu_sora_i2v:
                summary: Vidu Sora i2v
                value:
                  model: viduq3-turbo
                  mode: i2v
                  content:
                    - type: text
                      text: A stylish ad clip from a product still image
                    - type: image_url
                      image_url: https://example.com/product.jpg
                  duration: 5
                  resolution: 720p
              vidu_sora_first_last:
                summary: Vidu Sora i2v_first_last
                value:
                  model: viduq3-turbo
                  mode: i2v_first_last
                  content:
                    - type: text
                      text: A smooth shot from start frame to end frame
                    - type: image_url
                      image_url: https://example.com/start.jpg
                    - type: image_url
                      image_url: https://example.com/end.jpg
                  duration: 5
                  resolution: 720p
              vidu_sora_reference:
                summary: Vidu Sora reference_images
                value:
                  model: viduq3-turbo
                  mode: reference_images
                  content:
                    - type: text
                      text: Maintain product style and character consistency
                    - type: image_url
                      image_url: https://example.com/ref-1.jpg
                    - type: image_url
                      image_url: https://example.com/ref-2.jpg
                    - type: image_url
                      image_url: https://example.com/ref-3.jpg
                  duration: 5
                  resolution: 720p
                  aspect_ratio: '16:9'
              hailuo_text2video:
                summary: Hailuo 2.3 First-frame image-to-video
                value:
                  model: hailuo-2.3
                  prompt: A cinematic travel shot over the coastline
                  image_urls:
                    - https://example.com/first-frame.jpg
                  duration: 6
                  resolution: 768p
                  aspect_ratio: '16:9'
      responses:
        '200':
          description: Video task created successfully
          content:
            application/json:
              schema:
                type: object
                properties:
                  id:
                    type: string
                  task_id:
                    type: string
                    description: Legacy compatibility; to be deprecated
                  object:
                    type: string
                  model:
                    type: string
                  status:
                    type: string
                    description: >-
                      Should use VideoStatus constants: VideoStatusQueued,
                      VideoStatusInProgress, VideoStatusCompleted,
                      VideoStatusFailed
                    enum:
                      - queued
                      - in_progress
                      - completed
                      - failed
                      - video_url
                      - url
                      - completed_at
                  progress:
                    type: integer
                  created_at:
                    type: integer
                  completed_at:
                    type: integer
                  expires_at:
                    type: integer
                  seconds:
                    type: string
                  size:
                    type: string
                  remixed_from_video_id:
                    type: string
                  error:
                    type: object
                  metadata:
                    type: object
                    additionalProperties: true
                  output:
                    type: object
                  request_id:
                    type: string
                required:
                  - created_at
                  - id
                  - model
                  - object
                  - progress
                  - status
              example:
                id: video_abc123
                object: video
                model: sora-2
                status: queued
                progress: 0
                created_at: 1764347090922
                seconds: '8'
          headers: {}
        '400':
          description: Invalid request parameters
          content:
            application/json:
              schema:
                $ref: '#/components/schemas/ErrorResponse'
          headers: {}
      deprecated: false
      security:
        - BearerAuth: []
components:
  schemas:
    ErrorResponse:
      type: object
      properties:
        error:
          type: object
          properties:
            message:
              type: string
              description: Error message
            type:
              type: string
              description: Error Type
            param:
              type: string
              description: Related Parameters
              nullable: true
            code:
              type: string
              description: Error Codes
              nullable: true
  securitySchemes:
    BearerAuth:
      type: http
      scheme: bearer
      description: |-
        Use Bearer Token authentication.
        Format: `Authorization: Bearer sk-xxxxxx`

````