> ## Documentation Index
> Fetch the complete documentation index at: https://docs.budgetpixel.com/llms.txt
> Use this file to discover all available pages before exploring further.

# Generate video with Wan 3.0

> Alibaba Wan 3.0 — all-in-one video generation up to 30 seconds at 30fps. Text-to-video; image-to-video (first + optional last frame); and reference-to-video mixing up to 10 reference images, up to 5 reference video clips (15s combined) for editing/replication/extension, and up to 5 reference audio clips (15s combined). In reference mode the prompt can address assets by order ("Image 1", "Video 1"). Reference media can't be combined with start/end frames. 480p/720p/1080p, priced per OUTPUT second by resolution (input media is free — see `resolution_pricing` in GET /v1/models); any duration 2-30 seconds (input video + output ≤ 30s); optional audio.

**Asynchronous.** Returns a job `id` (the video is not in this response). Poll [`GET /v1/videos/{id}`](/api-reference/videos/get-video-job-status) until `status` is `succeeded` — the video URL is in that response's `video_url`.



## OpenAPI

````yaml /openapi.yaml post /videos/wan-3.0-video
openapi: 3.1.0
info:
  contact:
    email: support@budgetpixel.com
    name: BudgetPixel Support
  description: >
    The BudgetPixel developer API for programmatic image, video, music and

    sound-effect generation.


    Generation is **asynchronous**: you create a job, then poll its status until
    it

    reaches a terminal state (`succeeded` / `failed`). You are charged in
    credits

    only on success — never for failures, timeouts, or content blocked before

    generation.


    **Authentication.** All requests require an API key sent as a bearer token:

        Authorization: Bearer bpx_live_xxx

    Create and manage keys from your BudgetPixel account dashboard. The
    developer

    API is available on the **Premium** plan and above.


    **Pricing.** API usage is metered in credits at each model's published price

    (see `GET /v1/models`), and your plan's model discounts and active
    promotions

    apply to API charges exactly as they do on the web — `POST /v1/cost` quotes

    your discounted price. Pro and Ultra plans include a daily free-models

    allowance (a shared pool of free generations per day across qualified image

    models such as FLUX 2 Klein and Qwen-Image) that covers API

    requests too — the standard price applies once the day's pool is used.

    Other promotional free daily generations remain web/app-only. Failed

    generations are never charged.


    **Rate limits.** Requests are limited per minute: **600 requests/min per API

    key** and **1200 requests/min per source IP** (defaults; subject to tuning).

    Every response carries `X-RateLimit-Limit`, `X-RateLimit-Remaining`, and

    `X-RateLimit-Reset` (seconds until the window resets). Exceeding a limit

    returns HTTP `429` with a `Retry-After` header — wait that long before

    retrying. Rejected requests still count against the window, so a tight retry

    loop keeps you limited; back off instead. Generation jobs additionally have
    a

    per-plan concurrency cap (the same one as the web workshop), independent of

    these request limits. `POST /v1/uploads` also has a dedicated **60
    uploads/min

    per account** cap (it's free/unmetered) on top of these limits.
  title: BudgetPixel API
  version: '2026-06-20'
servers:
  - description: Production
    url: https://api.budgetpixel.com/v1
security:
  - ApiKeyAuth: []
tags:
  - description: Account and credit balance
    name: Account
  - description: Discover available models and pricing
    name: Models
  - description: Upload input media for image/video generation
    name: Uploads
  - description: Image job status
    name: Images
  - description: FLUX image models
    name: Black Forest Labs
  - description: SeeDream image models
    name: Bytedance
  - description: Video job status
    name: Videos
  - description: Kuaishou Kling video models
    name: Kling
  - description: SeeDance video models
    name: ByteDance (SeeDance)
  - description: Music and sound-effect job status
    name: Audios
  - description: Music generation models
    name: Music
  - description: Text-to-sound-effect models (priced per second)
    name: Sound Effects
  - description: Format conversion for images, video, and audio
    name: Conversions
  - description: Publish posts to your BudgetPixel feed
    name: Social
  - description: Content moderation — NSFW rating and CSAM detection
    name: Moderation
  - description: Content classification — music genre detection
    name: Classification
paths:
  /videos/wan-3.0-video:
    post:
      tags:
        - Alibaba
      summary: Generate video with Wan 3.0
      description: >-
        Alibaba Wan 3.0 — all-in-one video generation up to 30 seconds at 30fps.
        Text-to-video; image-to-video (first + optional last frame); and
        reference-to-video mixing up to 10 reference images, up to 5 reference
        video clips (15s combined) for editing/replication/extension, and up to
        5 reference audio clips (15s combined). In reference mode the prompt can
        address assets by order ("Image 1", "Video 1"). Reference media can't be
        combined with start/end frames. 480p/720p/1080p, priced per OUTPUT
        second by resolution (input media is free — see `resolution_pricing` in
        GET /v1/models); any duration 2-30 seconds (input video + output ≤ 30s);
        optional audio.


        **Asynchronous.** Returns a job `id` (the video is not in this
        response). Poll [`GET
        /v1/videos/{id}`](/api-reference/videos/get-video-job-status) until
        `status` is `succeeded` — the video URL is in that response's
        `video_url`.
      operationId: createVideo_wan_3_0_video
      requestBody:
        content:
          application/json:
            schema:
              properties:
                aspect_ratio:
                  default: '16:9'
                  description: >-
                    Aspect ratio for text-to-video. With any image/video/audio
                    input the frame adapts to the inputs and this is ignored.
                  enum:
                    - '16:9'
                    - '9:16'
                    - '1:1'
                    - '4:3'
                    - '3:4'
                  type: string
                end_image:
                  description: >-
                    Optional last frame, used together with `image` (the first
                    frame) to interpolate the video between the two frames. Same
                    input forms as `image`.
                  type: string
                generate_audio:
                  default: true
                  description: >-
                    Generate audio with the video (default true; no price
                    impact).
                  type: boolean
                image:
                  description: >-
                    Optional first frame for image-to-video. Provide a public
                    image URL, a data URI, raw base64, or an uploaded-file URL
                    from POST /v1/uploads. Omit for text-to-video. Can't be
                    combined with reference media.
                  type: string
                length_seconds:
                  default: 5
                  description: >-
                    Output video length in seconds (2-30). With reference
                    videos, input duration + output length must not exceed 30
                    seconds.
                  maximum: 30
                  minimum: 2
                  type: integer
                prompt:
                  description: >-
                    Text description of the video, or the edit instruction when
                    reference videos are supplied. In reference mode, address
                    assets by order: "Image 1", "Video 1", "Audio 1".
                  type: string
                reference_audios:
                  description: >-
                    Optional reference audio clips (up to 5; each 2-15s, 15s
                    combined; WAV/MP3) that guide sound/voice. Free. Each item
                    is a public audio URL or an uploaded-file URL from POST
                    /v1/uploads. Can't be combined with `image`/`end_image`.
                  items:
                    type: string
                  maxItems: 5
                  type: array
                reference_images:
                  description: >-
                    Optional reference images (up to 10, free) that guide
                    identity/style/scene in reference-to-video mode. Each item
                    is a public image URL, a data URI, raw base64, or an
                    uploaded-file URL from POST /v1/uploads. Can't be combined
                    with `image`/`end_image`.
                  items:
                    type: string
                  maxItems: 10
                  type: array
                reference_videos:
                  description: >-
                    Optional reference video clips (up to 5; each 2-15s, 15s
                    combined; MP4/MOV, ≤50MB each) for editing, effect/camera
                    replication, and extension. Each item is a public video URL
                    or an uploaded-file URL from POST /v1/uploads (videos are
                    passed by URL, not inlined). Free — only OUTPUT seconds are
                    billed — but input video duration + `length_seconds` must
                    not exceed 30. Can't be combined with `image`/`end_image`.
                  items:
                    type: string
                  maxItems: 5
                  type: array
                resolution:
                  default: 720p
                  description: >-
                    Output resolution. Pricing varies by resolution — see
                    `resolution_pricing` in GET /v1/models.
                  enum:
                    - 480p
                    - 720p
                    - 1080p
                  type: string
                seed:
                  description: Seed for reproducible generation. Omit for random.
                  type: integer
              required:
                - prompt
              type: object
        required: true
      responses:
        '200':
          content:
            application/json:
              schema:
                $ref: '#/components/schemas/CreateVideoResponse'
          description: Job accepted.
        '400':
          $ref: '#/components/responses/BadRequest'
        '401':
          $ref: '#/components/responses/Unauthorized'
        '403':
          $ref: '#/components/responses/Forbidden'
        '429':
          $ref: '#/components/responses/TooManyRequests'
components:
  schemas:
    CreateVideoResponse:
      properties:
        id:
          description: Opaque job id — use it to poll status.
          type: string
        message:
          type: string
        model:
          type: string
        status:
          $ref: '#/components/schemas/JobStatus'
      type: object
    JobStatus:
      description: Lifecycle state. `succeeded`/`failed`/`timeout` are terminal.
      enum:
        - pending
        - starting
        - processing
        - completing
        - succeeded
        - failed
        - timeout
      type: string
    Error:
      properties:
        error:
          properties:
            code:
              description: Stable machine-readable code.
              example: model_not_available
              type: string
            message:
              type: string
            type:
              description: Error category.
              example: invalid_request_error
              type: string
          required:
            - type
            - code
            - message
          type: object
      required:
        - error
      type: object
    ModerationBlocked:
      description: >-
        Returned (with HTTP 400) when the input content moderation gate blocks a
        generation request. The block is a property of the request's prompt or
        input media — reword the prompt or change the input and retry. Branch on
        `restriction_reason`, which is stable and machine-readable.
      properties:
        error:
          description: Human-readable explanation of the block.
          type: string
        restriction_reason:
          description: Stable machine-readable block reason.
          enum:
            - input_csam
            - input_explicit_adult
            - input_upload_nudity
            - input_celebrity_likeness
            - strict_model_nsfw
          type: string
      required:
        - error
        - restriction_reason
      type: object
  responses:
    BadRequest:
      content:
        application/json:
          schema:
            oneOf:
              - $ref: '#/components/schemas/Error'
              - $ref: '#/components/schemas/ModerationBlocked'
      description: >-
        The request was malformed, referenced an unavailable model, or was
        blocked by input content moderation (moderation blocks carry a
        `restriction_reason` — see ModerationBlocked).
    Unauthorized:
      content:
        application/json:
          schema:
            $ref: '#/components/schemas/Error'
      description: Missing or invalid API key.
    Forbidden:
      content:
        application/json:
          schema:
            $ref: '#/components/schemas/Error'
      description: >-
        Authenticated but not permitted (plan gate, ownership, account
        restriction). Content-moderation blocks are 400, not 403.
    TooManyRequests:
      content:
        application/json:
          schema:
            $ref: '#/components/schemas/Error'
      description: Rate or queue limit reached. Retry after a short delay.
  securitySchemes:
    ApiKeyAuth:
      bearerFormat: bpx_live_*
      description: 'API key as a bearer token: Authorization: Bearer bpx_live_xxx'
      scheme: bearer
      type: http

````