# Alibaba | Wan | 3.0 | Image to Video Alibaba Wan 3.0 Image-to-Video turns first-frame or first-and-last-frame images into AI video with motion, audio, duration, and resolution controls. ## API Information - **Model Slug:** alibaba-wan-3-0-image-to-video - **Branded URL:** https://www.eachlabs.ai/alibaba/wan-3-0/alibaba-wan-3-0-image-to-video - **Provider:** Alibaba - **Category:** Image to Video - **Output Type:** video - **Status:** active - **Base Cost:** $0.05–$0.20 per second, depending on resolution - **Estimated Processing Time:** 60 seconds - **Interactive Demo:** https://www.eachlabs.ai/ai-models/alibaba-wan-3-0-image-to-video ## Pricing - **Charge Type:** dynamic - **Estimate:** $0.05–$0.20 per second, depending on resolution - **Pricing Details:** 1080p: $0.2 - **Pricing Details:** 720p: $0.1 - **Pricing Details:** 480p: $0.05 - **Pricing Details:** default: $0.2 ## Input Schema | Parameter | Type | Required | Default | Constraints | Description | |-----------|------|----------|---------|-------------|-------------| | prompt | string | No | - | - | Text description to guide video generation. Optional but recommended for better results. Maximum 20,000 characters. | | first_frame | string | Yes | - | - | URL of the image to use as the first frame. Supported formats: JPEG, JPG, PNG without alpha, BMP, WEBP. Resolution: 240-8000px per side. Aspect ratio: 1:8 to 8:1. Maximum file size: 20 MB. | | last_frame | string | No | - | - | Optional URL of the image to use as the last frame. When provided, first-and-last-frame video generation is used. | | ratio | string | No | adaptive | adaptive, 16:9, 4:3, 1:1, 3:4, 9:16 | Output video aspect ratio. Adaptive lets the model choose a suitable ratio based on the input image. | | resolution | string | No | 1080P | 480P, 720P, 1080P | Output video resolution. | | duration | string | No | 5 | auto, 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20, 21, 22, 23, 24, 25, 26, 27, 28, 29, 30 | Generated video duration. Select Auto to let the model determine the duration automatically. | | audio | boolean | No | true | - | Whether to generate an audio track for the output video. | | prompt_extend | boolean | No | true | - | Enable automatic prompt rewriting. Enabled by default. Disabling it may reduce latency but can reduce generation quality. | | seed | integer | No | - | 0–2147483647 | Random seed for reproducibility. The same seed may produce similar, but not identical, results. | ## Example Request ```bash curl -X POST https://api.eachlabs.ai/v1/prediction/ \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "alibaba-wan-3-0-image-to-video", "input": { "first_frame": "https://cdn-us.eachlabs.ai/defaults/5b3f13fefafc42238966352a7f9cf013.png" } }' ``` ## Output Schema Response returned by `GET /v1/prediction/{id}` when the job completes: ```json { "status": "success", "predictionID": "string", "output": "string (URL of generated video)", "metrics": { "predict_time": "number (seconds)" } } ``` ## Polling ```bash curl https://api.eachlabs.ai/v1/prediction/{PREDICTION_ID} \ -H "Authorization: Bearer YOUR_API_KEY" ``` | Status | Meaning | |--------|---------| | `processing` | Still running — poll again | | `success` | Done — read `output` | | `error` | Failed — read `message` / `details` | ## Webhook (alternative to polling) Pass `"webhook_url": "https://your.host/path"` in the create request. Eachlabs POSTs this payload when the job ends: ```json { "exec_id": "prediction-uuid", "status": "succeeded", "output": "https://...", "error": "" } ``` `status` is `"succeeded"` or `"failed"`. `exec_id` equals the `predictionID` from create. Return 2xx within 30 seconds. ## Errors Error body: `{ "status": "error", "message": "...", "details": "..." }` | Code | Meaning | |------|---------| | `400` | Invalid input | | `401` | Missing / invalid `Authorization` bearer token | | `404` | Unknown model or prediction id | | `429` | Rate limit — 100 creates / min, 10 concurrent per key | | `5xx` | Retry with backoff | ## Overview **Alibaba | Wan | 3.0 | Image to Video Overview** Alibaba | Wan | 3.0 | Image to Video is an advanced image-to-video mode within Alibaba’s Wan 3.0 multimodal video generation family, designed to turn still images into continuous AI video with native audio in a single pass. It animates a first frame, or both a first and last frame, into directed clips of up to 30 seconds without stitching separate segments. The model’s primary differentiator is its combination of long single-pass duration, strong source fidelity for image-to-video, and synchronized audio, all controlled through prompts and reference inputs. Built by Alibaba’s Tongyi Lab and exposed via Wan 3.0 endpoints and partner platforms, it fits seamlessly into each::labs as a powerful option for creators who need cinematic motion, camera control, and sound from still imagery. ## Usage Notes - API Base URL: `https://api.eachlabs.ai/v1` - Authentication: send `Authorization: Bearer YOUR_API_KEY`. Generate a key from the Eachlabs dashboard at https://www.eachlabs.ai/dashboard/api-keys. - File-typed parameters (`*_url`, `image_url`, `video_url`, `audio_url`, etc.) accept publicly-reachable HTTPS URLs only. Upload your asset first (GCS / S3 / your CDN) and pass the resulting URL. Data-URIs and localhost URLs are rejected. - For structured parameters (arrays / objects) send real JSON values, not stringified payloads. - Monetary values are reported in USD; per-token / per-megapixel rates may be billed in micro-cents internally. - Prefer `webhook_url` over polling for long-running predictions — see the Webhook Callback section.