# wan-3.0/text-to-video

> Wan 3.0 (Text-to-Video) turns a prompt alone into native video up to 30 seconds with synchronized dialogue, music, and sound effects — no reference media required. Supports multi-scene narratives at up to 1080p with strong instruction-following.

- **Provider:** Alibaba
- **Category:** text-to-video
- **Price:** $0.2000 per run

## Key Features

- Generates video from a prompt alone — no reference image required, up to 30 seconds.
- Native audio-video synthesis — synchronized dialogue, background music and sound effects generated automatically.
- Multi-scene narrative support — a single generation can carry more than one beat or camera setup.
- Strong instruction-following — accurately renders detailed, multi-element prompts.
- Multiple resolution tiers — generate at 480p, 720p, or 1080p.
- Bilingual prompt support — supports both English and Chinese prompts natively.

## Parameters

| Parameter | Required | Description |
| --- | --- | --- |
| `prompt` | Yes | Detailed description of the scene, action, and dialogue (English or Chinese). |
| `duration` | No | Video length in seconds: 2-30 (default: 5). |
| `resolution` | No | Output resolution: 480p, 720p (default), or 1080p. |
| `aspect_ratio` | No | Output format: 16:9 (default), 9:16, 4:3, 3:4, 1:1, or adaptive. |

## How to Use

1. Write a prompt — describe subject, action, camera movement, lighting and mood in English or Chinese.
2. Select your aspect ratio (e.g., 16:9 for widescreen).
3. Choose a duration up to 30 seconds.
4. Generate and download your video with synchronized dialogue, music, and sound effects.

## Code Examples

### Python

```python
import os
import requests

response = requests.post(
    "https://aircube.ai/api/v3/wan-3.0/text-to-video",
    headers={
        "Authorization": "Bearer " + os.environ["AIRCUBE_API_KEY"],
        "Content-Type": "application/json",
    },
    json={
    "prompt": "A chef plates a dessert in a busy restaurant kitchen, steam rising, camera pushes in slowly, ambient kitchen noise",
    "duration": 8,
    "aspect_ratio": "16:9",
    "resolution": "1080p"
},
    timeout=300,
)
data = response.json()

if data["success"]:
    print("ID:", data["data"]["id"], "Status:", data["data"]["status"])
else:
    print("Error:", data["error"]["message"])
```

### Node.js

```javascript
const response = await fetch("https://aircube.ai/api/v3/wan-3.0/text-to-video", {
  method: "POST",
  headers: {
    "Authorization": "Bearer " + process.env.AIRCUBE_API_KEY,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
  "prompt": "A chef plates a dessert in a busy restaurant kitchen, steam rising, camera pushes in slowly, ambient kitchen noise",
  "duration": 8,
  "aspect_ratio": "16:9",
  "resolution": "1080p"
}),
});

const data = await response.json();

if (data.success) {
  console.log("ID:", data.data.id, "Status:", data.data.status);
} else {
  console.error("Error:", data.error.message);
}
```

### cURL

```curl
curl -X POST "https://aircube.ai/api/v3/wan-3.0/text-to-video" \
  -H "Authorization: Bearer $AIRCUBE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "prompt": "A chef plates a dessert in a busy restaurant kitchen, steam rising, camera pushes in slowly, ambient kitchen noise",
  "duration": 8,
  "aspect_ratio": "16:9",
  "resolution": "1080p"
}'
```

### Python (Async)

```python
import os
import time
import requests

# 1. Submit
response = requests.post(
    "https://aircube.ai/api/v3/wan-3.0/text-to-video",
    headers={
        "Authorization": "Bearer " + os.environ["AIRCUBE_API_KEY"],
        "Content-Type": "application/json",
    },
    json={
    "prompt": "A chef plates a dessert in a busy restaurant kitchen, steam rising, camera pushes in slowly, ambient kitchen noise",
    "duration": 8,
    "aspect_ratio": "16:9",
    "resolution": "1080p"
},
    timeout=300,
)
data = response.json()

if not data["success"]:
    print("Error:", data["error"]["message"])
    exit(1)

generation_id = data["data"]["id"]
print(f"Submitted: {generation_id}")

# 2. Poll until completed
while True:
    time.sleep(5)
    r = requests.get(
        f"https://aircube.ai/api/v3/status/{generation_id}",
        headers={"Authorization": "Bearer " + os.environ["AIRCUBE_API_KEY"]},
    )
    result = r.json()["data"]

    if result["status"] == "completed":
        print(result["output_url"])
        break
    elif result["status"] == "failed":
        print("Generation failed")
        break
```

### Node.js (Async)

```javascript
// 1. Submit
const response = await fetch("https://aircube.ai/api/v3/wan-3.0/text-to-video", {
  method: "POST",
  headers: {
    "Authorization": "Bearer " + process.env.AIRCUBE_API_KEY,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
  "prompt": "A chef plates a dessert in a busy restaurant kitchen, steam rising, camera pushes in slowly, ambient kitchen noise",
  "duration": 8,
  "aspect_ratio": "16:9",
  "resolution": "1080p"
}),
});
const data = await response.json();

if (!data.success) {
  console.error("Error:", data.error.message);
  process.exit(1);
}

const generationId = data.data.id;
console.log("Submitted:", generationId);

// 2. Poll until completed
while (true) {
  await new Promise((r) => setTimeout(r, 5000));
  const res = await fetch(
    `https://aircube.ai/api/v3/status/${generationId}`,
    { headers: { "Authorization": "Bearer " + process.env.AIRCUBE_API_KEY } },
  );
  const result = (await res.json()).data;

  if (result.status === "completed") {
    console.log(result.output_url);
    break;
  } else if (result.status === "failed") {
    console.error("Generation failed");
    break;
  }
}
```

### cURL (Async)

```curl
# 1. Submit
RESPONSE=$(curl -s -X POST "https://aircube.ai/api/v3/wan-3.0/text-to-video" \
  -H "Authorization: Bearer $AIRCUBE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "prompt": "A chef plates a dessert in a busy restaurant kitchen, steam rising, camera pushes in slowly, ambient kitchen noise",
  "duration": 8,
  "aspect_ratio": "16:9",
  "resolution": "1080p"
}')

ID=$(echo "$RESPONSE" | jq -r '.data.id')
echo "Submitted: $ID"

# 2. Poll until completed
while true; do
  sleep 5
  STATUS_RES=$(curl -s "https://aircube.ai/api/v3/status/$ID" \
    -H "Authorization: Bearer $AIRCUBE_API_KEY")
  STATUS=$(echo "$STATUS_RES" | jq -r '.data.status')

  if [ "$STATUS" = "completed" ]; then
    echo "$STATUS_RES" | jq -r '.data.output_url'
    break
  elif [ "$STATUS" = "failed" ]; then
    echo "Generation failed"; break
  fi
done
```

## Pricing

| Resolution | Duration | Cost |
| --- | --- | --- |
| 480p | 4s | $0.20 |
| 480p | 5s | $0.25 |
| 480p | 6s | $0.30 |
| 480p | 8s | $0.40 |
| 480p | 10s | $0.50 |
| 480p | 12s | $0.60 |
| 480p | 15s | $0.75 |
| 480p | 20s | $1.00 |
| 480p | 25s | $1.25 |
| 480p | 30s | $1.50 |
| 720p | 4s | $0.40 |
| 720p | 5s | $0.50 |
| 720p | 6s | $0.60 |
| 720p | 8s | $0.80 |
| 720p | 10s | $1.00 |
| 720p | 12s | $1.20 |
| 720p | 15s | $1.50 |
| 720p | 20s | $2.00 |
| 720p | 25s | $2.50 |
| 720p | 30s | $3.00 |
| 1080p | 4s | $0.80 |
| 1080p | 5s | $1.00 |
| 1080p | 6s | $1.20 |
| 1080p | 8s | $1.60 |
| 1080p | 10s | $2.00 |
| 1080p | 12s | $2.40 |
| 1080p | 15s | $3.00 |
| 1080p | 20s | $4.00 |
| 1080p | 25s | $5.00 |
| 1080p | 30s | $6.00 |

### Billing Rules

- 480p / 4s: $0.20.
- 720p / 4s: $0.40.
- 1080p / 4s: $0.80.
- Longer durations scale proportionally.
- Failed generations are not charged.

## Best Use Cases

- Short-form video content — create engaging clips for social media platforms with native audio.
- Storyboarding — visualize written concepts as animated multi-scene sequences.
- Promotional content — generate video ads with spoken dialogue directly from a brief.
- Multi-scene narrative — script scenes with more than one camera setup in a single 30-second generation.
- Chinese-language projects — native Chinese prompt support.

## Pro Tips

- Include specific camera directions: tracking shot, dolly zoom, static wide angle.
- Describe lighting explicitly: golden hour, dramatic shadows, soft diffused light.
- Write dialogue in quotes for spoken lines — Wan 3.0 generates synchronized audio automatically.
- Use the 30-second ceiling for scripts with multiple scenes or beats.
- Wan 3.0 supports both English and Chinese prompts natively.

## Notes

- Bilingual prompt support: English and Chinese.
- Output in MP4 format with native synchronized audio.
- Duration range: 2-30 seconds.
- Successor to Wan 2.7 with a longer duration ceiling and native multi-scene narratives.

## FAQ

**Q: What is the Wan 3.0 Text to Video API?**

A text-to-video generation API from Alibaba that creates video up to 30 seconds long with native synchronized dialogue, music, and sound effects, available through AirCube.

**Q: Do I need an image?**

No — Wan 3.0 Text to Video generates video purely from text descriptions.

**Q: How long can videos be?**

2 to 30 seconds per generation, double the ceiling of Wan 2.7.

**Q: Does it generate audio?**

Yes — dialogue, background music, and sound effects are generated natively and synchronized with the picture.

**Q: How much does it cost?**

Starting at $0.25 for a 480p 5-second clip, scaling with resolution and duration.

**Q: Can I use outputs commercially?**

Yes — videos are yours to use commercially.
