New — access 450+ AI models through one unified API.Start building free

seedance-2.0-mini/text-to-video

Seedance 2.0 Mini Text to Video is ByteDance's faster, lower-cost text-to-video model for cinematic multi-shot videos. It generates narrative sequences from text prompts with AI camera control, consistent characters across scenes, 480P / 720P output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

text-to-video$0.1400/ run

Prompt

Enhancer
0/2000
PreviewJSON

Your result will appear here

Write a prompt and hit Generate.

Examples

A realistic cinematic close-up shot of a beautiful young blonde European woman blowing a bubble gum bubble. She has long blonde hair, soft natural makeup, fair skin, and a stylish casual outfit. Warm natural daylight, shallow depth of field, film grain texture.

A cinematic ocean wave at sunrise, highly detailed.

Related models

seedance-2.0 / image-to-video15% OFFimage-to-video
seedance-2.0 / image-to-video

Seedance 2.0 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.

Seedance$0.4080$0.4800
seedance-2.0 / text-to-video15% OFFtext-to-video
seedance-2.0 / text-to-video

Seedance 2.0 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics.

Seedance$0.4080$0.4800
seedance-2.0 / reference-to-video15% OFFreference-to-video
seedance-2.0 / reference-to-video

Seedance 2.0 (Reference-to-Video) generates cinematic videos guided by up to 12 reference files spanning images, videos, and audio clips. Use @Image1, @Video1, @Audio1 tags in your prompt to assign roles — style transfer, lip-sync, motion transfer, character consistency, and multi-scene composition. Supports 480P / 720P / 1080P / 4K output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Seedance$0.4080$0.4800
seedance-2.0 / image-to-video-spicy15% OFFimage-to-video
seedance-2.0 / image-to-video-spicy

Seedance 2.0 Spicy Image to Video is a fast AI image-to-video generation model that creates high-quality cinematic clips from images, optimized for scalable content generation with smooth animations and stable aesthetics. Ready-to-use REST inference API for animating images, social media clips, product videos, advertising creatives, visual storytelling, and professional image-to-video workflows with simple integration, no coldstarts, and affordable pricing.

Seedance$0.5100$0.6000
seedance-2.0 / text-to-video-spicy15% OFFtext-to-video
seedance-2.0 / text-to-video-spicy

Seedance 2.0 Spicy Text to Video generates high-quality cinematic clips from text prompts, optimized for scalable content generation with smooth animations and stable aesthetics. Ready-to-use REST inference API for creating social media clips, product videos, advertising creatives, visual storytelling, and professional text-to-video workflows with simple integration, no coldstarts, and affordable pricing.

Seedance$0.5100$0.6000
seedance-2.0 / reference-to-video-spicy15% OFFreference-to-video
seedance-2.0 / reference-to-video-spicy

Seedance 2.0 Spicy Reference to Video generates cinematic videos guided by up to 12 reference files spanning images, videos, and audio clips. Use @Image1, @Video1, @Audio1 tags in your prompt to assign roles — style transfer, lip-sync, motion transfer, character consistency, and multi-scene composition. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Seedance$0.5100$0.6000

Seedance 2.0 Mini Text To Video

Seedance 2.0 Mini Text to Video is ByteDance's faster, lower-cost text-to-video model for cinematic multi-shot videos. It generates narrative sequences from text prompts with AI camera control, consistent characters across scenes, 480P / 720P output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Key Features

  • Generate video directly from text prompts — no source image required.
  • Cinematic quality — produce smooth, coherent video with consistent subjects and scenes.
  • Flexible creative control — specify camera angles, lighting, subject actions and mood.
  • Multiple duration and resolution options for preview through production workflows.
  • Coherent storytelling — maintains narrative consistency across the generated frames.
  • Diverse styles — photorealistic, animated, stylized, or abstract visual outputs.

Parameters

ParameterRequiredDescription
promptYesDetailed text description of the video scene to generate.
durationNoVideo length in seconds (default varies by model).
aspect_ratioNoOutput ratio: 16:9, 9:16, 4:3, 1:1 (default: 16:9).
resolutionNoOutput resolution: 480p, 720p (default: 720p).

How to Use

  1. Write a detailed prompt describing your video scene — include subject, action, setting, lighting and camera movement.
  2. Select aspect ratio and resolution appropriate for your use case.
  3. Set the desired duration for your video clip.
  4. Generate and download your video — iterate on the prompt for better results.

Code Examples

import os
import requests

response = requests.post(
    "https://aircube.ai/api/v3/seedance-2.0-mini/text-to-video",
    headers={
        "Authorization": "Bearer " + os.environ["AIRCUBE_API_KEY"],
        "Content-Type": "application/json",
    },
    json={
    "prompt": "Aerial shot of a coastal city at golden hour, camera slowly descending toward the waterfront",
    "duration": 8,
    "aspect_ratio": "16:9"
},
    timeout=300,
)
data = response.json()

if data["success"]:
    print("ID:", data["data"]["id"], "Status:", data["data"]["status"])
else:
    print("Error:", data["error"]["message"])

Pricing

ResolutionDurationCost
480p4s$0.14
480p5s$0.18
480p6s$0.22
480p8s$0.29
480p10s$0.35
480p12s$0.43
480p15s$0.53
720p4s$0.38
720p5s$0.47
720p6s$0.56
720p8s$0.75
720p10s$0.94
720p12s$1.13
720p15s$1.41

Billing rules

  • 480p / 4s: $0.14.
  • 720p / 4s: $0.38.
  • Longer durations scale proportionally.
  • Failed generations are not charged.

Best Use Cases

  • Content creation — generate video clips for social media, ads and marketing campaigns.
  • Storyboard visualization — bring written concepts to life as video previews.
  • Music video production — create visual sequences from lyrical descriptions.
  • Educational content — generate explainer or demonstration clips from descriptions.
  • Game and film pre-visualization — prototype scenes before full production.

Pro Tips

  • Write prompts like a film director — include specific camera angles (close-up, wide shot, tracking shot).
  • Describe lighting conditions (golden hour, dramatic shadows, neon glow) for mood control.
  • Keep the scene description focused — one clear action or sequence per generation works best.
  • Include temporal cues (slowly, suddenly, gradually) to guide the pacing of motion.
  • Start with 16:9 for cinematic content and 9:16 for social media verticals.

Notes

  • Video output is typically MP4 format.
  • Generation time scales with duration and resolution.
  • Complex prompts with multiple subjects may require iteration to achieve desired results.

Seedance 2.0 Mini Text To Video — Frequently asked questions