~/WAN 3.0 Text-to-Video
Text to Video

WAN 3.0 Text-to-Video

Newest WAN model: up to 30s videos, audio, and 1080p.

Schema

Run the model to see output

Your request will cost $0.50 per run. ($0.10/s) For $10 you can run this model approximately 20 times.

One more thing:

README

Alibaba / WAN 3.0 Text-to-VideoText to Video (wan-3.0-t2v)

WAN 3.0 Text-to-Video is the newest generation of Alibaba's WAN video model family, extending maximum duration to 30 seconds with native audio output and up to 1080p resolution.

Highlights

  • Long-form generation Videos from 4 up to 30 seconds.
  • Native audio Optionally generate a synchronized audio track.
  • 480p, 720p, and 1080p output.
  • Smart prompt rewriting Built-in LLM prompt expansion for better quality.

Parameters

  • prompt*Text description of the video to generate
  • resolutionOutput video resolution
    • 480p
    • 720p
    • 1080p
  • durationVideo duration in seconds (4-30)
  • aspect_ratioOutput aspect ratio
    • 16:9
    • 9:16
    • 4:3
    • 3:4
    • 1:1
    • 21:9
  • enable_audioGenerate the video with an audio track
  • enable_prompt_expansionEnable smart prompt rewriting for better quality

Pricing

$0.50 per generation

ResolutionPrice
480p$0.07
720p$0.10
1080p$0.20

How to Use

  1. 1.Write a prompt describing the video.
  2. 2.Choose resolution (480p-1080p) and duration (4-30s).
  3. 3.Submit and poll the generation until it completes.

Pro Tips

  • Be specific about camera movements for cinematic results.
  • Longer durations benefit from prompts describing a sequence of actions.
  • Keep prompt expansion enabled unless you need exact prompt control.

More Models to Try

Content creators needing longer AI video clips.
Marketing teams producing high-resolution campaigns.
Storytellers and filmmakers.

Frequently Asked Questions

What is the WAN 3.0 Text-to-Video API?
Newest WAN model: up to 30s videos, audio, and 1080p.
How much does WAN 3.0 Text-to-Video cost via API?
WAN 3.0 Text-to-Video costs $0.5000 per generation through Renderful's API. No subscription required — pay only for what you use.
How do I use WAN 3.0 Text-to-Video via API?
Sign up for a free Renderful API key, then send a POST request to the /v1/predictions endpoint with model "wan-3.0-t2v". See the documentation at renderful.ai/docs for code examples in Python, JavaScript, and cURL.
What type of content does WAN 3.0 Text-to-Video generate?
WAN 3.0 Text-to-Video is a text to video model by Alibaba. Key features include: 4-30s videos, Native audio output, Up to 1080p resolution.
Is the WAN 3.0 Text-to-Video API fast?
WAN 3.0 Text-to-Video has medium generation speed. Results are delivered via polling or webhook callback for seamless integration.