Back to models

Wan 3.0

VideoImage to video
alibaba/wan-3.0/reference-to-videoView on fal.ai

Parameters

Dictation ready.

Text prompt directing how the reference media is used. Reference media can be addressed positionally, e.g. 'the subject in Image 1 walks past Video 1'.

Up to 10 reference image URLs.

Up to 5 reference video URLs totaling at most 15 seconds. Each clip must be at least 16 fps.

Up to 5 reference audio URLs totaling at most 15 seconds.

Output aspect ratio, or adaptive selection.

Output duration in seconds. Set to null for smart duration, which lets the model pick a length from the prompt and reference media.

Result

Run the model to see the result here.