What is the difference between wan2.6-r2v and regular image-to-video?
It uses the character identity in a reference video as the basis for creation, applying the person's appearance to a new scene; image-to-video uses a static image as the generation starting point. If you already have a video character and want to continue telling their story, r2v is a better fit; if you mainly want to animate a single image, consider wan2.6-i2v.
How long can the generated videos be, and at what resolution?
The native output specifications of wan2.6-r2v are 2–10 seconds, with support for 720P and 1080P, and videos are 30 fps MP4 files. When designing scripts, arrange actions around short shots, and do not treat the 15-second limit of wan2.6-t2v and wan2.6-i2v as the output range for this model.
How do I submit a reference video and obtain the generated result?
Submit model=wan2.6-r2v to POST /wan/videos, provide the reference video URL with reference_video_urls, and use prompt to describe the new scene. Set async=true to first obtain a task_id, then query the final status through /wan/tasks; after success, obtain the video URL for playback or download.
How should prompts be written for character transfer?
First explain the identity of the reference character in the new clip, then describe the environment, action sequence, and shot intent. For multi-character tasks, distinguish who does what, and avoid using only abstract terms such as “cinematic feel.” Keep the plot focused on one clear event, and check whether the generated character appearance and actions meet the creative goal.
Does it support audio, and is it equivalent to voice-over editing?
wan2.6-r2v has native audio synchronization capabilities and can be used for short narratives with sound, but this does not make it an independent voice-over editing, precise lip-sync correction, or voice cloning tool. When creating, describe audio requirements together with scene events, and check the synchronization effect after completion; when precise dialogue or audio processing is needed, arrange post-production afterward.