视频对口型 Video Retalk

Tongyi VideoRetalk lip sync / lip-sync (mouth sync, dubbing) video model — takes a talking-person video plus a voice audio track and regenerates the video so the speaker's mouth/lips match the new audio. Use this for lip syncing a person video to new speech. Optionally provide a reference face image to pick the target person when the video contains multiple faces. 通义声动人像 VideoRetalk 口型同步(对口型、lip sync / lip-sync、配音对嘴)视频模型,输入一段人物讲话视频与一段人声音频,生成讲话口型与音频匹配的新视频;适用于让人物视频的口型对上新的语音。当视频中存在多张人脸时,可额外提供人脸参考图来指定要替换口型的目标人物。

Install

openclaw skills install @dlazyai/dlazy-videoretalk