This application allows you to generate a cloned voice from a given text and speaker audio, and create a lip-synced video by combining the generated audio with a video or image. You need to provide...
Upload a video or image and an audio file to create a lip-synced video. Choose a checkpoint and adjust padding and resizing options to get the best results.
Upload a reference audio and a transform audio to change the tone of the transform audio to match the reference audio. You'll receive the modified audio as a result.
[NVIDIA ONLY] Generate Video Progressively. FramePack is a next-frame (next-frame-section) prediction neural network structure that generates videos progressively. https://github.com/lllyasviel/FramePack