Installable

Scenema Audio

Zero-shot expressive voice cloning and speech generation. Clone any voice from 10-20s of audio and direct its emotional performance with stage directions. NVIDIA GPU with 16GB+ VRAM required.

github.com/vdruts/scenema-audio.pinokioUpdated 2mo ago
Check-in
Loading community details…

Community

Create post
Ask or share about Scenema Audio…
Loading...