Fully offline document-to-speech converter using Coqui TTS models (XTTS v2, Bark, VITS, YourTTS) and StyleTTS 2. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio with voice cloning support. No API keys or cloud services required.
Fully offline document-to-speech converter using OpenVoice V2. Tone color conversion and voice cloning via MeloTTS. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.
Reconstruct 3D Gaussian Splatting worlds from video using non-rigid alignment. Supports fast and extensive modes with 2DGS/3DGS rendering. https://github.com/lukasHoel/video_to_world
BazedFrog/SongGeneration-Studiov3.7updated 1mo ago
AI Song Generation with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model. [NVIDIA ONLY]
🌍 Capture speech, translate it instantly, and playback audio in a selected language with this easy-to-use real-time voice translator built on Python and Streamlit.
Offline Vietnamese text-to-speech (TTS) and voice cloning - fully local on your own NVIDIA GPU (works on 4GB VRAM), no internet, no API keys. Clean web UI + CLI. Built on OmniVoice/KhanhTTS. | TTS tiếng Việt + clone giọng, chạy offline 100%.
CapCap is a Windows desktop app for short-form video localization and dubbing. It brings transcription, translation, subtitle styling, voice generation, timeline editing, preview, and export into a single project workflow for Vietnamese-focused content production.