Upload a clean 20 seconds WAV file of the vocal persona you want to mimic, type your text-to-speech prompt and hit submit! A local version of https://huggingface.co/spaces/fffiloni/instant-TTS-Bark-cloning
Projects by @morpheus
74 total[NVIDIA ONLY] High-Quality and Efficient 3D Mesh Generation from a Single Image (Minimum requirements 12GB VRAM / 24GB RAM)
Video to 3D: 4D Face Reconstruction from any Video or Image Sequence. Normal Map, Depth Map and 3D Mesh Generation.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface. https://github.com/comfyanonymous/ComfyUI
Create 3D Meshes of Body Poses from Images.
Automatically remove watermarks from videos generated by Sora AI.
[NVIDIA Only] Dead simple web UI for training FLUX LoRA with LOW VRAM support (From 12GB)
[NVIDIA ONLY] Text-driven, intelligent restoration, blending AI technology with creativity to give every image a brand new life https://supir.xpixel.group
An open-source, modern-design ChatGPT/LLMs UI/Framework. Supports speech-synthesis, multi-modal, and extensible (function call) plugin system. https://github.com/lobehub/lobe-chat
User-friendly WebUI for LLMs, supported LLM runners include Ollama and OpenAI-compatible APIs https://github.com/open-webui/open-webui
All in one Gradio interface for chatterbox. Voice cloning from uploaded audio samples, automatic text processing for long content and real-time speech generation with configurable parameters. (Minimum Requirements 4GB VRAM / Recommended Requirements 8GB VRAM)
[NVIDIA ONLY] Gradio demo for Flux Kontext based on Diffusers with single and multiple images.
Automatically create music videos. Synchronize the cuts to the music's beat.
Kimodo generates high-quality 3D human and robot motions and is controlled through text prompts
(WINDOWS)NVIDIA, Hallo2: Long-Duration and High-Resolution Audio-driven Portrait Image Animation
A professional, Suno-like music generation studio for HeartLib. https://github.com/fspecii/HeartMuLa-Studio
Fooocus powered by FastAPI
Generating Consistent Long Depth Sequences for Open-world Videos
Local GPU-accelerated music video generator: Gradio UI, analysis, SDXL backgrounds, NVENC output.
High-Quality Text-to-Speech for Indian Languages
