AI Song Generation on Mac Apple Silicon, with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model.
Enter a face image and transform it to any other image. Demo for the h94/IP-Adapter-FaceID model https://huggingface.co/spaces/multimodalart/Ip-Adapter-FaceID
Text-to-Speech for 16 Indian languages: Assamese, Bengali, Bodo, English (Indian accent), Hinglish, Gujarati, Hindi, Kannada, Malayalam, Manipuri, Marathi, Odia, Punjabi, Rajasthani, Tamil, and Telugu. SOTA models based on FastPitch and HiFi-GAN V1.
[NVIDIA ONLY] Generate Video Progressively. FramePack is a next-frame (next-frame-section) prediction neural network structure that generates videos progressively. https://github.com/lllyasviel/FramePack
A unified image generation model that you can use to perform various tasks, including but not limited to text-to-image generation, subject-driven generation, Identity-Preserving Generation, and image-conditioned generation. https://huggingface.co/spaces/Shitao/OmniGen
hoodtronik/musubi-tuner.pinokiov2.0updated 3mo ago
Train LoRA / LoHa / LoKr for Wan2.2, FLUX.2, Z-Image, HunyuanVideo, and more — one-click install of kohya-ss/musubi-tuner with its built-in Gradio GUI.