An all-in-one, 100% local AI creative studio, director, and multi-track editor. Generate with MiniMax H3, LTX-2.5/2.3, Wan, Flux, Qwen, and more; turn an idea or song into a planned production; then finish it on the timeline. Requires an NVIDIA GPU (6GB+ VRAM).
Local generative video, image, and character training on Apple Silicon. Train face + voice LoRAs in-app. Q8 HQ for character clips. MLX native — no cloud, no API key.
1-click WanGP Launcher. Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP
pinokiofactory/stable-diffusion-webui-forgev2.0updated 2d ago
[NVIDIA ONLY] The most efficient way to run FLUX (Optimized to run even on low memory machines, as low as 3GB VRAM with 512x512 resolution) https://github.com/lllyasviel/stable-diffusion-webui-forge
YUE2 // GROOVE is the latest music studio built on the open-source Yue2 model and its inference stack. Generate high-quality full songs from style and lyrics with an editable score plan — powered by the latest YuE model — cover from audio with SheetSage2 and MERT2, refine and compare edits, and keep your works in a reusable, easy-to-manage library. You get high-quality creation with real creative control. Hardware: an NVIDIA GPU with 24 GB VRAM on Linux (YuE2's recommended setup, validated end-to-end on NVIDIA L4 hosts; a 16 GB memory budget runs everything the app can produce, and 12 GB runs the unquantized model at CFG 1.0 or for shorter songs) or an Apple Silicon Mac with 32 GB+ unified memory (where this app is developed and tested). Windows is best effort: install, launch, Cover and a full-length song verified on Windows 11 (RTX 2070, 8 GB); song generation there runs through the GGUF engine. Cards under 16 GB (and Windows) get the optional GGUF engine: the same model through yue2.cpp with an 8-bit backbone, 8.2 GB peak for a full song, measured indistinguishable from the reference rendering in a blind ABX, a different take for the same seed; the reference PyTorch configuration stays the default wherever it fits.
Mac-optimized YuE2 music generation using native MLX inference. Generate complete songs (melody, chords, vocals, accompaniment) from text prompts. Features an LLM Writing Room for lyric composition via OpenAI-compatible APIs (LM Studio, etc.). Pure Apple Silicon — no PyTorch needed.
[AMD ONLY] Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video, Flux and more. (On Windows supported by all dedicated AMD GPUs from RDNA 2 - RDNA 4)
Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP
A fully local, cross-platform audio visualizer editor. Create reactive music videos with layered graphics, AI-transcribed lyrics, and frame-perfect MP4 exports — all running in your browser
MiniMax H3 omni-modal video generation in ComfyUI. Text/image/video/audio in, video with native 32kHz stereo audio out (768p default, 1080p+ supported). Disk-optimized: pruned INT8 + NVFP4 weights (~63GB instead of ~290GB). NVIDIA only.
Uncensored deepfakes for images and videos, no training required. Advanced masking, batch processing, and face enhancement powered by InsightFace and ONNX. Supports NVIDIA (CUDA/TensorRT), AMD (DirectML/ROCm), Apple Silicon, and CPU.