An all-in-one, 100% local AI creative studio, director, and multi-track editor. Generate with MiniMax H3, LTX-2.5/2.3, Wan, Flux, Qwen, and more; turn an idea or song into a planned production; then finish it on the timeline. Requires an NVIDIA GPU (6GB+ VRAM).
MiniMax H3 omni-modal video generation in ComfyUI. Text/image/video/audio in, video with native 32kHz stereo audio out (768p default, 1080p+ supported). Disk-optimized: pruned INT8 + NVFP4 weights (~63GB instead of ~290GB). NVIDIA only.
[AMD ONLY] Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video, Flux and more. (On Windows supported by all dedicated AMD GPUs from RDNA 2 - RDNA 4)
Video translation & dubbing with voice cloning — 100% local, zero API. Supports 30 languages (VoxCPM 2), YouTube SEO Studio, Viral Shorts Studio (9:16), and WordPress SEO Blog Post Generator.
Unofficial ChatGPT desktop app for Linux (formerly the Codex app), built locally from OpenAI's official Linux package. Includes Chat, Work, and Codex. https://github.com/ilysenko/codex-desktop-linux
Pixal3D image-to-3D generation with a clean, simplified web UI powered by ComfyUI. Turn a single reference image into a textured 3D model (GLB) — tuned for 8-12GB VRAM with RAM offloading. https://github.com/PrimeEcto/PIXAL-3D
pinokiofactory/stable-diffusion-webui-forgev2.0updated 9d ago
[NVIDIA ONLY] The most efficient way to run FLUX (Optimized to run even on low memory machines, as low as 3GB VRAM with 512x512 resolution) https://github.com/lllyasviel/stable-diffusion-webui-forge
1-click WanGP Launcher. Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP
cocktailpeanut/stable-audio-3-small.pinokiov7.0updated 12d ago
Launcher for Stable Audio 3 Small Music, Small SFX, and NVIDIA Medium using public cocktailpeanut Hugging Face mirrors. https://github.com/Stability-AI/stable-audio-3
Fully offline speech-to-text transcription with 11 local AI models. Generates styled subtitles (SRT, ASS) and burns them directly onto video. Supports Whisper, Parakeet, Canary, Moonshine, SenseVoice, Vosk, and more. No API keys required.