MiniMax H3 omni-modal video generation in ComfyUI. Text/image/video/audio in, video with native 32kHz stereo audio out (768p default, 1080p+ supported). Disk-optimized: pruned INT8 + NVFP4 weights (~63GB instead of ~290GB). NVIDIA only.
An all-in-one, 100% local AI creative studio, director, and multi-track editor. Generate with MiniMax H3, LTX-2.5/2.3, Wan, Flux, Qwen, and more; turn an idea or song into a planned production; then finish it on the timeline. Requires an NVIDIA GPU (6GB+ VRAM).
cocktailpeanut/stabledaw.pinokiov7.0updated 18d ago
Browser-based AI audio DAW for Stable Audio 3 with text-to-audio, inpainting, LoRA training, FFmpeg effects, waveform editing, sequencer, piano roll, and persistent library. https://github.com/gantasmo/stabledaw
[AMD ONLY] Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video, Flux and more. (On Windows supported by all dedicated AMD GPUs from RDNA 2 - RDNA 4)
cocktailpeanut/stable-audio-3-small.pinokiov7.0updated 12d ago
Launcher for Stable Audio 3 Small Music, Small SFX, and NVIDIA Medium using public cocktailpeanut Hugging Face mirrors. https://github.com/Stability-AI/stable-audio-3
cocktailpeanut/worldmirror.pinokiov7.0updated 4mo ago
[NVIDIA] Pinokio launcher for the released WorldMirror 2.0 reconstruction app from HY-World 2.0. Uses a cu128 PyTorch baseline with gsplat from PyPI/JIT. https://github.com/Tencent-Hunyuan/HY-World-2.0
Professional-grade browser-based video editor with multi-track editing, keyframe animations, real-time preview, and high-quality exports. No uploads — everything runs locally.
Blizaine/Qwen3-TTS-MLX-WebUI-Enhancedv5.0updated 1mo ago
High-quality text-to-speech with Beautiful Web UI & API, optimized for Apple Silicon using MLX. Features include Custom Voice (preset speakers), Voice Design (natural language), and Voice Cloning. With enhanced features for saving custom voices and long-form / endless TTS streaming.