Local-first AI video intelligence platform. Index your video library with multi-modal analysis (YOLO, DeepFace, Whisper), search semantically via natural language, Docker-ready.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
[NVIDIA ONLY] Super Optimized Gradio UI for Wan2.1 video for GPU poor machines (5GB+ VRAM). Generate up to 12 sec videos https://github.com/deepbeepmeep/Wan2GP
A cross-platform desktop application for running AI models from [WaveSpeedAI](https://wavespeed.ai), as well as many free local AI models including Z-Image.
ElWalki/ProdIA_Max-Ace-Step-UI_Ace-Step-v1.5updated 4mo ago
ProdIA-MAX is an advanced fork of fspecii/ace-step-ui paired with ACE-Step v1.5. Adds AI Music Chat Assistant, visual Chord Progression Editor, floating LoRA Manager, Demucs vocal separation, automatic ID3 tagging, full i18n support (EN/ES/ZH/JA/KO), extended presets, section-by-section generation, and multiple backend reliability fixes.