[NVIDIA ONLY] [RTX 50 Support] Image generation, image editing and free-form manipulation with a VLM (Minimum Requirements 12GB VRAM / 32GB RAM Recommended Requirements 24GB VRAM / 48GB RAM)
Creations by @morpheus
105 total(WINDOWS)NVIDIA, Hallo2: Long-Duration and High-Resolution Audio-driven Portrait Image Animation
An intelligent, interactive Image Editing System. Easily erase and add objects on a user-friendly interface.
create a story by generating consistent images https://github.com/HVision-NKU/StoryDiffusion
An Open Source Model for Audio Samples and Sound Design https://github.com/Stability-AI/stable-audio-tools
Unify Efficient Fine-Tuning of 100+ LLMs https://github.com/hiyouga/LLaMA-Factory
[NVIDIA ONLY] High-Quality and Efficient 3D Mesh Generation from a Single Image (Minimum requirements 12GB VRAM / 24GB RAM)
Automatically create music videos. Synchronize the cuts to the music's beat. (Compatible with all OS and GPU's)
A simple, high-quality image generation tool to create stunning illusions.
Frontier Open-Source Text-to-Speech
All-in-One Video Creation and Editing. Move-Anything, Swap-Anything, Reference-Anything, Expand-Anything, Animate-Anything.
[NVIDIA ONLY] End-to-end multimodal SVG generator capable of generating complex and detailed SVGs, from simple icons to intricate anime characters. (Minimum Requirements 12GB VRAM / 32GB RAM, Recommended Requirements 24GB VRAM / 24GB RAM)
Fooocus powered by FastAPI
diffusers InstantID + ControlNet inspired by face-to-many from fofr (https://x.com/fofrAI) - a localized Version of https://huggingface.co/spaces/multimodalart/face-to-all
[NVIDIA ONLY] Gradio demo for Flux Kontext based on Diffusers with single and multiple images.
Official Implementations for Paper - MagicQuill: An Intelligent Interactive Image Editing System
Unified Image Understanding and Generation. Text-to-Image Generation, In-context Generation, Instruction-guided Image Editing, Visual Understanding (Minimum Requirements 12GBV RAM / 48GB RAM, Recommended Requirements 24GB VRAM / 32GB RAM)
Image inpainting tool powered by SOTA AI models. Remove any unwanted object, defect, or even people from your pictures, and replace (powered by stable diffusion) anything in your pictures. https://www.iopaint.com/
LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation https://huggingface.co/spaces/ashawkey/LGM
[NVIDIA ONLY] Compose anything in video diffusion transformers. (Requires 24GB VRAM)
