An Open Source Model for Audio Samples and Sound Design https://github.com/Stability-AI/stable-audio-tools
Creations by @morpheus
104 totalUnify Efficient Fine-Tuning of 100+ LLMs https://github.com/hiyouga/LLaMA-Factory
[NVIDIA ONLY] High-Quality and Efficient 3D Mesh Generation from a Single Image (Minimum requirements 12GB VRAM / 24GB RAM)
Automatically create music videos. Synchronize the cuts to the music's beat. (Compatible with all OS and GPU's)
[NVIDIA ONLY] Text-driven, intelligent restoration, blending AI technology with creativity to give every image a brand new life https://supir.xpixel.group
A simple, high-quality image generation tool to create stunning illusions.
Stable Diffusion Trainer: https://github.com/bmaltais/kohya_ss
Frontier Open-Source Text-to-Speech
All-in-One Video Creation and Editing. Move-Anything, Swap-Anything, Reference-Anything, Expand-Anything, Animate-Anything.
[NVIDIA ONLY] End-to-end multimodal SVG generator capable of generating complex and detailed SVGs, from simple icons to intricate anime characters. (Minimum Requirements 12GB VRAM / 32GB RAM, Recommended Requirements 24GB VRAM / 24GB RAM)
Fooocus powered by FastAPI
diffusers InstantID + ControlNet inspired by face-to-many from fofr (https://x.com/fofrAI) - a localized Version of https://huggingface.co/spaces/multimodalart/face-to-all
[NVIDIA ONLY] Gradio demo for Flux Kontext based on Diffusers with single and multiple images.
Official Implementations for Paper - MagicQuill: An Intelligent Interactive Image Editing System
Unified Image Understanding and Generation. Text-to-Image Generation, In-context Generation, Instruction-guided Image Editing, Visual Understanding (Minimum Requirements 12GBV RAM / 48GB RAM, Recommended Requirements 24GB VRAM / 32GB RAM)
Image inpainting tool powered by SOTA AI models. Remove any unwanted object, defect, or even people from your pictures, and replace (powered by stable diffusion) anything in your pictures. https://www.iopaint.com/
LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation https://huggingface.co/spaces/ashawkey/LGM
[NVIDIA ONLY] Compose anything in video diffusion transformers. (Requires 24GB VRAM)
[NVIDIA ONLY] Remove Objects in videos with inpainting. Recommended requirements 16 - 24 GB VRAM / 48 GB RAM, Minimal requirements 12GB VRAM / 32 GBRAM
An enhanced version of Fooocus giving you access to all of the latest AI image generation models
