A Gradio app with Rerun visualization for Microsoft's TRELLIS.2-4B model that generates textured 3D assets (GLB) from text or images using a two-stage pipeline: text-to-image (Z-Image-Turbo) then image-to-3D (TRELLIS.2).
Global radar
An AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction, etc.
Fabric is an open-source framework for augmenting humans using AI. It provides a modular system for solving specific problems using a crowdsourced set of AI prompts that can be used anywhere.
Upload an audio file to improve its quality and reduce background noise. Choose settings for better results.
Clone a voice in 5 seconds to generate arbitrary speech in real-time
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Effortless data labeling with AI support from Segment Anything and other awesome models.
A Conversational Speech Generation Model with Gradio UI and OpenAI compatible API. UI and API support CUDA, MLX and CPU devices.
Enterprise-grade AI marketing automation for Claude Code, Cursor, GitHub Copilot, and any AI assistant supporting agents & skills
Convert any text to a graph of knowledge. This can be used for Graph Augmented Generation or Knowledge Graph based QnA
This application lets you upload an image and generate a caption tailored to your choice of style and length. You can select from options like descriptive, informal, or specific formats like traini...
