We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Global radar
From Images to High-Fidelity 3D Assets with Production-Ready PBR Material
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Official repo for paper "Structured 3D Latents for Scalable and Versatile 3D Generation" (CVPR'25 Spotlight).
Use OCR in Windows quickly and easily with Text Grab. With optional background process and notifications.
[SIGGRAPH 2025] One Model to Rig Them All: Diverse Skeleton Rigging with UniRig
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
Image edit, text to image, image upscale, remove watermark
Generate realistic dialogue from a script, using Dia!
[SIGGRAPH Asia 2022] VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
[SIGGRAPH Asia 2022] VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild
Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
Generate responsive pages and apps on HTML, Tailwind, Flutter and SwiftUI.
A Gradio-based web UI for voice cloning and voice design, powered by Qwen3-TTS & VibeVoice. Can use Whisper or VibeVoice-ASR for automatic transcription.
WeShopAI Fashion Model Pose Change. Change poses in photos.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
