Wanted
1,906 projectsNon-launcher projects without a Pinokio launcher yet.
LiveCC: Learning Video LLM with Streaming Speech Transcription at Scale (CVPR 2025)
Aplikasi ini digunakan untuk menghasilkan suara berbasis teks dengan berbagai pilihan pembicara. Teknologi yang digunakan meliputi model text-to-speech (TTS) yang canggih dengan konversi teks ke fonem. Model yang dipakai dilatih khusus untuk bahasa Indonesia, Jawa dan Sunda.
<⚡️> SuperAGI - A dev-first open source autonomous AI agent framework. Enabling developers to build, manage & run useful autonomous agents quickly and reliably.
Get up and running with GLM-4.7, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Contribute to sukebenet/instruct-pix2pix development by creating an account on GitHub.
Fully local AI vtuber that can see your screen and talk in real time
Contribute to justinram11/qwen3-tts-rocm development by creating an account on GitHub.
Code for SIGGRAPH 2020 paper "RigNet: Neural Rigging for Articulated Characters"
Native and Compact Structured Latents for 3D Generation
Magenta: Music and Art Generation with Machine Intelligence
Self-hosted AI audio transcription
NVIDIA FastGen: Fast Generation from Diffusion Models
Di♪♪Rhythm 2: Efficient And High Fidelity Song Generation Via Block Flow Matching
AIMedia 是一款自动抓取热点,AI创作文章,自动发布的集成软件。支持头条,小红书,公众号等
自动化上传视频到社交媒体:抖音、小红书、视频号、tiktok、youtube、bilibili
Optimized Retrieval-based Voice Conversion WebUI for Apple Silicon Macs (M1/M2/M3). Real-time, high-quality voice conversion with an easy web interface. All models included!
This is an official PyTorch implementation of our NeurIPS 2023 paper "GeoCLIP: Clip-Inspired Alignment between Locations and Images for Effective Worldwide Geo-localization"
Contribute to pinokiocomputer/python development by creating an account on GitHub.
Python script for extracting faces from video files
Creates 256 cropped, high quality facesets from a batch of mp4 videos end to end
