Minimal launcher to verify macOS uv-runtime FFmpeg dylib exposure.
Projects on Pinokio
Secure Workflow Automation for Technical Teams
A minimalist todo list with a lightweight JSON API and local storage.
Contribute to peanutcocktail/prototype development by creating an account on GitHub.
github
Roblox Foundation Model for 3D Intelligence
EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
Select a portrait, click to move the head around (please use your own space / GPU!)
Text-to-video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
Various AI scripts. Mostly Stable Diffusion stuff.
Instant voice cloning by MIT and MyShell.
MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model
馃 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
Generative models for conditional audio generation
High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.
[CVPR 2024] Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data. Foundation Model for Monocular Depth Estimation
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
