Launcher updates

More
dangvansam/viet-ttsupdated 8mo ago
VietTTS: An Open-Source Vietnamese Text to Speech
0 check-insNVIDIAAMDApple
travisvn/chatterbox-tts-apiupdated 8mo ago
Local, OpenAI-compatible text-to-speech (TTS) API using Chatterbox, enabling users to generate voice cloned speech anywhere the OpenAI API is used (e.g. Open WebUI, AnythingLLM, etc.)
0 check-insNVIDIAAMDApple
leafspark/AutoGGUFupdated 8mo ago
automatically quant GGUF models
0 check-insNVIDIAAMDApple
platomav/MEAnalyzerupdated 8mo ago
Intel Engine & Graphics Firmware Analysis Tool
0 check-insNVIDIAAMDApple
changetheconstants/seedvarianceenhancerupdated 8mo ago
A ComfyUI node that adds random noise to text embeddings.
0 check-insNVIDIAAMDApple
V-Sekai-fire/pinokio-image-to-3dv1.0.0updated 8mo ago
ComfyUI with TRELLIS2, GeometryPack, and UniRig custom nodes for image-to-3D generation
1 check-inNVIDIAAMDApple
serpotapov/stable-diffusion-portableupdated 8mo ago
Stable Diffusion Portable
0 check-insNVIDIAAMDApple
devnen/Chatterbox-TTS-Serverupdated 8mo ago
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.
0 check-insNVIDIAAMDApple
6Morpheus6/photomaker2v3.7updated 8mo ago
Customizing Realistic Human Photos via Stacked ID Embedding https://huggingface.co/spaces/TencentARC/PhotoMaker-V2
@morpheus1 check-inNVIDIAAMDApple
tonykipkemboi/ollama_pdf_ragupdated 8mo ago
A full-stack demo showcasing a local RAG (Retrieval Augmented Generation) pipeline to chat with your PDFs.
0 check-insNVIDIAAMDApple
PierrunoYT/Kokoro-TTS-Localupdated 8mo ago
A local implementation of the Kokoro Text-to-Speech model, featuring dynamic module loading, automatic dependency management, and a web interface.
@pierrunoyt0 check-insNVIDIAAMDApple
facebookresearch/sam-3d-bodyupdated 8mo ago
The repository provides code for running inference with the SAM 3D Body Model (3DB), links for downloading the trained model checkpoints and datasets, and example notebooks that show how to use the model.
0 check-insNVIDIAAMDApple
SakanaAI/AI-Scientist-v2updated 8mo ago
The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
0 check-insNVIDIAAMDApple
SakanaAI/AI-Scientistupdated 8mo ago
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery 🧑‍🔬
0 check-insNVIDIAAMDApple
apple/ml-sharpupdated 8mo ago
Sharp Monocular View Synthesis in Less Than a Second
0 check-insNVIDIAAMDApple
AIM-Intelligence/video2robotupdated 8mo ago
End-to-end pipeline converting generative videos (Veo, Sora) to humanoid robot motions
0 check-insNVIDIAAMDApple
MeiGen-AI/InfiniteTalkupdated 8mo ago
​​Unlimited-length talking video generation​​ that supports image-to-video and video-to-video generation
0 check-insNVIDIAAMDApple
MeiGen-AI/MultiTalkupdated 8mo ago
[NeurIPS 2025] Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
0 check-insNVIDIAAMDApple
meigen-ai/longcat-video-avatarupdated 8mo ago
The homepage of LongCat-Video-Avatar
0 check-insNVIDIAAMDApple
Ordinary0x/The-3rd-Eyeupdated 8mo ago
The 3rd Eye is a modular OSINT (Open Source Intelligence) framework built on an agent-based, graph-driven architecture. It automates public information discovery, identity correlation, and exposure analysis across multiple platforms, and generates structured intelligence reports. The system follows a LangGraph agent design.
0 check-insNVIDIAAMDApple