Launcher updates

More
devnen/Chatterbox-TTS-Serverupdated 7mo ago
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.
0 check-insNVIDIAAMDApple
6Morpheus6/photomaker2v3.7updated 7mo ago
Customizing Realistic Human Photos via Stacked ID Embedding https://huggingface.co/spaces/TencentARC/PhotoMaker-V2
@morpheus1 check-inNVIDIAAMDApple
tonykipkemboi/ollama_pdf_ragupdated 7mo ago
A full-stack demo showcasing a local RAG (Retrieval Augmented Generation) pipeline to chat with your PDFs.
0 check-insNVIDIAAMDApple
PierrunoYT/Kokoro-TTS-Localupdated 7mo ago
A local implementation of the Kokoro Text-to-Speech model, featuring dynamic module loading, automatic dependency management, and a web interface.
@pierrunoyt0 check-insNVIDIAAMDApple
facebookresearch/sam-3d-bodyupdated 7mo ago
The repository provides code for running inference with the SAM 3D Body Model (3DB), links for downloading the trained model checkpoints and datasets, and example notebooks that show how to use the model.
0 check-insNVIDIAAMDApple
SakanaAI/AI-Scientist-v2updated 7mo ago
The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
0 check-insNVIDIAAMDApple
SakanaAI/AI-Scientistupdated 7mo ago
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery 🧑‍🔬
0 check-insNVIDIAAMDApple
apple/ml-sharpupdated 7mo ago
Sharp Monocular View Synthesis in Less Than a Second
0 check-insNVIDIAAMDApple
AIM-Intelligence/video2robotupdated 7mo ago
End-to-end pipeline converting generative videos (Veo, Sora) to humanoid robot motions
0 check-insNVIDIAAMDApple
MeiGen-AI/InfiniteTalkupdated 7mo ago
​​Unlimited-length talking video generation​​ that supports image-to-video and video-to-video generation
0 check-insNVIDIAAMDApple
MeiGen-AI/MultiTalkupdated 7mo ago
[NeurIPS 2025] Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
0 check-insNVIDIAAMDApple
meigen-ai/longcat-video-avatarupdated 7mo ago
The homepage of LongCat-Video-Avatar
0 check-insNVIDIAAMDApple
Ordinary0x/The-3rd-Eyeupdated 7mo ago
The 3rd Eye is a modular OSINT (Open Source Intelligence) framework built on an agent-based, graph-driven architecture. It automates public information discovery, identity correlation, and exposure analysis across multiple platforms, and generates structured intelligence reports. The system follows a LangGraph agent design.
0 check-insNVIDIAAMDApple
hendrybui/facefusionupdated 7mo ago
Industry leading face manipulation platform
0 check-insNVIDIAAMDApple
Tencent-Hunyuan/HunyuanWorld-1.0updated 7mo ago
Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model
0 check-insNVIDIAAMDApple
bytedance/Dolphinupdated 7mo ago
The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.
0 check-insNVIDIAAMDApple
zyddnys/manga-image-translatorupdated 7mo ago
Translate manga/image 一键翻译各类图片内文字 https://cotrans.touhou.ai/ (no longer working)
0 check-insNVIDIAAMDApple
mmehmetisik/ai-text-to-image-generatorupdated 7mo ago
AI-powered image generation tool using Hugging Face API and Stable Diffusion. Create images from text prompts with multiple style options.
0 check-insNVIDIAAMDApple
Light-x02/ComfyUI-Civitai-Discovery-Hubupdated 7mo ago
This ComfyUI node lets you browse the Civitai gallery directly within the interface, featuring infinite scroll, advanced filters (including NSFW), and favorites management. It also allows you to retrieve prompts, metadata, and images/videos to seamlessly reuse them in your workflows.
0 check-insNVIDIAAMDApple
Stability-AI/generative-modelsupdated 7mo ago
Generative Models by Stability AI
0 check-insNVIDIAAMDApple