A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows.
Enter a description for your character (using the trigger word “img”) and optionally upload face photos, then provide a list of prompts—one per line—for each scene. The app creates a series of imag...
Open-source alternative to Higgsfield AI — Free AI image generation & cinema studio with 20+ models (Flux, SDXL, Midjourney, Ideogram). Self-hosted, customizable, MIT licensed.
Open-source alternative to Higgsfield AI — Free AI image generation & cinema studio with 20+ models (Flux, SDXL, Midjourney, Ideogram). Self-hosted, customizable, MIT licensed.
An offline-first, AI-powered medical scribe application that transcribes consultations and generates structured SOAP notes entirely on your device. Built with React, Vite, and the RunAnywhere SDK, it demonstrates the power of private, high-performance local AI.
ThonburianTTS, a finetuned Thai TTS based on the E2-TTS and F5-TTS architectures, designed to improve pronunciation accuracy, alignment robustness, and zero-shot speaker adaptation for the Thai language