@sup3rmass1ve
Creations by @sup3rmass1ve
59 totalJanus Pro 7B is a powerful multimodal AI model designed for advanced image understanding and text-to-image generation.
Generate realistic and expressive speech with natural language voice design.
A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics,
SongBloom, a novel framework for full-length song generation
(WINDOWS)NVIDIA, Hallo2: Long-Duration and High-Resolution Audio-driven Portrait Image Animation
NeuTTS Air is the world’s first super-realistic, on-device, TTS speech language model with instant voice cloning. Built off a 0.5B
Fast and High-Quality Zero-Shot voice clone Text-to-Speech with Flow Matching
Fast and High-Quality Zero-Shot voice clone Text-to-Speech with Flow Matching Multilingual
SoTA open-source TTS
A powerful tool for extending images to different aspect ratios using Stable Diffusion XL.
DreamO: A Unified Framework for Image Customization
Real Time Speech Transcription
interacting with the Ovis2-8B model. The script allows users to load the model, process image and video inputs, and generate text-based responses using a conversational chatbot.
MegaTTS app
Higgs Audio Text-to-Speech Playground
Simple, scalable AI model deployment on GPU clusters
