Store

Type:api
Platform:All
GPU:All
Tag:#captionx
mikecastrodemaria/moondream-3-improvedv5.0updated 3mo ago
A web interface for the Moondream3 vision-language model featuring image captioning, visual question answering, object detection, and object pointing.
@supersoniquestudio2 check-insNVIDIAAMDApple
manat0912/AI-Video-Clipper-LoRA-Pinokiov5.2updated 28d ago
Automatically clip videos and generate captions for LoRA training using advanced vision models like Gemma-3, Qwen3-VL, and Qwen2-VL.
@manatheturipa1 check-inNVIDIAAMDApple