Enter a task description and optionally upload one piece of media—an audio clip, a picture, or a short video. The app sends your inputs to the Qwen3.5 Omni model, which replies with a streamed text...
mikecastrodemaria/Captionz.pinokiov7.0updated 16d ago
Batch image captioning with Ollama vision models: composed prompts (type × length × options), image preview, editable captions, skip/overwrite/append, live progress. Web UI (NiceGUI) or Gradio UI on the same core. https://github.com/mikecastrodemaria/Captionz
This is an extension based on sd-webui, aimed at improving the user experience of the prompt/negative prompt input box. It has a more intuitive and powerful input interface function, and provides automatic translation, history record, and bookmarking functions. 这是一个基于 sd-webui 的扩展,旨在提高提示词/反向提示词输入框的使用体验。它拥有更直观、强大的输入界面功能,它提供了自动翻译、历史记录和收藏等功能。