Caption interface using ollama
Batch image captioning with a vision model served by Ollama. One core, four front ends: a command line, a Tkinter desktop app (standard library only), a NiceGUI web UI (--ui web), and a Gradio app made for Hugging Face Spaces. The Ollama layer reuses proven patterns from crispz-studio (cz_ollama.py): vision detection through /api/show with a name-based fallback, JPEG downscaling before upload, stripping of blocks from "thinking" models, keep_alive / CPU mode so the model does not hog VRAM. Features Connect to any Ollama server (configurable URL, local or remote) Automatic list of installed mod
