mikecastrodemaria/Captionz.pinokiov7.0updated 1h ago
Batch image captioning with Ollama vision models: composed prompts (type × length × options), image preview, editable captions, skip/overwrite/append, live progress. Web UI (NiceGUI) or Gradio UI on the same core. https://github.com/mikecastrodemaria/Captionz
mikecastrodemaria/crispz-klein.pinokiov7.0updated 2h ago
FLUX.2 Klein 4B studio (Apache 2.0, 4 steps), Fooocus-style, 100% local: txt2img, multi-reference editing in the SAME pipeline (up to 4 refs, no second model), inpaint/outpaint, ESRGAN+refine upscale, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Face Swap. ~15 GB VRAM for everything. Note: klein is distilled, so negative prompts and CFG have no effect. Fork of crispz-qwen-edit. https://github.com/mikecastrodemaria/crispz-klein
Local generative video, image, and character training on Apple Silicon. Train face + voice LoRAs in-app. Q8 HQ for character clips. MLX native — no cloud, no API key.