mikecastrodemaria/Captionz.pinokiov7.0updated 1h ago
Batch image captioning with Ollama vision models: composed prompts (type × length × options), image preview, editable captions, skip/overwrite/append, live progress. Web UI (NiceGUI) or Gradio UI on the same core. https://github.com/mikecastrodemaria/Captionz
mikecastrodemaria/crispz-klein.pinokiov7.0updated 1h ago
FLUX.2 Klein 4B studio (Apache 2.0, 4 steps), Fooocus-style, 100% local: txt2img, multi-reference editing in the SAME pipeline (up to 4 refs, no second model), inpaint/outpaint, ESRGAN+refine upscale, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Face Swap. ~15 GB VRAM for everything. Note: klein is distilled, so negative prompts and CFG have no effect. Fork of crispz-qwen-edit. https://github.com/mikecastrodemaria/crispz-klein
Local generative video, image, and character training on Apple Silicon. Train face + voice LoRAs in-app. Q8 HQ for character clips. MLX native — no cloud, no API key.
An all-in-one, 100% local AI creative studio, director, and multi-track editor. Generate with MiniMax H3, LTX-2.5/2.3, Wan, Flux, Qwen, and more; turn an idea or song into a planned production; then finish it on the timeline. Requires an NVIDIA GPU (6GB+ VRAM).