A personal fork of lllyasviel/Fooocus v2.5.5 with quality-of-life features: Save Preset, CivitAI Model Settings, LoRA trigger words, Embeddings panel, Wildcards editor, Vary-with-aspect-ratio, Custom Resolution, Asset Browser, Restart UI button.
Apps by @supersoniquestudio
13 totalQwen-Image studio + instruction image editing (Qwen-Image-Edit-2509), Fooocus-style, 100% local: txt2img, ESRGAN+Qwen refine upscale, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Reframe/outpaint, Face Swap, and an Edit tab (image + prompt). Fork of crispz-studio. https://github.com/mikecastrodemaria/crispz-qwen-edit
Standalone local hi-res fix: enlarge images with Real-ESRGAN and reinject clean detail with Z-Image Turbo img2img. 100% local — no ComfyUI, no SwarmUI, no cloud.
FLUX.1 Krea [dev] txt2img studio (Fooocus-style, 100% local): high-aesthetic text-to-image, ESRGAN+Flux refine upscale, single-file/Civitai Flux models, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Reframe/outpaint, Face Swap. Fork of crispz-studio. https://github.com/mikecastrodemaria/crispz-krea
Z-Image txt2img + upscaler/detailer studio (Fooocus-style, 100% local): txt2img, ESRGAN+Z-Image refine upscale, single-file/Civitai models, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Reframe/outpaint, Face Swap. https://github.com/mikecastrodemaria/crispz-studio
Pre-mastering & audio enhancement for AI-generated music. 12-stage processing chain with platform presets (Suno, Udio), before/after spectrogram, and broadcast-ready LUFS normalization.
Krea 2 txt2img studio (Fooocus-style, 100% local): txt2img, pure-ESRGAN upscale, single-file/Civitai models, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Face Swap. https://github.com/mikecastrodemaria/crispz-krea2
Pre-mastering & audio enhancement for AI-generated music. 12-stage processing chain with platform presets (Suno, Udio), before/after spectrogram, and broadcast-ready LUFS normalization.
A web interface for the Moondream3 vision-language model featuring image captioning, visual question answering, object detection, and object pointing.
FLUX.2 Klein 4B studio (Apache 2.0, 4 steps), Fooocus-style, 100% local: txt2img, multi-reference editing in the SAME pipeline (up to 4 refs, no second model), inpaint/outpaint, ESRGAN+refine upscale, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Face Swap. ~15 GB VRAM for everything. Note: klein is distilled, so negative prompts and CFG have no effect. Fork of crispz-qwen-edit. https://github.com/mikecastrodemaria/crispz-klein
Customizing Realistic Human Photos via Stacked ID Embedding with SDXL model switching support. Load any local SDXL .safetensors checkpoint directly from the UI.
Local, engine-agnostic comic book workshop: script -> flatplan -> pages -> balloons -> PDF. Draws through the crispz family apps (crispz-studio, crispz-klein, crispz-qwen-edit...) over the CLI protocol; panel shapes, break-the-frame cutouts, frameless elements, Visual Bible, storyboard, .md bundles. https://github.com/mikecastrodemaria/comics2crispz
Batch image captioning with Ollama vision models: composed prompts (type × length × options), image preview, editable captions, skip/overwrite/append, live progress. Web UI (NiceGUI) or Gradio UI on the same core. https://github.com/mikecastrodemaria/Captionz
