Z-Image txt2img + upscaler/detailer studio (Fooocus-style, 100% local): txt2img, ESRGAN+Z-Image refine upscale, single-file/Civitai models, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Reframe/outpaint, Face Swap. https://github.com/mikecastrodemaria/crispz-studio
Creations by @supersoniquestudio
16 totalQwen-Image-Edit instruction-based image editing studio — fork of crispz-studio
FLUX.1 Krea [dev] txt2img studio — fork of crispz-studio (Fooocus-style, 100% local)
Qwen-Image studio + instruction image editing (Qwen-Image-Edit-2509), Fooocus-style, 100% local: txt2img, ESRGAN+Qwen refine upscale, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Reframe/outpaint, Face Swap, and an Edit tab (image + prompt). Fork of crispz-studio. https://github.com/mikecastrodemaria/crispz-qwen-edit
FLUX.1 Krea [dev] txt2img studio (Fooocus-style, 100% local): high-aesthetic text-to-image, ESRGAN+Flux refine upscale, single-file/Civitai Flux models, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Reframe/outpaint, Face Swap. Fork of crispz-studio. https://github.com/mikecastrodemaria/crispz-krea
A personal fork of lllyasviel/Fooocus v2.5.5 with quality-of-life features: Save Preset, CivitAI Model Settings, LoRA trigger words, Embeddings panel, Wildcards editor, Vary-with-aspect-ratio, Custom Resolution, Asset Browser, Restart UI button.
Z-Image txt2img + upscaler/detailer studio (Fooocus-style fork of crispz)
Contribute to mikecastrodemaria/RhythmBeatDetection development by creating an account on GitHub.
Pre-mastering & audio enhancement for AI-generated music. 12-stage processing chain with platform presets (Suno, Udio), before/after spectrogram, and broadcast-ready LUFS normalization.
A standalone, fully local image upscaler and denoiser — a "hi-res fix" tool that enlarges images and reinjects clean detail without altering the composition. No ComfyUI, no SwarmUI, no cloud. Everything runs on your own machine, which matters for sensitive client work.
Standalone local hi-res fix: enlarge images with Real-ESRGAN and reinject clean detail with Z-Image Turbo img2img. 100% local — no ComfyUI, no SwarmUI, no cloud.
A web interface for the Moondream3 vision-language model featuring image captioning, visual question answering, object detection, and object pointing.
Focus on prompting and generating
Pre-mastering & audio enhancement for AI-generated music. 12-stage processing chain with platform presets (Suno, Udio), before/after spectrogram, and broadcast-ready LUFS normalization.
Pre-mastering & audio enhancement for AI-generated music. 12-stage processing chain with platform presets (Suno, Udio), before/after spectrogram, and broadcast-ready LUFS normalization.
Customizing Realistic Human Photos via Stacked ID Embedding with SDXL model switching support. Load any local SDXL .safetensors checkpoint directly from the UI.
