An all-in-one, 100% local AI video, image & music studio. Its Director mode turns a single prompt into a full music video or short film — LLM-planned, shot by shot. Built on the WanGP pipeline (Wan 2.1/2.2, LTX-2.3, Qwen, Hunyuan Video, Flux). Requires an NVIDIA GPU (6GB+ VRAM).
[AMD ONLY] All-in-one local AI video, image & music studio built on the WanGP pipeline (Wan 2.1/2.2, LTX-2.3, Qwen, Hunyuan Video, Flux). ROCm-powered. Supported on Windows and Linux for RDNA 2-4 dGPUs (RX 6000/7000/8000/9000) and gfx1150-53 APUs (Strix Point/Halo, Krackan Point/Halo). 6 GB+ VRAM recommended. For NVIDIA GPUs use the upstream Maestro app. Credits: the app itself is Maestro by Blizaine (github.com/Blizaine/Maestro) — this is just an AMD installer wrapper around it. The clone-at-install-time approach used to port it to AMD was pioneered by 6Morpheus6's wan2gp-amd (github.com/6Morpheus6/wan2gp-amd).
MiniMax H3 omni-modal video generation in ComfyUI. Text/image/video/audio in, video with native 32kHz stereo audio out (768p default, 1080p+ supported). Disk-optimized: pruned INT8 + NVFP4 weights (~63GB instead of ~290GB). NVIDIA only.
timoncool/VoxCPM2_portable-pinokiov6.0.0updated 20h ago
ElevenLabs at home. Multilingual TTS with Voice Design, Voice Cloning, and end-to-end LoRA fine-tuning straight from a video or podcast. Built on VoxCPM2 by OpenBMB. 30 languages incl. Russian.
Local LoRA trainer for MiniMax H3, Flux2 Klein 9B & Krea 2 — trains against the exact models you deploy. Desktop GUI + headless CLI. NVIDIA only (RTX 30/40/50, driver 555+).
A god roleplaying sandbox — AI characters talk to each other with automated conversations, character development, and god interference. Works with any OpenAI-compatible local LLM (e.g. LM Studio).
Local generative video, image, and character training on Apple Silicon. Train face + voice LoRAs in-app. Q8 HQ for character clips. MLX native — no cloud, no API key.
Describe an image, get a 100% schema-valid Ideogram 4 JSON prompt — generated fully locally with an embedded llama.cpp (no Ollama or LM Studio required).
Uncensored deepfakes for images and videos, no training required. Advanced masking, batch processing, and face enhancement powered by InsightFace and ONNX. Supports NVIDIA (CUDA/TensorRT), AMD (DirectML/ROCm), Apple Silicon, and CPU.
1-click WanGP Launcher. Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP