An all-in-one, 100% local AI creative studio, director, and multi-track editor. Generate with MiniMax H3, LTX-2.5/2.3, Wan, Flux, Qwen, and more; turn an idea or song into a planned production; then finish it on the timeline. Requires an NVIDIA GPU (6GB+ VRAM).
Local generative video, image, and character training on Apple Silicon. Train face + voice LoRAs in-app. Q8 HQ for character clips. MLX native — no cloud, no API key.
Build, curate, caption and clean LoRA training datasets in your browser, then train, compare and iterate — local Flask app, your files stay on your disk. https://github.com/perfectgf/lora-dataset-studio
timoncool/VoxCPM2_portable-pinokiov6.0.0updated 13h ago
ElevenLabs at home. Multilingual TTS with Voice Design, Voice Cloning, and end-to-end LoRA fine-tuning straight from a video or podcast. Built on VoxCPM2 by OpenBMB. 30 languages incl. Russian.
1-click WanGP Launcher. Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP
pinokiofactory/stable-diffusion-webui-forgev2.0updated 2d ago
[NVIDIA ONLY] The most efficient way to run FLUX (Optimized to run even on low memory machines, as low as 3GB VRAM with 512x512 resolution) https://github.com/lllyasviel/stable-diffusion-webui-forge
YUE2 // GROOVE is the latest music studio built on the open-source Yue2 model and its inference stack. Generate high-quality full songs from style and lyrics with an editable score plan — powered by the latest YuE model — cover from audio with SheetSage2 and MERT2, refine and compare edits, and keep your works in a reusable, easy-to-manage library. You get high-quality creation with real creative control. Hardware: an NVIDIA GPU with 24 GB VRAM on Linux (YuE2's recommended setup, validated end-to-end on NVIDIA L4 hosts; a 16 GB memory budget runs everything the app can produce, and 12 GB runs the unquantized model at CFG 1.0 or for shorter songs) or an Apple Silicon Mac with 32 GB+ unified memory (where this app is developed and tested). Windows is best effort: install, launch, Cover and a full-length song verified on Windows 11 (RTX 2070, 8 GB); song generation there runs through the GGUF engine. Cards under 16 GB (and Windows) get the optional GGUF engine: the same model through yue2.cpp with an 8-bit backbone, 8.2 GB peak for a full song, measured indistinguishable from the reference rendering in a blind ABX, a different take for the same seed; the reference PyTorch configuration stays the default wherever it fits.