Enhanced background remove and replace app built around BRIA-RMBG-2.0 https://huggingface.co/briaai/RMBG-2.0
Global radar
An all-in-one, 100% local AI video, image & music studio. Its Director mode turns a single prompt into a full music video or short film — LLM-planned, shot by shot. Built on the WanGP pipeline (Wan 2.1/2.2, LTX-2.3, Qwen, Hunyuan Video, Flux). Requires an NVIDIA GPU (6GB+ VRAM).
Differential Diffusion modifies an image according to a text prompt, and according to a map that specifies the amount of change in each region https://differential-diffusion.github.io/
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
AI-powered model downloader, workflow bridge, timeline video editor, and manager for ComfyUI.
A self-hosted Vietnamese Text-to-Speech tool that runs entirely on your machine. No subscriptions, no usage quotas, no data sent to external servers. Your text and audio never leave your computer.
High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean https://github.com/myshell-ai/MeloTTS
DreamID-V: Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer
[NVIDIA GPU required] Unified S^3 fields for animatable 3D asset generation from a single image. 6GB VRAM needed (Low VRAM option)
📹 A more flexible framework that can generate videos at any resolution and creates videos from images.
Fast and High-Quality Zero-Shot voice clone Text-to-Speech with Flow Matching Multilingual
Based on BFS - Best Face Swap, VisoMaster, and SwapAnyHead.
