Followed1d ago
Ideoprompt
github.com/cocktailpeanut

Describe an image, get a 100% schema-valid Ideogram 4 JSON prompt — generated fully locally with an embedded llama.cpp (no Ollama or LM Studio required).

Followed1d ago
Stable Diffusion web UI
github.com/cocktailpeanut

One-click launcher for Stable Diffusion web UI (AUTOMATIC1111/stable-diffusion-webui)

Followed1d ago
watermark-remover-ai
github.com/nihannihu

Free & unlimited AI video and image watermark remover. Uses Florence-2 and LaMA to seamlessly erase watermarks and logos offline

Followed1d ago
HeartMuLa Studio
github.com/cocktailpeanut

A professional, Suno-like music generation studio for HeartLib. https://github.com/fspecii/HeartMuLa-Studio

Followed1d ago
TRELLIS.2
github.com/Deathdadev

One-click installer for Microsoft TRELLIS.2: High-quality 3D asset generation from images with PBR textures.

Followed1d ago
Applio
github.com/pinokiofactory

A simple, high-quality voice conversion tool focused on ease of use and performance.

Followed1d ago
SoulX-Singer
github.com/SUP3RMASS1VE

Official inference code for SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis

Followed1d ago
lavie
github.com/cocktailpeanut

Text-to-Video (T2V) generation framework from Vchitect https://github.com/Vchitect/LaVie

Followed1d ago
Bark Voice Cloning
github.com/cocktailpeanutlabs

Upload a clean 20 seconds WAV file of the vocal persona you want to mimic, type your text-to-speech prompt and hit submit! A local version of https://huggingface.co/spaces/fffiloni/instant-TTS-Bark-cloning

Followed1d ago
Maestro
github.com/yo-steven

An all-in-one, 100% local AI video, image & music studio. Its Director mode turns a single prompt into a full music video or short film — LLM-planned, shot by shot. Built on the WanGP pipeline (Wan 2.1/2.2, LTX-2.3, Qwen, Hunyuan Video, Flux). Requires an NVIDIA GPU (6GB+ VRAM).

Followed1d ago
AllTalk-TTS v2
github.com/6Morpheus6

[NVIDIA ONLY] AllTalk-TTS is a unified UI for E5-TTS, XTTS, Vite TTS, Piper TTS, Parler TTS and RVC, based on CoquiTTS, including a finetune mode.

Followed1d ago
Maestro
github.com/blizaine

Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/Blizaine/Maestro

Followed1d ago
AI Video Clipper & LoRA Captioner
github.com/manat0912

Automatically clip videos and generate captions for LoRA training using advanced vision models like Gemma-3, Qwen3-VL, and Qwen2-VL.

Followed1d ago
XTTS
github.com/6Morpheus6

clone voices into different languages by using just a quick 3-second audio clip. (a local version of https://huggingface.co/spaces/coqui/xtts)

Followed1d ago
Stable Diffusion Forge
github.com/cocktailpeanutlabs

Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion WebUI (based on Gradio) to make development easier, optimize resource management, and speed up inference. https://github.com/lllyasviel/stable-diffusion-webui-forge?tab=readme-ov-file

Followed1d ago
Faceswap
github.com/manat0912

Deepfakes Software For All

Followed1d ago
PhotoMaker
github.com/cocktailpeanutlabs

Customizing Realistic Human Photos via Stacked ID Embedding https://github.com/TencentARC/PhotoMaker

Followed1d ago
autogpt
github.com/pinokiofactory

AutoGPT is a powerful tool that lets you create and run intelligent agents https://github.com/Significant-Gravitas/AutoGPT

Followed1d ago
flux-webui
github.com/6Morpheus6

Minimal Flux Web UI powered by Gradio & Diffusers (Flux Schnell + Flux Merged)

Followed1d ago
MatAnyone2
github.com/ai-anchorite

Remove backgrounds from videos and images with precision AI matting. Runs locally on 12GB VRAM — Windows, Linux, and macOS.