Followed8d ago
Bark Voice Cloning
github.com/6Morpheus6

Upload a clean 20 seconds WAV file of the vocal persona you want to mimic, type your text-to-speech prompt and hit submit! A local version of https://huggingface.co/spaces/fffiloni/instant-TTS-Bark-cloning

Followed8d ago
llama.cpp
github.com/ggerganov

LLM inference in C/C++

Followed8d ago
OpenMontage
github.com/calesthio

World's first open-source, agentic video production system. 12 pipelines, 52 tools, 500+ agent skills. Turn your AI coding assistant into a full video production studio.

Discovered8d ago
google/gemma-4-E4B-it · Hugging Face
huggingface.co/google

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Followed8d ago
ComfyUI Image to 3D
github.com/V-Sekai-fire

ComfyUI with TRELLIS2, GeometryPack, and UniRig custom nodes for image-to-3D generation

Followed8d ago
OpenReel Video
github.com/cocktailpeanut

Professional browser-based video editor. Open source CapCut alternative. 100% browser-based, no cloud uploads, no watermarks.

Followed8d ago
SwarmUI
github.com/MasterX1582

A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.

Followed8d ago
maestro-seedvc
github.com/blizaine

seed-vc voice conversion adapted for Maestro (GPL-3.0). Cloned automatically at install time by github.com/Blizaine/Maestro - not for standalone use.

Discovered8d ago
lanshu-create-ai-presenter-video
github.com/cclank

Provider-neutral Codex Skill for producing verified AI presenter videos from a script and an authorized presenter image.

Followed8d ago
LTX-Desktop-WanGP
github.com/deepbeepmeep

LTX-Desktop powered By WanGP itself powered among other things by LTX-2 Engine

Followed8d ago
Ten Forward
github.com/wsimon98

Your own AI radio: it writes the lyrics, sings them with YuE2 on your own GPU, and plays them on channels you define. Android app included.

Discovered8d ago
OpenNoMark
github.com/nanmicoder

Local-first, open-source AI watermark removal for Gemini, Doubao, Qwen, Jimeng, Kling, Tencent Yuanbao, Baidu, and more. Smart localization, Gemini reverse-alpha recovery, and LaMa inpainting. Includes a Web UI, CLI, and Python API; supports CUDA, Apple Silicon, and CPU.

Followed8d ago
RVC-WebUI
github.com/bloomsirenix
Followed8d ago
Musubi Tuner
github.com/hoodtronik

Train LoRA / LoHa / LoKr for Wan2.2, FLUX.2, Z-Image, HunyuanVideo, and more — one-click install of kohya-ss/musubi-tuner with its built-in Gradio GUI.

Followed8d ago
e2-f5-tts
github.com/heiredjio-beep

F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching https://huggingface.co/spaces/mrfakename/E2-F5-TTS

Discovered8d ago
YuE2-Studio
github.com/timoncool

Local AI song generator with an editable score — YuE2 on your GPU: full songs with vocals, sheet music, covers, exact replay. Native Windows app, no Python, installer with auto-update.

Followed8d ago
Forge Neo
github.com/6Morpheus6

[NVIDIA ONLY] Stable Diffusion WebUI Forge supporting Flux, Qwen, wan, nunchaku and more in a lightweight WebUI. https://github.com/Haoming02/sd-webui-forge-classic/tree/neo

Discovered8d ago
ABot-Recon
github.com/amap-cvlab

Streaming 3D reconstruction from only video input: Revisiting Local Context for Long-Horizon Streaming 3D Reconstruction

Followed8d ago
Z-Image Fusion
github.com/DenisJunio

Fast, high-quality image generation using comfyui via a Gradio UI