Followed8h ago
pinokio-zh
github.com/jiwannian

Pinokio 桌面版简体中文汉化:视图覆盖 + 运行时翻译引擎,支持一键安装/卸载与原生菜单汉化(适配 8.2.0)

Followed8h ago
OneTrainer
github.com/supersonic13

The script utilizes various deep learning models to create detailed character cards, including names, summaries, personalities, greeting messages, and character avatars.

Followed8h ago
XTTS
github.com/cocktailpeanut

clone voices into different languages by using just a quick 3-second audio clip. (a local version of https://huggingface.co/spaces/coqui/xtts)

Discovered9h ago
X-AnyLabeling
github.com/umd81568-lab

X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data, combining versatile built-in tools with state-of-the-art AI models and flexible multi-format export.

Followed9h ago
Maestro
github.com/lamardealmaker

An all-in-one, 100% local AI video, image & music studio. Its Director mode turns a single prompt into a full music video or short film — LLM-planned, shot by shot. Built on the WanGP pipeline (Wan 2.1/2.2, LTX-2.3, Qwen, Hunyuan Video, Flux). Requires an NVIDIA GPU (6GB+ VRAM).

Discovered9h ago
AvatarAI
github.com/Washim-8

AvatarAI is a fully local AI application that transforms a single portrait into a realistic talking avatar video. It combines text-to-speech, lip-sync (Wav2Lip), and video processing—running entirely offline with no APIs, ensuring privacy, zero cost, and high-quality output.

Followed10h ago
Hunyuan3D-2-LowVRAM
github.com/pinokiofactory

Text/Image to 3D (Cross Platform: Mac + Windows + Linux): High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models. https://github.com/deepbeepmeep/Hunyuan3D-2GP

Followed10h ago
Comfy LTX Desktop
github.com/ArtDesignAwesome

Pinokio launcher for Comfy LTX Desktop with GGUF and INT8 support.

Followed10h ago
GeoLibre
github.com/cocktailpeanut

A free, open-source GIS for visualizing, exploring, and analyzing geospatial data locally in your browser.

Followed10h ago
Kokoro-TTS
github.com/pinokiofactory

Welcome to Kokoro, a high-quality text-to-speech synthesis program powered by deep learning. This tool converts any text into high-fidelity speech in just a few seconds. Simply input text, select a voice, adjust the speed, and enjoy the generated audio.

Followed11h ago
Qwen3-TTS
github.com/SUP3RMASS1VE

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team

Followed11h ago
OmniVoice
github.com/PierrunoYT

Zero-shot multilingual TTS (600+ languages) with voice cloning and voice design — Gradio UI (app/app.py)

Followed11h ago
Qwen3-TTS (AMD GPU)
github.com/trevortai

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, with AMD GPU support via ROCm. Windows and Linux.

Followed12h ago
Youtube2MP3
github.com/PierrunoYT

🎵 YouTube to MP3 downloader with a simple Gradio UI and bundled FFmpeg. Paste a YouTube link to download MP3.

Followed12h ago
Heartmorrow
github.com/HMDSimDev

A local-first, AI-powered dating and world simulator game that runs in your browser.

Followed12h ago
instantstyle
github.com/cocktailpeanutlabs

Upload the picture of an image, and generate images with that image style. Instant generation with no LoRA required https://huggingface.co/spaces/InstantX/InstantStyle

Followed12h ago
LivePortrait
github.com/pinokiofactory

Bring portraits to life! https://github.com/KwaiVGI/LivePortrait

Followed12h ago
MuseTalk
github.com/manat0912

MuseTalk is a cutting-edge video-to-video (V2V) lip-sync solution engineered to deliver highly accurate and natural mouth movements synchronized to audio input. Precision LipSync: Realistic and seamless synchronization of speech audio to facial movements. Efficiently designed to run on 8–12 GB VRAM,

Followed12h ago
SongGeneration Studio
github.com/6Morpheus6

AI Song Generation with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model. [NVIDIA ONLY]