Wanted
1,937 projectsNon-launcher projects without a Pinokio launcher yet.
Official inference repo for FLUX.2 models
Stable Diffusion without the safety filter and invisible watermark
Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.
Contribute to cubiq/ComfyUI_IPAdapter_plus development by creating an account on GitHub.
Forked from Wan2GP, a fast AI Video Generator, for the Apple Silicon.
Translate the video from one language to another and embed dubbing & subtitles.
Industry leading face manipulation platform
Contribute to veo-3/veo-3 development by creating an account on GitHub.
🎙️ Qwen3-TTS-DubFlow: An open-source, human-in-the-loop AI dubbing workbench for novels, games, podcasts, and more. Features a "Design-then-Clone" workflow powered by Qwen3-TTS to achieve consistent identity and context-aware emotional performance.
Modified version of Chatterbox that accepts text files as input and no character restrictions. I use it to make audiobooks, especially for my kids.
[CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
A User Interface for XTTS-2 Text-Based Voice Cloning using only 10 seconds of speech
DeepFaceLab is the leading software for creating deepfakes.
From Images to High-Fidelity 3D Assets with Production-Ready PBR Material
Stable Diffusion web UI
Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
A Gradio-based web UI for voice cloning and voice design, powered by Qwen3-TTS & VibeVoice. Can use Whisper or VibeVoice-ASR for automatic transcription.
Simple and easy to use DDNS. Support Aliyun, Tencent Cloud, Dnspod, Cloudflare, Callback, Huawei Cloud, Baidu Cloud, Porkbun, GoDaddy, Namecheap, NameSilo...
MOVA: Towards Scalable and Synchronized Video–Audio Generation
Send files from one device to many in real-time.
