Followed7mo ago
pyvideotrans
github.com/jianchang512

Translate the video from one language to another and embed dubbing & subtitles.

Discovered7mo ago
insightface
github.com/deepinsight

State-of-the-art 2D and 3D Face Analysis Project

Discovered7mo ago
inference
github.com/xorbitsai

Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.

Discovered7mo ago
LivePortrait
github.com/cocktailpeanut

Bring portraits to life!

Followed7mo ago
facefusion-comfyui
github.com/facefusion

Industry leading face manipulation platform

Discovered7mo ago
VibeVoice
github.com/microsoft

Open-Source Frontier Voice AI

Discovered7mo ago
shannon
github.com/KeygraphHQ

Fully autonomous AI hacker to find actual exploits in your web apps. Shannon has achieved a 96.15% success rate on the hint-free, source-aware XBOW Benchmark.

Followed7mo ago
Song Generation - a Hugging Face Space by tencent
huggingface.co/spaces

This app lets you create original songs by providing lyrics and optional audio/text prompts. You can choose a music genre or let it auto-select. The system generates a complete song based on your i...

Followed7mo ago
Qwen3-TTS-DubFlow
github.com/zlh123123

🎙️ Qwen3-TTS-DubFlow: An open-source, human-in-the-loop AI dubbing workbench for novels, games, podcasts, and more. Features a "Design-then-Clone" workflow powered by Qwen3-TTS to achieve consistent identity and context-aware emotional performance.

Discovered7mo ago
facefusion/models-3.5.0 · Hugging Face
huggingface.co/facefusion

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Discovered7mo ago
fish-speech-SUP3R
github.com/SUP3RMASS1VE

SOTA Open Source TTS

Followed7mo ago
Chatterbox-TTS-Extended
github.com/petermg

Modified version of Chatterbox that accepts text files as input and no character restrictions. I use it to make audiobooks, especially for my kids.

Followed7mo ago
SwarmUI
github.com/mcmonkeyprojects

SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.

Followed7mo ago
SadTalker
github.com/OpenTalker

[CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation

Discovered7mo ago
GLM-Image
github.com/zai-org

GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image Generation.

Discovered7mo ago
zai-org/GLM-Image · Hugging Face
huggingface.co/zai-org

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Followed7mo ago
Qwen3-TTS Demo - a Hugging Face Space by Qwen
huggingface.co/spaces

This app lets you create natural-sounding speech from text in multiple ways. You can describe a custom voice using text, clone someone's voice from an audio sample, or use predefined speakers. Just...