We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Global radar
Modified version of Chatterbox that accepts text files as input and no character restrictions. I use it to make audiobooks, especially for my kids.
SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.
[CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image Generation.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
This app lets you create natural-sounding speech from text in multiple ways. You can describe a custom voice using text, clone someone's voice from an audio sample, or use predefined speakers. Just...
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
This repository detects your facial expressions and matches you with a famous meme.
Fast high quality video with audio generation with FA3
Discover amazing ML apps made by the community
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
A User Interface for XTTS-2 Text-Based Voice Cloning using only 10 seconds of speech
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
