Followed7mo ago
Qwen3-ASR
github.com/QwenLM

Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music/song recognition, language detection and timestamp prediction.

Followed7mo ago
ClippedAI
github.com/Shaarav4795

Open-source alternative to OpusClip - AI-powered YouTube Shorts generator.
100% free and unlimited.

Discovered7mo ago
buzz
github.com/chidiwilliams

Buzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.

Discovered7mo ago
natively-cluely-ai-assistant
github.com/evinjohnn

Natively is an open-source, privacy-first AI meeting assistant and the best alternative to Cluely. It supports both local models and cloud AI providers , delivering real-time assistance during meetings, interviews, and conversations while remaining invisible in screen shares and recordings.

Followed7mo ago
open-lovable
github.com/firecrawl

🔥 Clone and recreate any website as a modern React app in seconds

Discovered7mo ago
SDF
github.com/DavesDx

Generador de Imagenes

Followed7mo ago
LivePortrait-AudioDriven
github.com/Hekenye

Audio driven talkinghead generation based on LivePortrait.

Discovered7mo ago
magicresume-pinokio
github.com/LoneWolfVPS

Contribute to LoneWolfVPS/magicresume-pinokio development by creating an account on GitHub.

Discovered7mo ago
Final2x
github.com/EutropicAI

a cross-platform image super-resolution tool

Followed7mo ago
picoclaw
github.com/sipeed

Tiny, Fast, and Deployable anywhere — automate the mundane, unleash your creativity

Followed7mo ago
SkyReels-V1
github.com/SkyworkAI

SkyReels V1: The first and most advanced open-source human-centric video foundation model

Discovered7mo ago
Generative-Media-Skills
github.com/SamurAIGPT

Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.

Discovered7mo ago
Fooocus-API
github.com/mrhan1993

FastAPI powered API for Fooocus

Followed7mo ago
stable-fast-3d
github.com/Stability-AI

SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement

Discovered7mo ago
Text2Video-Zero
github.com/Picsart-AI-Research

[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators

Followed7mo ago
StyleTTS2
github.com/yl4579

StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models

Discovered7mo ago
coqui/XTTS-v1 · Hugging Face
huggingface.co/coqui

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Discovered7mo ago
Far-Better-Voice-Clone-Studio
github.com/martinobettucci

A Gradio-based web UI for voice cloning and voice design, powered by Qwen3-TTS & VibeVoice. Can use Whisper or VibeVoice-ASR for automatic transcription. Improved from the FrankyB origial version