Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech application
Global radar
Ad-free YouTube frontend with Ollama AI summarization (forked from christian-fei/my-yt)
A single Gradio + React WebUI with extensions for ACE-Step, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Audio, MMS, StyleTTS2, MAGNet, AudioGen, ...
Repository for the CVPR 2026 paper MeshFlow Efficient Artistic Mesh Generation via MeshVAE and Flow-based Diffusion Transformer by Weiyu Li, Antoine Toisoul, Tom Monnier, Roman Shapovalov, Rakesh Ranjan, Ping Tan and Andrea Vedaldi.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models
Local GPU-accelerated music video generator: Gradio UI, analysis, SDXL backgrounds, NVENC output.
[NVIDIA ONLY] Generate an image from multiple images https://github.com/bytedance/UNO
Zero-Shot Text-Based Audio Editing Using DDPM Inversion https://huggingface.co/spaces/hilamanor/audioEditing
AI Song Generation on Mac Apple Silicon, with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model.
YouTube Auto-Subscriber Bot | Optimize your audience engagement workflows. A smart channel management tool designed to automate subscription interactions and handle multi-profile sync operations.
Youtube auto likes and subscribes with multi accounts using python
Youtube bot for various features such as subscribe certain channel or increase views
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Style Aligned Image Generation via Shared Attention https://style-aligned-gen.github.io/
Open source agent built on local models, with its own inference engine. 100% private and offline
Expressive TTS with voice cloning, prompt-driven speech synthesis built on LTX-2.3 by Resemble AI
