One-command fine-tuning for Qwen3-TTS text-to-speech model with custom voice samples
Global radar
Official implementations for paper: Zero-shot Image Editing with Reference Imitation
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Contribute to visomaster/visomaster-assets development by creating an account on GitHub.
An advanced singing voice synthesis system with high fidelity, expressiveness, controllability and flexibility based on DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism
Contribute to mpalatsi/open-webui-discord-bot development by creating an account on GitHub.
AutoClip : AI-powered video clipping and highlight generation · 一款智能高光提取与剪辑的二创工具
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
text to video, image to video, video extend
[ICLR 2026] UniVideo: Unified Understanding, Generation, and Editing for Videos
[ICCV 2023] DDColor: Towards Photo-Realistic Image Colorization via Dual Decoders
Open-source CLI for unrestricted AI - Access powerful models without censorship
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Code repository for Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
The fastest repo in history to surpass 50K stars ⭐, reaching the milestone in just 2 hours after publication. Better Harness Tools that make real things done. Now writing in Rust using oh-my-codex.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
[ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing
