Demonstration for the Qwen-Image-Edit-2511 model with lazy-loaded LoRA adapters for advanced single- and multi-image editing. Supports 7+ specialized LoRAs including photo-to-anime, multi-angle camera control, pose transfer (Any-Pose), upscaling, style transfer, light migration, and manga tone. Features fast inference (4 steps default).
Global radar
AI Song Generation with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model. [NVIDIA ONLY]
Hermes Agent made portable desktop for Windows — 100 tools, GUI, local models via LM Studio, TTS, Music, ComfyUI, workflows, tool maker. No install. No Docker. No admin rights.
Apple Silicon TTS — curated MLX-first models for cloning, multilingual speech, narration, and expressive voices.
Auris-BadBaDaki is Offline audiobook reader for EPUB, PDF, and TXT with local OmniVoice TTS, character-aware voices, per-book narrator control, and synced text highlighting. Everything runs locally after setup. No API keys. No hosted TTS dependency.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Enter the text you want spoken, choose a language/style, and upload a brief recording of the speaker you wish to imitate. After agreeing to the MIT license, the app creates an audio file that reads...
Fully offline document-to-speech converter using OpenVoice V2. Tone color conversion and voice cloning via MeloTTS. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.
Fast and High-Quality Zero-Shot voice clone Text-to-Speech with Flow Matching
Unified Image Understanding and Generation. Text-to-Image Generation, In-context Generation, Instruction-guided Image Editing, Visual Understanding (Minimum Requirements 12GBV RAM / 48GB RAM, Recommended Requirements 24GB VRAM / 32GB RAM)
GUI for Faster‑Whisper‑XXL transcription tool: download YouTube audio, transcribe local files, manage models, and export multiple formats with themes and auto yt‑dlp updates.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
[NVIDIA ONLY] Remove Objects in videos with inpainting. Recommended requirements 16 - 24 GB VRAM / 48 GB RAM, Minimal requirements 12GB VRAM / 32 GBRAM
YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open
4D face reconstruction & tracking from an image sequence
StyleTTS2 trained on Ukrainian multispeaker dataset
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Visualize, collaborate, and evolve the software architecture with always actual and live diagrams from your code
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
