Global radar
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Demo of the Collection of Qwen Image Edit LoRAs
Fully offline document-to-speech converter using Tortoise TTS. High-quality autoregressive synthesis with voice cloning. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface https://github.com/comfyanonymous/ComfyUI
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface. https://github.com/comfyanonymous/ComfyUI
Upload a clean 20 seconds WAV file of the vocal persona you want to mimic, type your text-to-speech prompt and hit submit! A local version of https://huggingface.co/spaces/fffiloni/instant-TTS-Bark-cloning
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Stealth headless browser for AI agents - bypass Cloudflare, bot detection, and anti-scraping. Drop-in Puppeteer/Playwright replacement.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Use multiple FLUX.2-Klein LoRAs
[NVIDIA ONLY] High-Quality and Efficient 3D Mesh Generation from a Single Image (Minimum requirements 12GB VRAM / 24GB RAM)
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
PDFCraft is a free, privacy-focused PDF toolkit that runs entirely in your browser. With 90+ professional tools, you can edit, convert, merge, split, and secure your PDF files without ever uploading them to a server.
Temporally consistent video editing. A local version of https://huggingface.co/spaces/weizmannscience/tokenflow
