Local-first, open-source AI watermark removal for Gemini, Doubao, Qwen, Jimeng, Kling, Tencent Yuanbao, Baidu, and more. Smart localization, Gemini reverse-alpha recovery, and LaMa inpainting. Includes a Web UI, CLI, and Python API; supports CUDA, Apple Silicon, and CPU.
Git Source Code Mirror - This is a publish-only repository but pull requests can be turned into patches to the mailing list via GitGitGadget (https://gitgitgadget.github.io/). Please follow Documentation/SubmittingPatches procedure for any of your improvements.
Fully offline document-to-speech converter using Tortoise TTS. High-quality autoregressive synthesis with voice cloning. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.
Fully offline document-to-speech converter using Sesame CSM-1B. Conversational speech model with Llama backbone and voice cloning. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.