clone voices into different languages by using just a quick 3-second audio clip. (a local version of https://huggingface.co/spaces/coqui/xtts)
Global radar
Zero-shot voice conversion, real-time voice conversion, and singing voice conversion. https://github.com/Plachtaa/seed-vc
Pure MLX port of LTX-2 (Lightricks LTX-2.3) for Apple Silicon — video + audio generation
Describe any song in plain English, compose it locally with an embedded Qwen GGUF model, and generate it with ACE-Step v1.5.
A full Hermes skin manager for browsing, editing, saving, and activating CLI skins directly from Pinokio.
An open-source, modern-design ChatGPT/LLMs UI/Framework. Supports speech-synthesis, multi-modal, and extensible (function call) plugin system. https://github.com/lobehub/lobe-chat
The ultimate space for work and life — to find, build, and collaborate with agent teammates that grow with you. We are taking agent harness to the next level — enabling multi-agent collaboration, effortless agent team design, and introducing agents as the unit of work interaction.
An open-source, modern-design ChatGPT/LLMs UI/Framework. Supports speech-synthesis, multi-modal, and extensible (function call) plugin system. https://github.com/lobehub/lobe-chat
Image Upscale is an AI-powered application designed to enhance and upscale images using advanced techniques like Stable Diffusion and Tile ControlNet. It provides high-quality image enhancement with options for HDR effects and customizable settings.
holehe allows you to check if the mail is used on different sites like twitter, instagram and will retrieve information on sites with the forgotten password function.
Dub & translate any short video — locally, offline. Voice clone / per-speaker cast / voice packs, on-screen text localized in place, subtitle styling, blur-or-solid mask covers, funny re-dub. One process (FastAPI serves the React SPA), 6 UI languages.
Higgs Audio v3 TTS + AI text director, voice cloning, podcast & audiobook (multi-speaker). 100+ languages, offline, NVIDIA GPU.
Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.
Fully offline speech-to-text transcription with 11 local AI models. Generates styled subtitles (SRT, ASS) and burns them directly onto video. Supports Whisper, Parakeet, Canary, Moonshine, SenseVoice, Vosk, and more. No API keys required.
AI Song Generation with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model. [NVIDIA ONLY]
