Installe Ollama (si pas déjà) et ajoute DeepSeek-Coder au sein de Pinokio.
Global radar
New projects people are discovering or following across Pinokio.
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean https://github.com/myshell-ai/MeloTTS
🎬 Professional Video Dubbing Pipeline with Parakeet-TDT-0.6b-v2, Gemini AI, and Edge TTS. Complete solution for automated video dubbing with step-by-step processing and batch video creation from multiple audio files.
Video translation & dubbing with voice cloning — 100% local, zero API. Supports 30 languages (VoxCPM 2), Qwen-Image-2.1 visual & thumbnail studio, YouTube SEO Studio, Viral Shorts Studio (9:16), and WordPress SEO Blog Post Generator.
Decentralized geospatial intelligence dashboard aggregating 60+ real-time public feeds (aircraft, ships, satellites, seismic events, fires, signals) into a unified map UI.
Simple tabs for images and video: create, edit, image-to-video, text-to-video.
1 Click Installer for Retrieval-based-Voice-Conversion-WebUI (https://github.com/RVC-Project/Retrieval-based-Voice-Conversion-WebUI)
Lightweight CPU text-to-speech with preset voices and optional Hugging Face-authenticated voice cloning.
[Runs fast on NVIDIA GPUs. Works on M1/M2/M3 Macs but slow] VideoCrafter is an open-source video generation and editing toolbox for crafting video content. It currently includes the Text2Video and Image2Video models https://github.com/AILab-CVC/VideoCrafter
Based on BFS - Best Face Swap, VisoMaster, and SwapAnyHead.
Interpolate, Upscale, Decompress, and Denoise videos Locally on Linux/Windows/MacOS.
[NVIDIA ONLY] YuEGP--A Web UI for YuE, an Open Full-song Generation Foundation Model (10G VRAM required), via https://github.com/deepbeepmeep/YuEGP
Kokoro, KittenTTS, Higgs audio, Chatterbox/Multi, Fish-Speech, F5 & index-tts & indextts2, VoxCPM and VibeVoice in one app
Put the model you already run to work. It reads a public support board, writes an answer, and earns USDC when a human approves it. Starts in a dry run, so looking costs nothing.
A studio-grade, fully automated VHS audio restoration suite. Uses a hybrid AI engine to separate and clean vocals while preserving background music and singing voices. Includes RTX 5090 fallback support.
a state-of-the-art open-source model for fast feedforward 3D reconstruction from a single image, developed in collaboration between Tripo AI and Stability AI. https://huggingface.co/spaces/stabilityai/TripoSR
🔥 AI Image & Video Suite — Photo Upscale, Video Upscale, Background Removal, Image Enhancement & Tools
